Patronus AI launches API to prevent AI hallucinations in real-time

Patronus AI, a San Francisco startup, has launched a self-serve API that detects and prevents AI failures, such as hallucinations and unsafe responses, in real-time. According to CEO Anand Kannappan in an interview with VentureBeat, the platform introduces several innovations, including “judge evaluators” that allow companies to create custom rules in plain English and Lynx, …

Read more

Google launches real-time search for Gemini AI

Google has introduced “Grounding with Google Search” for its Gemini AI platform, allowing developers to enhance their AI applications with current information from Google Search. As reported by VentureBeat’s Michael Nuñez, the service launched just hours before OpenAI’s consumer-focused ChatGPT Search. Google’s offering targets developers and costs $35 per 1,000 queries, while OpenAI’s service is …

Read more

OpenAI expands Realtime API with new voices and reduces costs for developers

OpenAI has updated its Realtime API, currently in beta, with five new expressive voices for speech-to-speech applications and reduced costs for developers by introducing prompt caching. According to OpenAI’s API documentation cited in an article by VentureBeat, the native speech-to-speech feature enables low latency and nuanced output. The company showcased three of the new voices …

Read more

Amphion: open-source toolkit for audio, music and speech generation

Amphion is an open-source toolkit designed to support research and development in audio, music and speech generation. According to the project’s GitHub site, it offers unique visualizations of classic models and architectures to help junior researchers and engineers better understand them. The toolkit supports various individual generation tasks such as text-to-speech (TTS), singing voice synthesis …

Read more

Cerebras Inference achieves breakthrough performance for Llama 3.1-70B

Cerebras has announced a major update to its Cerebras Inference platform, which now runs the Llama 3.1-70B language model at an impressive 2,100 tokens per second – a threefold performance increase compared to the previous release. According to James Wang from the official Cerebras blog, this performance is 16 times faster than the fastest GPU …

Read more

Hugging Face helps companies develop AI

New York-based AI startup Hugging Face is teaming up with Amazon and Google to launch new open-source software aimed at lowering the cost of developing chatbots and other AI systems, Stephen Nellis reports for Reuters. The offering, called “HUGS” (Hugging Face for Generative AI Services), automates the implementation of AI models and will be available …

Read more

Cohere’s Embed 3 now searches for images

AI company Cohere has added multimodal capabilities to its Embed 3 embedding model, allowing images to be included in RAG-based company searches. This is reported by Emilia David for VentureBeat. The new version can create embeddings for both images and text, with both formats stored in a unified database. According to Cohere, this allows companies …

Read more

“Computer Use”: Anthropic’s Claude can now control your PC

Anthropic has unveiled an updated version of its AI model Claude 3.5 Sonnet. According to the company, the model can now control desktop applications and perform PC tasks. It uses a new “Computer Use” feature, which is in public beta. Anthropic emphasizes that the technology is still error-prone and recommends developers initially test it only …

Read more

IBM launches Granite 3.0 models for enterprise

IBM has launched its Granite 3.0 large language models (LLMs), expanding its enterprise AI offerings, Sean Michael Kerner reports for VentureBeat. The new open-source models, available under the Apache 2.0 license, are designed for various enterprise applications, including customer service, IT automation, and cybersecurity. IBM claims the models outperform competitors like Google and Anthropic, having …

Read more

Sana is a small and extremely fast AI image generator

A new text-to-image framework called Sana can efficiently and quickly generate high-resolution images up to 4096 x 4096 pixels. The system uses a deep compression autoencoder, linear attention, and a decoder-based text encoder to optimize performance. According to the developers, Sana-0.6B can compete with state-of-the-art large diffusion models, but is 20 times smaller and over …

Read more

×