Google updates Flow and Flow Music with new AI model and mobile apps

Google has announced a series of updates to its AI-powered creative platforms, Google Flow and Google Flow Music. The changes were revealed at Google I/O, the company’s annual developer conference, and include a new AI model, an agentic assistant, custom tool-building and mobile applications. Google Flow is an AI creative studio that lets users generate, …

Read more

Alexa+ now runs its own AI podcast shows on demand

Amazon has added an AI-generated podcast feature to its Alexa+ voice assistant, allowing users to create audio content on any topic using two synthetic co-hosts. Todd Spangler reports for Variety that the feature, called Alexa Podcasts, produces episodes through artificial intelligence without any human editorial involvement. Users request a topic by voice. Alexa+ then outlines …

Read more

OpenAI adds advanced reasoning, translation, and transcription to its voice API

OpenAI has released three new voice models for developers through its Realtime API. Each model focuses on a different capability: reasoning, translation, and transcription. The first model, GPT-Realtime-2, brings GPT-5-class reasoning to live voice conversations. According to OpenAI, it can handle complex requests, manage interruptions, and call external tools while keeping a conversation going naturally. …

Read more

AI music floods streaming platforms as services scramble to respond

Artificial intelligence is generating music at a scale that is reshaping the streaming industry. French streaming service Deezer reports that 75,000 AI-generated tracks are uploaded to its platform every day, accounting for around 44 percent of all daily uploads. Spotify removed more than 75 million spam tracks in a single year. The surge is driven …

Read more

Gemini for content professionals: Google’s AI ecosystem explained

Gemini is the label for Google’s AI offerings. The company has developed a dazzling array of tools and services under it. In combination, they are arguably the best AI has to offer today. Even considered on their individual merits, they are often best in class. But Google’s Gemini universe is so vast, that it is …

Read more

Google’s new voice AI lets you direct speech like a film director

Google has released Gemini 3.1 Flash TTS, a new text-to-speech model that the company describes as its most natural and expressive to date. The model is available in preview through the Gemini API, Google AI Studio, Vertex AI for enterprise users, and Google Vids for Workspace users. The model supports more than 70 languages and …

Read more

Google’s Lyria 3 Pro brings structure to AI-generated music

Google has expanded its AI music generation capabilities with the launch of Lyria 3 Pro, a model that creates tracks up to three minutes long. Myriam Hamed Torres writes for Google DeepMind that the model understands musical structure, allowing users to prompt for specific elements such as intros, verses, choruses, and bridges. Lyria 3 Pro …

Read more

Google launches Gemini 3.1 Flash Live voice model

Google has released Gemini 3.1 Flash Live, its latest real-time voice and audio model. Valeria Wu and Yifan Ding write in the Google Blog that the model offers faster responses and improved natural conversation compared to its predecessor. The model is available in several Google products. Developers can access it via the Gemini Live API …

Read more

Mistral releases open-weight text-to-speech model Voxtral TTS

French AI company Mistral has released Voxtral TTS, an open-weight text-to-speech model aimed at enterprise use cases such as customer support, sales, and real-time translation. Unlike competitors such as ElevenLabs, Deepgram, and OpenAI, Mistral is releasing the full model weights, allowing companies to run the system on their own infrastructure without sending data to a …

Read more

×