Image created with gemini-3.1-flash-image-preview with claude-sonnet-4-5. Image prompt: Using the provided reference image, preserve the exact weathered wooden crate construction with horizontal reddish-brown slats, iron hardware, three-panel face layout, and hand-painted black lettering style, but replace the address text with a bold stencil illustration of a vintage phonograph horn rendered in the same thick black paint with slight drips and irregularities, placed on a foggy wooden dock in early spring with soft raking morning light emphasizing the wood grain and peeling paint texture.

US man pleads guilty to defrauding music streamers out of millions using AI | AI (artificial intelligence) | The Guardian https://www.theguardian.com/us-news/2026/mar/21/man-pleads-guilty-music-streaming-fraud-ai

Gemini 3.1 Flash Live is our highest-quality audio and voice model yet. Voice capabilities have come a long way and are a big part of how we interact with AI to get things done. 3.1 Flash Live’s improved precision and reasoning make those interactions more natural and intuitive.
https://x.com/sundarpichai/status/2037189971359261081

Gemini 3.1 Flash Live: Google’s latest AI audio model https://blog.google/innovation-and-ai/models-and-research/gemini-models/gemini-3-1-flash-live/

Gemini’s audio and voice capabilities just got an upgrade with Gemini 3.1 Flash Live. Our new high-quality audio and voice model comes with: ⚡️ Faster response times 💬 More helpful, natural dialogue 🧵 2x longer conversation memory in Gemini Live 🌍 Multilingual support for
https://x.com/Google/status/2037190616061284353

Google has released Gemini 3.1 Flash Live Preview, achieving #2 in our Big Bench Audio Speech to Speech model benchmark, and now features configurable thinking levels With thinking level set to high, it scores 95.9% on Big Bench Audio, making it the second-highest scoring speech
https://x.com/ArtificialAnlys/status/2037195442489090485

I had access to the new Google Lyria 3 Pro music AI. Its quite good. I’ve been ruining(?) Rilke by giving the AI the First Elegy & asking it to make it “”more 1990s boy band”” (“”oooo the beginning of terror, girl””) Catchy! It is also nuts that you can ask an AI to do this & it can
https://x.com/emollick/status/2036836310447452606

Introducing Gemini 3.1 Flash Live, our new realtime model to build voice and vision agents!! We have spent more than a year improving the model + infra + experience, the results? A step function improvement in quality, reliability, and latency.
https://x.com/OfficialLoganK/status/2037187750005240307

Introducing Lyria 3 Pro and Lyria 3 Clip, our full song and 30 second music models, available starting today in the Gemini API and our all new music experience in @GoogleAIStudio!!
https://x.com/OfficialLoganK/status/2036848277333622956

Last month we launched Lyria 3. Today, we’re introducing Lyria 3 Pro: our most advanced music model yet, from @GoogleDeepMind. 🎶 Now you can create tracks up to 3 minutes long with more creative control. We’re also bringing Lyria to more Google products starting today.
https://x.com/Google/status/2036836307612119488

Longer tracks are here with Lyria 3 Pro in Gemini! From experimenting with different styles to generating tracks with complex transitions, Lyria 3 Pro makes it easier to bring your full vision to life. Rolling out today to Google AI Plus, Pro, and Ultra users. Learn more 🧵
https://x.com/GeminiApp/status/2036836190431711500

Lyria 3 expands to more Google products, adds more features https://blog.google/innovation-and-ai/technology/ai/lyria-3-pro/

Say hello to Gemini 3.1 Flash Live. 🗣️ Our latest audio model delivers more natural conversations with improved function calling – making it more useful and informed. Here’s what’s new 🧵
https://x.com/GoogleDeepMind/status/2037190678883524716

Today, we’re releasing Lyria 3 music generation models (Pro/Clip) in @GoogleAIStudio and Gemini API! 🎵 – Lyria 3 Pro generates full songs (minutes, controllable via prompt), $0.08/song. – Lyria 3 Clip creates 30-second audio clips, $0.04/song. – Control tempo, time-aligned
https://x.com/_philschmid/status/2036841210770333998

You can now create longer tracks with Lyria 3 Pro. 🎶 Map out intros, verses, choruses, and bridges to build high-fidelity compositions up to 3 minutes long. 🎹
https://x.com/GoogleDeepMind/status/2036836176233918707

Speaking of Voxtral | Mistral AI https://mistral.ai/news/voxtral-tts

ElevenLabs CLI is now agent-first! We made it non-interactive by default, so that your agent and automations can easily interact with it. Rich interactive experience based on Ink UI is now behind –human-friendly flag. > npm install -g @elevenlabs/cli
https://x.com/ElevenLabsDevs/status/2036802792061333989?s=20

A New Framework for Evaluating Voice Agents (EVA) https://huggingface.co/blog/ServiceNow-AI/eva

Suno v5.5: More Expressive. More You. – Suno https://suno.com/blog/v5-5

Cohear 👂 Cohere’s first audio model. Apache 2.0. #1 on the Open ASR leaderboard. Multilingual transcription across 14 languages.
https://x.com/aidangomez/status/2037172942803701838

🔊Introducing Voxtral TTS: our new frontier open-weight model for natural, expressive, and ultra-fast text-to-speech 🎭Realistic, emotionally expressive speech. 🌍Supports 9 languages and accurately captures diverse dialects. ⚡Very low latency for time-to-first-audio. 🔄Easily
https://x.com/MistralAI/status/2037183026539483288

Mistral AI released Voxtral TTS, a 3-billion-parameter text-to-speech model with open weights that the company says outperformed ElevenLabs Flash v2.5 in human preference tests roughly 63% of the time on standard voices and nearly 70% on voice customization. The model runs on
https://x.com/kimmonismus/status/2037149838023024753

Our first speech model, Voxtral TTS, is out. It delivers SOTA performance while significantly reducing cost compared to existing solutions, and it operates with very low latency. It uses a new architecture that combines auto-regressive generation of semantic speech tokens with
https://x.com/GuillaumeLample/status/2037274172607594609

Leave a Reply

Trending

Discover more from Ethan B. Holland

Subscribe now to keep reading and get access to the full archive.

Continue reading