Image created with Gemini. Image prompt: A flat-scanned 1920s Dada Merz collage on aged kraft board showing a phonograph horn constructed from concentric torn sheet music, radio programme clippings, and cut song lyrics unfurling from a black paper record, with wavy strips of newsprint and vermilion ticket stubs rippling outward as sound waves, coffee-stained and buckled with visible glue. The title AUDIO sits across the top in large mismatched cut-out letterpress letters glued at slight angles in ink black and faded vermilion, palette of newsprint cream, kraft brown, slate blue, and oxidized ochre with flat even lighting.

Is the Strait of Hormuz open or closed? Came up over lunch so I voice noted my agent to “boot up god’s eye view and check.” It sent me back this timelapse clip — refreshed the AIS data, checked oil futures, mapped 120 days of vessel transits and rendered it out. TL;DR it’s”
https://x.com/bilawalsidhu/status/2069149876454314137

🚀 The 7-Day Voice AI Builder Challenge is Officially LIVE! Stop babysitting your terminal. 🗣️ The Challenge: Teach your AI coding agent to call you for backup–but only when human intervention is actually required. Real-time feedback? Yes. Live leaderboard? Absolutely. Epic”
https://x.com/DeepLearningAI/status/2069450429465854354

The Interactions API is now generally available. 🎉 The Interactions API is the simplest way to build with Gemini for humans and agents. – One API for Gemini models and agents. – Antigravity Agent with a isolated remote Linux sandbox. – Image gen with Nano Banana; music with”
https://x.com/_philschmid/status/2069108134044467487

Announcing the Artificial Analysis Speech to Speech Index, our new synthesis metric for native Speech to Speech model quality, comprising of Big Bench Audio, Full Duplex Bench, and 𝜏-Voice The index provides a single measure of how well native Speech to Speech models perform,”
https://x.com/ArtificialAnlys/status/2069436163065282737

The Millions of Songs Mashed Into AI-Generated Music – The Atlantic
https://www.theatlantic.com/technology/2026/06/ai-music-generators-suno-google-udio/687485/

Memoket | The Lightest AI Wristband That Remembers Context
https://memoket.ai/

OpenAI prepares bidirectional voice mode for rollout
https://www.testingcatalog.com/openai-prepares-bidirectional-voice-mode-for-rollout-on-chatgpt/

Grok TTS delivers the most human-like speech”
https://x.com/xai/status/2067654108123910495

Over the past six months, we shipped 30+ models, features, and upgraded tools for the API. Our changelog has been busy. Here’s what you may have missed for the API: New models • GPT-5.5 • GPT-5.4 mini • GPT-5.4 nano • GPT-Realtime-2 • GPT-Realtime-Whisper •”
https://x.com/OpenAIDevs/status/2069499656305090671

Leave a Reply

Trending

Discover more from Ethan B. Holland

Subscribe now to keep reading and get access to the full archive.

Continue reading