Image created with OpenAI GPT-Image-1. Image prompt: Cheesy late-night infomercial freeze-frame—workout studio with boom-box and cassette tapes featuring chrome headset “AUDIO AI AUTO-TUNE™”; rim sparkle, tape hiss, high-resolution

New Audio Overviews experiment in Search coming to Labs (comin for ya Amazon and Siri!) https://blog.google/products/search/audio-overviews-search-labs/

Search Live with voice in AI Mode on Google Search https://blog.google/products/search/search-live-ai-mode/

xAI tests Voice Mode and Tasks on Grok web with Grok 3.5 (Amazon, Siri) https://www.testingcatalog.com/xai-tests-voice-mode-and-tasks-on-grok-web-with-signs-of-grok-3-5-integration/

Record mode is rolling out today in ChatGPT to Pro, Enterprise, and Edu users. Available on macOS desktop app.”” / X https://x.com/OpenAI/status/1935419375600926971

Eleven v3 now supports Text to Speech in 41 new languages – bringing the total to over 70. This means you can now reach over 90% of the global population with ElevenLabs. https://x.com/elevenlabsio/status/1933557199294640171

Introducing ElevenLabs Conversational AI 2.0 – YouTube https://www.youtube.com/watch?v=TlclS4wLWgY

ElevenLabs Conversational AI now supports MCP — letting AI agents connect to services like Salesforce, HubSpot, Gmail and more instantly. No more manual tool definitions. No more setup friction. https://x.com/elevenlabsio/status/1934704151197597893

ElevenLabs said its Conversational AI offering now supports Anthropic’s open Model Context Protocol (MCP) The development will enable ElevenLabs users to build voice agents that can easily connect to data from third-party apps like Salesforce, HubSpot, and Gmail”” / X https://x.com/rowancheung/status/1934881619086586251

Apple (AAPL) Targets Spring 2026 for Release of Delayed Siri AI Upgrade – Bloomberg https://www.bloomberg.com/news/articles/2025-06-12/apple-targets-spring-2026-for-release-of-delayed-siri-ai-upgrade

I just built a Voice AI Agent that answers inbound calls, handles FAQs, and even books appointments.. the n8n workflow stores relevant information about the call in Airtable.. ✅ Works 24/7 ✅ Handles multiple calls at once ✅ Fully customizable to any business demo loading.. https://x.com/elewachii/status/1925256400260788559

7/ Omi world’s leading open-source AI wearable that captures conversations, gives summaries, action items and does actions for you.  Simply connect Omi to your mobile device and enjoy automatic, high-quality transcriptions of meetings + life @kodjima33 https://x.com/AtomSilverman/status/1932988839578206268

🎙️🤖 Local AI Podcast Generator Transform text into multilingual podcasts with this AI system! Built using LangChain + Ollama, it combines text summarization and speech generation for seamless podcast creation. Check out the tutorial here https://x.com/LangChainAI/status/1933917455560114287

Day 27/30 AI Automation series 🧙‍♂️ Built a powerful n8n workflow that helps you start a faceless YouTube channel in one click. ✅ Generates original Lo-fi music ✅ Handles content creation ✅ Ready to upload to YouTube All automated. Just press a button. https://x.com/Sharmakartikai1/status/1932014062172610985

Day 5/5 of #MiniMaxWeek: MiniMax Audio Dessert for Your Friday Voice Design- A breakthrough in voice generation: 🍰Any prompt, any voice, any emotion 🍩Fully customizable and multilingual Take a bite → https://x.com/MiniMax__AI/status/1936113656372379680

RT @reach_vb: WOW! DeepMind *just* dropped Magenta Real-time – Apache 2.0 licensed 🔥 > 800M params transformer, trained on ~190K hours of…”” / X https://x.com/osanseviero/status/1936182860228034902

So excited to welcome Google’s model #1000 at Hugging Face: Magenta Real Time!🤯 🎷Music generation model ⚡️Real-time 👀Permissive license 🤏800 million parameters Model: https://x.com/osanseviero/status/1936170526931615849

6/ LlamaFS is a self-organizing file manager. It automatically renames and organizes your files based on their content and well-known conventions (e.g., time).  It supports many file types , including images and audio. Super cool project by @AlexReibman! https://x.com/AtomSilverman/status/1932988830904365480

Gemini 2.5 models are sparse mixture-of-experts (MoE) transformers with native multimodal support for text, vision, and audio inputs. https://x.com/_philschmid/status/1935017208343634032

Chatterbox is a free open-source voice cloning model with emotional tone control https://the-decoder.com/chatterbox-is-a-free-open-source-voice-cloning-model-with-emotional-tone-control/

Get ready for the ultimate ASMR experience! 🔪✨ Watch as we slice through the most unexpected things, with our Sound Effects feature. A true healing moment for your senses. 🌿 #KlingAI #soundeffects #ASMR https://x.com/Kling_ai/status/1934894879773213059

Projects Update 📝 We’re adding more capabilities to projects in ChatGPT to help you do more focused work. ✅ Deep research support ✅ Voice mode support ✅ Improved memory to reference past chats in projects ✅ Upload files and access the model selector on mobile”” / X https://x.com/OpenAI/status/1933208575968752092

ChatGPT Record | OpenAI Help Center https://help.openai.com/en/articles/11487532-chatgpt-record

Trending

Discover more from Ethan B. Holland

Subscribe now to keep reading and get access to the full archive.

Continue reading