Image created with gemini-2.5-flash-image with claude-sonnet-4-5-20250929. Image prompt: A cinematic photograph of two tall birthday candles on a dark slate surface, flames glowing warmly, with vibrant concentric sound wave rings radiating outward in rich reds, deep blues, and whites against a high-contrast black background. The sound waves appear crisp and energetic, as if the audio itself is celebrating, modern and stylish lighting emphasizing the interplay of flame and frequency.

Researchers from AssemblyAI built a state-of-the-art model that: – transcribes speech across 99 languages. – works even if the audio has many speakers. – outperforms Deepgram and OpenAI models. And much more. (2-step setup below) https://x.com/_avichawla/status/1970376443629904154

Spotify to label AI music, filter spam and more in AI policy change | TechCrunch https://techcrunch.com/2025/09/25/spotify-updates-ai-policy-to-label-tracks-cut-down-on-spam/

We’re excited to share that NVIDIA is investing in ElevenLabs, with support from Jensen Huang. Last week’s U.S. state visit to the UK strengthened AI ties. With our roots growing deeper in both places, this partnership and conversation were the perfect way to cap it off. https://x.com/matistanis/status/1970185470182047788

🎙️ Meet Qwen3-TTS-Flash — the new text-to-speech model that’s redefining voice AI! Demo: https://x.com/Alibaba_Qwen/status/1970163551676592430

🔥 Qwen-Image-Edit-2509 IS LIVE — and it’s a GAME CHANGER. 🔥 We didn’t just upgrade it. We rebuilt it for creators, designers, and AI tinkerers who demand pixel-perfect control. ✅ Multi-Image Editing? YES. Drag in “person + product” or “person + scene” — it blends them like https://x.com/Alibaba_Qwen/status/1970189775467647266

🚀 Introducing Qwen3-LiveTranslate-Flash — Real‑Time Multimodal Interpretation — See It, Hear It, Speak It! 🌐 Wide language coverage — Understands 18 languages & 6 dialects, speaks 10 languages. 👁️ Vision‑Enhanced Comprehension — Reads lips, gestures, on‑screen text and https://x.com/Alibaba_Qwen/status/1970565641594867973

🚀 Introducing Qwen3-Omni — the first natively end-to-end omni-modal AI unifying text, image, audio & video in one model — no modality trade-offs! 🏆 SOTA on 22/36 audio & AV benchmarks 🌍 119L text / 19L speech in / 10L speech out ⚡ 211ms latency | 🎧 30-min audio https://x.com/Alibaba_Qwen/status/1970181599133344172

🚀 We’re thrilled to unveil Qwen3-VL — the most powerful vision-language model in the Qwen series yet! 🔥 The flagship model Qwen3-VL-235B-A22B is now open-sourced and available in both Instruct and Thinking versions: ✅ Instruct outperforms Gemini 2.5 Pro on key vision https://x.com/Alibaba_Qwen/status/1970594923503391182

🛡️ Meet Qwen3Guard — the Qwen3-based safety moderation model series built for global, real-time AI safety! 🌍 Supports 119 languages and dialects ✅ 3 sizes available: 0.6B, 4B, 8B ⚡ Low-latency, Real-time streaming detection with Qwen3Guard-Stream 📝 Robust Full-context safety https://x.com/Alibaba_Qwen/status/1970510193537753397

Alibaba Qwen officially achieves frontier lab status LFG https://x.com/zephyr_z9/status/1970587657421156622

Alibaba released Qwen3-Next-80B-A3B in Base, Instruct, and Thinking variants under an open-weights Apache 2.0 license, targeting faster long-context inference. The 80-billion-parameter mixture-of-experts design swaps most vanilla attention layers for Gated DeltaNet ones and the https://x.com/DeepLearningAI/status/1970254860416131146

Announcing the open-source release of Qwen3-VL! A powerful vision-language model that can operate GUIs, code https://t.co/ww8tsXcd1u charts from mockups, and recognize “”everything”” from daily life to specialized fields. Highlights: 🔹 Precise event location in videos up to 2 https://x.com/Ali_TongyiLab/status/1970665194390220864

NEW: Qwen 235B A22B Vision Language Model is OUTT! Apache 2.0 licensed and upto 1 Million context length 🤯 https://x.com/reach_vb/status/1970589927134937309

Qwen https://qwen.ai/blog?id=1675c295dc29dd31073e5b3f72876e9d684e41c6&from=research.research-list

Qwen https://qwen.ai/blog?id=241398b9cd6353de490b0f82806c7848c5d2777d&from=research.latest-advancements-list

Qwen https://qwen.ai/blog?id=99f0335c4ad9ff6153e517418d48535ab6d8afef&from=research.latest-advancements-list

Qwen https://qwen.ai/blog?id=b2de6ae8555599bf3b87eec55a285cdf496b78e4&from=research.latest-advancements-list

Qwen https://qwen.ai/blog?id=f0bbad0677edf58ba93d80a1e12ce458f7a80548&from=research.research-list

Qwen https://qwen.ai/blog?id=f50261eff44dfc0dcbade2baf1b527692bdca4cd&from=research.research-list

Qwen https://qwen.ai/blog?id=fdfbaf2907a36b7659a470c77fb135e381302028&from=research.research-list

Qwen3 VL might be the best multimodal (vision) model on the planet”” / X https://x.com/scaling01/status/1970591728433283354

Qwen3-Omni is new sota any-to-any model🔥 everything you have to know ⤵️ > a 30B MoE model with 3B active params, comes in three variants: instruct, thinking and captioner 🤩 thinking is for reasoning and captioner is for robust speech generation 🗣️ > it understands everything https://x.com/mervenoyann/status/1970444546216444022

Qwen3-Omni Technical Report A unified multimodal model that matches same-size Qwen text-only and vision-only baselines while pushing audio and audio-visual SOTA. Key technical details below: https://x.com/omarsar0/status/1970502225379381662

We’re excited to announce the upgrade of Qwen3-Coder, and the upgraded API `qwen3-coder-plus` is now available on Alibaba Cloud Model Studio with major improvements: 💻 Enhanced terminal task capabilities and better performance on Terminal Bench (w/ Qwen Code / Claude Code) 🏆 https://x.com/Alibaba_Qwen/status/1970582211993927774

🚨 New Models Update! 🔥 Qwen3 coming in hot into the Arena with three different models: 🔹Qwen3-VL-235b-a22b-thinking for Text & Vision 🔹Qwen3-VL-235b-a22b-instruct for Text & Vision 🔹Qwen3-Max-2025-9-23 for Text Check out the thread to learn more about them and get https://x.com/arena/status/1970920636957831611

Qwen just released Qwen3Guard-Gen-8B on Hugging Face This new safety moderation model offers three-tiered severity classification and multilingual support for AI content. https://x.com/HuggingPapers/status/1970504452466413639

Wow. Qwen Image Edit now has native support for ControlNet (depth maps, edge maps, keypoint maps etc)”” / X https://x.com/bilawalsidhu/status/1970193454505541755

Price analysis reveals trends in the Speech to Text market: Fireworks and Groq are the lowest-cost inference providers for Whisper Large v3, offering competitive access to OpenAI’s model. As word error rate decreases, pricing tends to increase, reflecting the performance-cost https://x.com/ArtificialAnlys/status/1971232403973943517

Voice Assist in Action | Try Our AI Voice Assistant | CallRail https://www.callrail.com/voice-assist-in-action

Neon, the No. 2 social app on the Apple App Store, pays users to record their phone calls and sells data to AI firms | TechCrunch https://techcrunch.com/2025/09/24/neon-the-no-2-social-app-on-the-apple-app-store-pays-users-to-record-their-phone-calls-and-sells-data-to-ai-firms/

AI Music Artist Xania Monet Signs Multimillion-Dollar Record Deal https://www.billboard.com/pro/ai-music-artist-xania-monet-multimillion-dollar-record-deal/#recipient_hashed=ba24e6d3da46405e909aa264349f277cfcfcea88a448e5b534db71220d79eb76&recipient_salt=d4583f86d36a2fd770e7e92d1dcb944618834a16695eb853934f633da567062e

Spotify Strengthens AI Protections for Artists, Songwriters, and Producers — Spotify https://newsroom.spotify.com/2025-09-25/spotify-strengthens-ai-protections/

We’re launching the Artificial Analysis Word Error Rate Index (AA-WER), our new synthesis benchmark for Speech to Text model accuracy comprising of 3 challenging datasets AA-WER comprises three challenging datasets aligned with real-world use cases: AMI-SDM (multi-speaker https://x.com/ArtificialAnlys/status/1971232397921534141

Introducing gemini-2.5-flash-native-audio-preview-09-2025 (sorry😅just latest Gemini Live 🔊) It brings – Natural sounding conversations (+ better at pauses/interruptions) – Much more robust function calling Try it now 👉 https://x.com/osanseviero/status/1970551996227674303

We did Google I/O back to back years with NotebookLM and both nights before the show I literally didn’t sleep. The first year of I/O (2023) we were going to live demo Notebook (Project Tailwind) and announce it to the world. Google Labs was fairly unknown and this was our one https://x.com/raizamrtn/status/1968508322329575452

.@Alibaba_Qwen shipping velocity is unmatched Avg 3.5 releases per month, or almost 1 release per week And the majority are open-weights models. Image credt @Smol_AI https://x.com/awnihannun/status/1970839682503348623

[23 Sept 2025] Alibaba Yunqi: 7 models released in 4 days (Qwen3-Max, Qwen3-Omni, Qwen3-VL) and $52B roadmap congrats @Alibaba_Qwen ! https://x.com/Smol_AI/status/1970842828512088486

📢 New Model(s) Drop: Qwen3 VL 235B A22B Instruct & Thinking are now on Yupp! These latest models are @Alibaba_Qwen’s most powerful vision-language models yet. https://x.com/yupp_ai/status/1970640795259851079

🚀 Qwen3-Max is here—no preview, just power! Qwen Chat: https://x.com/Alibaba_Qwen/status/1970599097297183035

🚀 Your Personal AI Travel Designer Is Here! 🍁 Stop wasting hours planning trips. Qwen Chat Travel Planner crafts complete, day-by-day itineraries tailored JUST for you — powered by Amap, Fliggy APIs and Search. ✅ Recommends perfect hotels & transport routes ✅ Builds https://x.com/Alibaba_Qwen/status/1970554287202935159

🚨 Top 10 Open Model Leaderboard Update New open models have entered the Text Arena, and the top 10 rankings by provider have shifted for September! 🔹Qwen-3-235b-a22b-instruct from @Alibaba_Qwen holds the crown at #1 🏆 🔹Longcat-flash-chat from @Meituan_LongCat makes a strong https://x.com/arena/status/1968705194868535749

ChatGPT – Qwen Aug–sep 2025 Timeline (interactive) https://chatgpt.com/canvas/shared/68d3972d363881918f24524394a87d87

Four new releases from Qwen https://simonwillison.net/2025/Sep/22/qwen/#atom-everything

Just enabled full cudagraphs by default on @vllm_project! This change should offer a huge improvement for low latency workloads on small models and efficient MoEs For Qwen3-30B-A3B-FP8 on H100 at bs=10 1024/128, I was able to see a speedup of 47% 🔥 https://x.com/mgoin_/status/1970601094142439761

qwen3-coder-plus is now available on Anycoder Enhanced terminal task capabilities and better performance on Terminal Bench (w/ Qwen Code / Claude Code) SWE-Bench performance up to 69.6 Safer code generation available as Qwen3-Coder-Plus-2025-09-23 https://x.com/_akhaliq/status/1970595669896503462

Qwen3-Max The Tau bench score is insane”” / X https://x.com/scaling01/status/1970599394337587671

Qwen3-Omni-30B-A3B: – instruct – thinking and -captioner https://x.com/scaling01/status/1970182151019659493

So far, Qwen3-Max seems impressive for a non-reasoning model, doing a good job at a lot of my weird tests that even some reasoners struggle with. https://x.com/emollick/status/1970847381966180685

the new Qwen3-Max is now available in anycoder as default as Qwen3-Max-2025-09-23 https://x.com/_akhaliq/status/1970618469344235677

Try the new Qwen models in the Arena!”” / X https://x.com/Alibaba_Qwen/status/1971097727477088717

Qwen (Qwen) https://huggingface.co/Qwen

Inference Providers @huggingface powered by @novita_labs supports Qwen3-VL, the bleeding-edge vision LM 🔥 the model is quite large (22B active 235B total params) so this makes it super easy to try 💚 https://x.com/mervenoyann/status/1971168938848551021

🚨 New Models Alert: WebDev 💻 GPT-5-Codex and Qwen3-Coder-Plus are both now available on WebDev Arena! In the WebDev Arena, you can test out all the best frontier AI coding models on web development tasks. Vote for your preferred response and see how they stack up on the https://x.com/arena/status/1970962780225507775

Trending

Discover more from Ethan B. Holland

Subscribe now to keep reading and get access to the full archive.

Continue reading