Image created with Ideogram v3. Image prompt: Late‑90s boy‑band cover “Free‑Love Code”: members tossing floppy disks like confetti; mismatched hacker tees; brick loft background; chrome stencil logo.

“We have optimized the Qwen3 models for coding and agentic capabilities, and also we have strengthened the support of MCP as well. Below we provide examples to show how Qwen3 thinks and interacts with the environment. https://x.com/Alibaba_Qwen/status/1916962100817367192

“The Leaderboard Illusion – Identifies systematic issues that have resulted in a distorted playing field of Chatbot Arena – Identifies 27 private LLM variants tested by Meta in the lead-up to the Llama-4 release https://x.com/arankomatsuzaki/status/1917400711882797144

[AINews] Llama 4’s Controversial Weekend Release • Buttondown https://buttondown.com/ainews/archive/ainews-llama-4s-controversial-weekend-release/

“Research reveals gaming of Chatbot Arena : companies test multiple private variants and cherry-pick results while hoarding 63% of community data. https://x.com/fdaudens/status/1917671335758594474

“BOOOOM! Qwen 3 235B MoE (22B Active) – beats R1, Grok, O1 AND Apache 2.0 licensed! 🔥 https://x.com/reach_vb/status/1916965315910553886

“Qwen3 exhibits scalable and smooth performance improvements that are directly correlated with the computational reasoning budget allocated. This design enables users to configure task-specific budgets with greater ease, achieving a more optimal balance between cost efficiency and https://x.com/Alibaba_Qwen/status/1916962091925442698

“Introducing Qwen3! We release and open-weight Qwen3, our latest large language models, including 2 MoE models and 6 dense models, ranging from 0.6B to 235B. Our flagship model, Qwen3-235B-A22B, achieves competitive results in benchmark evaluations of coding, math, general https://x.com/Alibaba_Qwen/status/1916962087676612998

Tiny Agents: an MCP-powered agent in 50 lines of code https://huggingface.co/blog/tiny-agents

“Is Qwen3-235B the new budget-friendly coding champ in Cline? Early user feedback is rolling in — it’s promising, but not perfect. Here’s what we’re hearing from the Cline community: 🧵” / X https://x.com/cline/status/1917708041857949983

“Spotify just announced ViSMaP on Hugging Face Unsupervised Hour-long Video Summarisation by Meta-Prompting https://x.com/_akhaliq/status/1915703054701044209

“Introducing Blazingly Fast LoRA powered by @FAL and Hugging Face Providers! 🔥 Bring any compatible LoRA and get generations at lightning fast speeds! ⚡ Lots more to come soon! https://x.com/reach_vb/status/1916803002322559444

Blog | Localforge | Localforge https://localforge.dev/blog/running-qwen3-macbook-mlx

“Qwen3-235B-A22B Superior to OpenAIs o3-mini in all the benchmarks 👀 Now it’s all about API pricing and further testing” / X https://x.com/scaling01/status/1916967634786029722

“Qwen3 is now out, but how do you run it locally? Spent the day (not knowing Qwen3 was today until an hour ago) getting @huggingface ChatUI + Qwen2.5 72B running 100% locally. Faced some headaches and confusing docs, so distilled it down for you all: https://x.com/TheZachMueller/status/1916969775525191684

microsoft/OmniParser-v2.0 · Hugging Face https://huggingface.co/microsoft/OmniParser-v2.0

“BOOOOM! you can now use the latest DeepSeek Prover V2 directly on the model page powered by @novita_labs 🔥 Open Source FTW! 💥 https://x.com/reach_vb/status/1917549921470972172

“Although I left DeepSeek quite a while ago, being able to scale up to 671B truly feels like a dream come true for me. I’m deeply grateful to ZZ, Zhihong and other colleagues at DeepSeek for their support, to Liang for the opportunity, and to everyone in the field who has” / X https://x.com/huajian_xin/status/1917603640124363090

“We just released DeepSeek-Prover V2. – Solves nearly 90% of miniF2F problems – Significantly improves the SoTA performance on the PutnamBench – Achieves a non-trivial pass rate on AIME 24 & 25 problems in their formal version Github: https://x.com/zhs05232838/status/1917600755936018715

“Next week, Grok 3.5 early beta release to SuperGrok subscribers only. It is the first AI that can, for example, accurately answer technical questions about rocket engines or electrochemistry. @Grok is reasoning from first principles and coming up with answers that simply don’t” / X https://x.com/elonmusk/status/1917099777327829386

“Elon Musk just confirmed that xAI will launch Grok 3.5 next week He said that the upcoming version will reason from ‘first principles’ and provide answers that don’t exist on the internet However, it’s only coming as an early beta to SuperGrok users! https://x.com/rowancheung/status/1917473832069169258

“The long-awaited Qwen3 is finally here! Our team has put tremendous effort into Qwen3, hoping to bring something fresh to the open LLM community. We’ve made significant progress in pretraining, large-scale reinforcement learning, and integration of reasoning modes. We believe” / X https://x.com/huybery/status/1916962562056524177

“BOOOOM: Today I’m dropping TINY AGENTS the 50 lines of code Agent in Javascript 🔥 I spent the last few weeks working on this, so I hope you will like it. I’ve been diving into MCP (Model Context Protocol) to understand what the hype was all about. It is fairly simple, but https://x.com/julien_c/status/1915790008599904660

“We also evaluated the preliminary performance of Qwen3-235B-A22B on the open-source coding agent Openhands. It achieved 34.4% on Swebench-verified, achieving competitive results with fewer parameters! Thanks to @allhands_ai for providing an easy-to-use agent. Both open models and https://x.com/Alibaba_Qwen/status/1917064282552078480

“🎥🤖 Multi-Modal RAG with Gemma 3 Build a powerful RAG system that processes mixed-content PDFs using Google’s Gemma 3 and LangChain. This implementation combines PDF processing with multi-modal support, powered by Streamlit and Ollama. Check out the tutorial 📚 https://x.com/LangChainAI/status/1916537826050498786

“This AI Agent can automatically apply to LinkedIn jobs that matches your resume using local AI models running with Ollama. 100% free and Opensource. https://x.com/Saboo_Shubham_/status/1914506210340168177

“Alibaba’s Qwen just unveiled Qwen3: a family of eight open models ranging from 600M to 235B params. — Flagship version rivals OpenAI o1 & DeepSeek-R1 — Hybrid “thinking” mode in all models — Boosted coding + agent performance — Supports 119 languages https://x.com/rowancheung/status/1917095301485052142

“Qwen3 and Qwen3 MoEs are already supported in the latest mlx-lm thanks to @Prince_Canuma and @ActuallyIsaak pip install -U mlx-lm Awesome that @Alibaba_Qwen ships a model for every device: -iPhone: 0.6B, 4B -Macbook: 8B, 30B, 3B/30B MoE -M2, M3 Ultra: 22B/235B MoE” / X https://x.com/AwniHannun/status/1916862553852203349

“Qwen3 and Qwen3 MoEs are already supported in the latest mlx-lm thanks to @Prince_Canuma and @ActuallyIsaak pip install -U mlx-lm Awesome that @Alibaba_Qwen ships a model for every device: -iPhone: 0.6B, 4B -Macbook: 8B, 30B, 3B/30B MoE -M2, M3 Ultra: 22B/235B MoE” / X https://x.com/awnihannun/status/1916862553852203349

“The Qwen Chat APP is now available for both iOS and Android users! It’s free to use and designed to assist with creativity, collaboration, and endless possibilities. Just ask, and let Qwen Chat handle the rest. Scan the QR code to quickly access the Qwen Chat APP! https://x.com/Alibaba_Qwen/status/1915761990703697925

“Feel free to download the Qwen Chat Android APP by scanning this QR code! https://x.com/Alibaba_Qwen/status/1915942739855937560

“So, how Qwen3 looks like imo: Main line: finegrained MoE, DeepSeek-like (V2 and V3-Lite scaled), GQA, trained with global-batch load balance, 25T tokens, 256K context, some improved GRPO (DAPO?), unified chat/reasoner, flagship is Sonnet 3.7 tier Dense models largely as before https://x.com/teortaxesTex/status/1916779853111509498

Hugging Face releases a 3D-printed robotic arm starting at $100 | TechCrunch https://techcrunch.com/2025/04/28/hugging-face-releases-a-3d-printed-robotic-arm-starting-at-100/

“NEW: You can now use Dia 1.6B SoTA Text-to-Speech model directly on Hugging Face via @FAL 🔥 You can get up-to 25 generations for less than a dollar 🤗 Run it 5 lines of code too: import requests API_URL = “https://router.huggingface. co/fal-ai/fal-ai/dia-tts” headers = { https://x.com/reach_vb/status/1915418386818834792

“That’s a wrap on the LlamaCon 2025 keynote! In just over two years, Llama has surpassed 1 billion downloads and established itself as the open ecosystem leader in AI. We’re continuing to support the growth and development of the Llama ecosystem with today’s announcements: The https://x.com/AIatMeta/status/1917278290441822674

“We just announced a major leap forward in AI inference: Groq is partnering with Meta to accelerate the official Llama API giving developers the fastest way to run the latest Llama models with no tradeoffs (starting with Llama 4). What developers get with Groq + Meta: 👉 Speeds https://x.com/JonathanRoss321/status/1917621705503080554

“At LlamaCon 2025, Meta announced: —Standalone Meta AI app with a social ‘discover’ feed to take on ChatGPT —Llama API free preview —Lama Guard 4 (12B), LlamaFirewall, and Prompt Guard —Colab with Groq and Cerebras for faster inference https://x.com/rowancheung/status/1917473779069968505

“Qwen3 is a win for open weights & efficiency – hybrid reasoning models that approach DeepSeek R1’s GPQA score with 1/3 the total parameters and a range of smaller models suited for compute limited environments Today, Alibaba announced eight hybrid reasoning models of varying https://x.com/ArtificialAnlys/status/1917246369510879280

LLM360/MegaMath · Datasets at Hugging Face https://huggingface.co/datasets/LLM360/MegaMath

“Really incredible detective work by @singhshiviii et al. at @Cohere_Labs and elsewhere documenting the ways in which @lmarena_ai works with companies to help them game the leaderboard. https://x.com/BlancheMinerva/status/1917445722380681651

DeepSeek available to download again in South Korea after suspension | Reuters https://www.reuters.com/sustainability/boards-policy-regulation/deepseek-available-download-again-south-korea-after-suspension-2025-04-28/

DeepSeek-R2: China’s Powerful New AI Model for 2025 https://deepseek.ai/blog/deepseek-r2-ai-model-launch-2025

“Major updates from LlamaCon! We’re advancing AI security with new open-source Llama protection tools and new AI- powered solutions for the defender community. Developers can now access: — Llama Guard 4, a customizable safeguard that supports protections for text and image” / X https://x.com/AIatMeta/status/1917271400118902860

“The Qwen team really cooked with this release Just incredible work all around: – 235B MoE that is comparable to o1, o3-mini, Gemini 2.5 Pro, etc. – trained on 36T tokens, covering 119 languages! Data extracted from PDFs, synthetic data, etc. – Thinking and non-thinking modes – https://x.com/iScienceLuvr/status/1916966249588002867

“QWEN-3 is finally out! > Matches Gemini 2.5 Pro performance > Outperforms OpenAI o1 > Open-sourced (Apache 2.0) > 119 languages, 32K–128K context https://x.com/LiorOnAI/status/1916998817725223240

“Xiaomi just dropped MiMo on Hugging Face Unlocking the Reasoning Potential of Language Model From Pretraining to Posttraining The final RL-tuned model, MiMo-7B-RL, achieves superior performance on mathematics, code and general reasoning tasks, surpassing the performance of https://x.com/_akhaliq/status/1917410882939715608

THUDM/CogView4-6B · Hugging Face https://huggingface.co/THUDM/CogView4-6B

allenai/OLMo-2-0425-1B · Hugging Face https://huggingface.co/allenai/OLMo-2-0425-1B

“Perception Encoder models and datasets: https://x.com/mervenoyann/status/1915723397272654194

“Fuck it, starting today you can run inference across 30,000+ Flux and SDXL LoRAs on the Hugging Face Hub via Inference Providers (powered by @FAL ⚡) And.. it gets better, you can generate over 40+ images in less than A DOLLAR! Go try it now on your favourite LoRA on HF 🤗 https://x.com/reach_vb/status/1915830938438717777

“Meet Solo Tech, one of the 10 international recipients of the second Llama Impact Grants. Solo Tech uses Llama to offer offline, multilingual AI support for underserved rural communities with limited internet access. This grant will help them to equip 50 rural centers with AI https://x.com/AIatMeta/status/1917727629601616030

“Today at LlamaCon, we announced the 10 international recipients of the second Llama Impact Grants! The Llama Impact Grants are aimed at fostering innovation and creating economic opportunities through open-source AI. This year’s recipients showcase a diverse range of solutions, https://x.com/AIatMeta/status/1917274585189568870

(4) Debug ML Deployments Faster: Test Databricks Locally https://decodingml.substack.com/p/how-to-debug-ml-deployments-20x-faster

“Pretty fucking incredible week so far: > Qwen3 – MoE (235B, 30B) + Dense (32, 14, 8, 4, 0.6B) > Xiaomi – MiMo 7B dense > Kyutai – Helium 2B dense > DeepSeek – Prover V2 671B MoE > Qwen2.5 Omni 3B > Microsoft – Phi4 14B Reasoning, Mini (3.8B) & Plus > JetBrains- Mellum 4B Dense” / X https://x.com/reach_vb/status/1917938596465750476

““We built this place on open source…” Meta Chief Product Officer Chris Cox took to the stage to kick off LlamaCon 2025, reflecting on our long legacy of open source contributions. 🧵 https://x.com/AIatMeta/status/1917353526088589409

Everything we announced at our first-ever LlamaCon https://ai.meta.com/blog/llamacon-llama-news/

“Qwen3-235B Base seems to be benefiting from its 94 layers compared to Llama-4 Mavericks 48 layers or DeepSeeks 61 layers, which are both much larger models https://x.com/scaling01/status/1916986267700506700

Meta previews an API for its Llama AI models | TechCrunch https://techcrunch.com/2025/04/29/meta-previews-an-api-for-its-llama-ai-models/

“meta.llama4-reasoning-17b-instruct-v1:0 https://x.com/btibor91/status/1917232574344384522

“The Leaderboard Illusion – Identifies systematic issues that have resulted in a distorted playing field of Chatbot Arena – Identifies 27 private LLM variants tested by Meta in the lead-up to the Llama-4 release https://x.com/arankomatsuzaki/status/1917400711882797144?s=46

“It turns out that Meta had 27 different models on LM Arena prior to the launch of Llama 4, but they announced it as if they had one model that topped the leaderboard. An extreme example of benchmark hacking (which other labs also do to lesser degrees). https://x.com/emollick/status/1917435868702257538

“> Qwen3 drops > 235B total / 22B active > neck to neck with LLaMA-4-Maverick zucc in absolute shambles https://x.com/ns123abc/status/1916971024509280503

“Dynamic Qwen3 GGUFs are here! – Run them in llama.cpp, lmstudio and ollama nowww! 💥 https://x.com/reach_vb/status/1916982114462900726

“Qwen’s distillation teacher MoE has fewer total parameters (235B) than Meta’s Llama 4 Behemoth has active (288B). As a result its still much smaller distills dunk on Scout viciously. Should have trained that behemoth ass to a fit condition first, huh https://x.com/teortaxesTex/status/1916971319800823932

Microsoft just released Phi 4 Reasoning (14b) : r/LocalLLaMA https://www.reddit.com/r/LocalLLaMA/comments/1kbvwsc/microsoft_just_released_phi_4_reasoning_14b/

Qwen 3: 0.6B to 235B MoE full+base models that beat R1 and o1 | AINews https://news.smol.ai/issues/25-04-28-qwen-3

“Seems like the community is right when feeling that companies overfit strongly to LMSYS! TLDR: closed source companies – get access to a lot of interaction data on models before release – can retract scores/select final model variants 🙄 – are in more battles than OSS models” / X https://x.com/clefourrier/status/1917488919450374383

“We’re publishing new queryable datasets to help researchers explore interpretable features in DeepSeek R1. https://x.com/GoodfireAI/status/1915802798513598490

DeepSeek-Prover-V2/DeepSeek_Prover_V2.pdf at main · deepseek-ai/DeepSeek-Prover-V2 https://github.com/deepseek-ai/DeepSeek-Prover-V2/blob/main/DeepSeek_Prover_V2.pdf

“The chatter is that DeepSeek R2 is going to be released soon…” / X https://x.com/iScienceLuvr/status/1916365312145924532

QwQ-32B claims to match DeepSeek R1-671B | AINews https://news.smol.ai/issues/25-04-16-ainews-qwq-32b-claims-to-match-deepseek-r1-671b

“DeepSeek R1T Chimera – Merge DeepSeek V3 & R1 – 40% fewer tokens, WITHOUT performance loss – MIT licensed 🔥 https://x.com/reach_vb/status/1916490086188736602

“YAYYY! MSFT released Phi 4 Reasoning & Reasoning plus on Hugging Face🔥 Architecture: > Dense decoder-only Transformer > 14B params > 32k context (extendable to 64k) Training: > SFT + RL on 16B tokens (8.3B unique) > 32 H100-80G GPUs for 2.5 days Benchmarks: > AIME 2025:” / X https://x.com/reach_vb/status/1917852036369916081

“Qwen 3 235B now on @togethercompute API! Qwen 3 is a reasoning model that has a non-reasoning instruct mode with allowance for setting a thinking budget. It’s efficient ($0.20/M input & $0.60/M output on our throughput optimized endpoint) and fantastic on a variety of” / X https://x.com/vipulved/status/1917777842466889873

Qwen 3: The new open standard – by Nathan Lambert https://www.interconnects.ai/p/qwen-3-the-new-open-standard

“Qwen-3-MoE vs DeepSeek V2 (original) their designs are superficially similar – but different This will be a very interesting test of a few scaling laws https://x.com/teortaxesTex/status/1916824004901359943

Qwen3: Think Deeper, Act Faster | Qwen https://qwenlm.github.io/blog/qwen3/

“🎉 Congrats @Alibaba_Qwen on releasing the new Qwen 3 family! – Qwen-3 from 0.6B to 235B all open under Apache 2.0 – Qwen3-235B-A22B competitive against the best proprietary models across hard benchmarks Challenge them with your toughest prompts in the Arena! https://x.com/lmarena_ai/status/1917245472521289815

“first impression is that Qwen 30B-3A will be the star of the show” / X https://x.com/teortaxesTex/status/1916918829050998981

“So Qwen 3-235B with thinking seems good, but not blowing away any of my weird frontier tests, some of which DeepSeek r1 did better. It did okay generating a p5js starship (though it had errors to correct), but failed the Lem Test and couldn’t do a twigl shader in many attempts. https://x.com/emollick/status/1917022882888142926

“⬆️ pip install -U vLLM vllm serve Qwen/Qwen3-235B-A22B-FP8 –enable-reasoning –reasoning-parser deepseek_r1 –tensor-parallel-size 4 vLLM introduce Day 0 support for @Alibaba_Qwen Qwen3 and Qwen3 MoE model architecture. Try it out: https://x.com/vllm_project/status/1917008899410215275

“You can now run inference directly on the Qwen 3 235B Hugging Face model page – powered by Together AI! https://x.com/togethercompute/status/1917616701249565120

“OpenRLHF is a pioneering framework to use vLLM for RLHF, driving many design and implementation of vLLM’s features for RLHF, making vLLM a popular choice for many RLHF frameworks. Learn more about the story at https://x.com/vllm_project/status/1915307134256091570

“Since a new version of Grok is coming out, is Xai going to release a system card? Will Grok 2 be made open weights? (Not that I think an open weights Grok 2 would be very competitive right now)” / X https://x.com/emollick/status/1917229176509116769

“New🚨 Grok 3.5 Early Beta is dropping next week to SuperGrok subscribers 🔥 A quick Reminder that, this is what Elon said about Grok 3 just few months ago. https://x.com/InnovationRapid/status/1917214321949495299

“@paulg A much improved Grok-powered algorithm is coming. Should help a lot.” / X https://x.com/elonmusk/status/1915983940399300686

“Elon Musk shared that an improved X algorithm powered by xAI’s Grok AI will be rolling out soon This came in response to Paul Graham complaining about X feed “drowning” in posts from either left or right-wing trolls https://x.com/rowancheung/status/1916726668774842424

“Qwen3 models have a cool feature: toggle thinking mode on and off. It’s a chat template option, so presumably works by including / excluding the `<think>` tokens. (IBMs Granite 3.3 models had a similar feature). Here’s how to use it with mlx-lm: https://x.com/awnihannun/status/1916932256578605246

“Nice UI for managing the thinking/non-thinking modes of Qwen3 https://x.com/fdaudens/status/1916981928009285858

“Qwen3 is out https://x.com/fdaudens/status/1916970577425846446

“🎉 Qwen3 235B is now on HuggingChat! https://x.com/fdaudens/status/1917317723547218352

“Qwen3 is out! SkyPilot is excited to be a close friend with the @Alibaba_Qwen team. Let’s spin up Qwen3 easily on your clusters or clouds with one SkyPilot command! https://x.com/skypilot_org/status/1916987145195295095

“Qwen3-30B-A3B is de facto on par with Qwen3-32B dense and the greatest vindication of finegrained MoEs the world has seen in the open. https://x.com/teortaxesTex/status/1916966009170251899

“Qwen3 models are supporting 119 languages and dialects. This extensive multilingual capability opens up new possibilities for international applications, enabling users worldwide to benefit from the power of these models. https://x.com/Alibaba_Qwen/status/1916962096346202468

“Super excited to introduce SO-101 today from @huggingface, in collaboration with @therobotstudio, Wowrobo, Seeedstudio & Partabot. Building on top of the insanely successful SO-100 (the most popular robot arms ever?), SO-101 are the first robot arms any AI builder should buy. https://x.com/ClementDelangue/status/1916859453917241761

Trending

Discover more from Ethan B. Holland

Subscribe now to keep reading and get access to the full archive.

Continue reading