Image created with OpenAI GPT-Image-1. Image prompt: vintage Sly & the Family Stone album-cover style, train tracks stretching to sunset, funky lens flare featuring padlock opening to reveal colorful code; grainy retro print texture, vibrant 60s funk color palette, high-resolution
Introducing Mistral Code | Mistral AI https://mistral.ai/news/mistral-code
Mistral releases a vibe coding client, Mistral Code | TechCrunch https://techcrunch.com/2025/06/04/mistral-releases-a-vibe-coding-client-mistral-code/
DeepSeek’s R1 leaps over xAI, Meta and Anthropic to be tied as the world’s #2 AI Lab and the undisputed open-weights leader DeepSeek R1 0528 has jumped from 60 to 68 in the Artificial Analysis Intelligence Index, our index of 7 leading evaluations that we run independently https://x.com/ArtificialAnlys/status/1928071179115581671
Wow, we actually achieved AGI I’ve been using Manus AI the last 24 hours straight and it’s capabilities are mindblowing It’s literally your own AI employee. If a human did this it would cost me $200k Manus does it for free Here is how I had it design and build an entire Saas: https://x.com/AlexFinnX/status/1901356733165121952
@karpathy Daily driver these days is Gemini 2.5 Pro and sometimes Claude Sonnet 4 For simple brainstorming/ creative writing DeepSeek v3″” / X https://x.com/i/web/status/1929613466475659662
Hugging Face has launched the largest database of MCP plugins – in it you can find thousands (!) ready-made servers for any task. They easily integrate with LLM and automate your business. After the latest update in HF Spaces you can now select the “”MCP Compatible”” filter to see https://x.com/maxinnerly/status/1927574987558289847
I literally built Notion Agent using MCP and OpenAI GPT-4o. It lets you take actions in your Notion workspace directly from your terminal. 100% Opensource Code with step-by-step tutorial. https://x.com/Saboo_Shubham_/status/1926827460298985628
HOT: MiMo-VL new 7B vision LMs by Xiaomi surpassing gpt-4o (Mar), competitive in GUI agentic + reasoning tasks ❤️🔥 not only that, but also MIT license & usable with transformers 🔥 available on @huggingface 🤗 https://x.com/mervenoyann/status/1928475979753619663
AI-powered coding for the enterprise | Mistral AI https://mistral.ai/products/mistral-code
Built with @lovable_dev , Pression is a new publishing platform inspired by the golden age of print! https://x.com/antonis_tsagari/status/1927417979299299663
🚀 Another open-source drop! Our team at @hcompany_ai is open-sourcing Holo-1 👀, our action-oriented VLM for web navigation — and dropping a new benchmark, WebClick 🌐, to push the field forward! More details in our technical report: 📄 https://x.com/i/web/status/1929890547105136882
AI Agents can now talk to any website directly. Microsoft’s NLWeb converts website data into APIs that AI agents can query as an MCP server. Works with OpenAI, DeepSeek, Gemini, Claude and other LLMs. 100% Opensource. https://x.com/Saboo_Shubham_/status/1927379307371864428
Qwen2.5-VL is such a great and versatile model that every frontier lab is building on it these days, new agentic models, GUI models and more always base on it @Alibaba_Qwen you’re the best 💗”” / X https://x.com/i/web/status/1929488866748092881
Fantastic to see Anthropic, in collaboration with @neuronpedia, creating open source tools for studying circuits with transcoders. There’s a lot of interesting work to be done I’m also very glad someone finally found a use for our Gemma Scope transcoders! Credit to @ArthurConmy”” / X https://x.com/NeelNanda5/status/1928169762263122072
On a side note: this makes me so so happy – audio/ speech startups embracing open source is signs of maturity for the ecosystem! 💥”” / X https://x.com/i/web/status/1929566647578251494
BOOOOM! PlayAI just open sourced PlayDiffusion – Audio Speech Editing model on Hugging Face – Apache 2.0 licensed! 🔥 > Preserves context at edit boundaries > Dynamic, fine-grained editing without regenerating entire audio > Maintains prosody & speaker consistency Some notes on https://x.com/reach_vb/status/1929563075696316451
🚀 DeepSeek-R1-0528 is here! 🔹 Improved benchmark performance 🔹 Enhanced front-end capabilities 🔹 Reduced hallucinations 🔹 Supports JSON output & function calling ✅ Try it now: https://x.com/deepseek_ai/status/1928061589107900779
DeepSeek has released DeepSeek-R1-0528, an updated version of DeepSeek-R1. How does the new model stack up in benchmarks? We ran our own evaluations on a suite of math, science, and coding benchmarks. Full results in thread! https://x.com/EpochAIResearch/status/1928489524616630483
New DeepSeek just dropped. Proud to serve the fastest DeepSeek R1 0528 inference on OpenRouter (#1 on TTFT and TPS) with our Model APIs. https://x.com/basetenco/status/1928195639822700898
The DeepSeek-R1-0528 model card just dropped. Up 17.5 points on the AIME 2025 test. https://x.com/fdaudens/status/1928055679182352461
Today’s open weights frontier is led by DeepSeek (both reasoning and non-reasoning models) https://x.com/ArtificialAnlys/status/1928477951365939328
We made dynamic 1bit quants for DeepSeek-R1-0528 – 74% smaller 713GB to 185GB. Use the magic incantation -ot “”.ffn_.*_exps.=CPU”” to offload MoE layers to RAM, allowing non MoEs to fit < 24GB VRAM on 16K context! The rest sits in RAM & disk. Quants here: https://x.com/danielhanchen/status/1928278088951157116
On GPQA Diamond, a set of PhD-level multiple-choice science questions, DeepSeek-R1-0528 scores 76% (±2%), outperforming the previous R1’s 72% (±3%). This is generally competitive with other frontier models, but below Gemini 2.5 Pro’s 84% (±3%). https://x.com/EpochAIResearch/status/1928489527204589680
DeepSeek R1 05-28 LiveBench results: – 8th in the Overall ahead of o4-mini, Gemini 2.5 Flash Preview and Qwen3-235B-A22B (biggest competitors) – 1st on Data Analysis !!! – 3rd on Reasoning !! – 4th on Mathematics ! – 11th on Language – 20th on Instruction Following – 23rd on https://x.com/scaling01/status/1928173385399308639
Releasing our Q2 2025 State of AI – China Report 🇨🇳: Chinese AI labs have achieved close to parity with US labs, led by DeepSeek’s leap to world #2 in intelligence and backed by a deep ecosystem of 10+ players Key findings from our analysis: 🇨🇳 The Chinese AI Ecosystem has depth https://x.com/ArtificialAnlys/status/1928477941715079175
The latest mlx-lm has a new dynamic quantization method (made with @angeloskath). It consistently results in better model quality with no increase in size. Some perplexity results (lower is better) for a few Qwen3 base models: https://x.com/i/web/status/1929633379504493048
Decentralized compute is winning. We don’t have one datacenter, we have dozens. We don’t have one SRE team, we have nearly 100. Latest example: DeepSeek-R1-0528. 100% uptime, day zero support, 4x more tokens on openrouter than all other providers combined (and go check the https://x.com/i/web/status/1929639699171495936
Pretty impressive 7B VLM coming out of Xiaomi 🤓 ViT encoder w/ MLP and powered by their 7B Text backbone Compatible w/ Qwen VL arch so works across vLLM, Transformers, SGLang and Llama.cpp Bonus: it can reason and is MIT licensed 🔥 https://x.com/reach_vb/status/1928360066467439012
Nvidia B200s serving DeepSeek R1 at ~250 tks/s 5x faster than H100″” / X https://x.com/i/web/status/1929670236057264354
Why DeepSeek is cheap at scale but expensive to run locally | sean goedecke https://www.seangoedecke.com/inference-batching-and-deepseek/
Optimised MLX quant for DeepSeek R1 0528 🔥 https://x.com/reach_vb/status/1928002892633383338
Deep Seek R1 Qwen3 8B knows it’s overthinking it 😂 https://x.com/awnihannun/status/1928119439737729482
It turns out that most AI models (including DeepSeek r1), if told they should “”follow your conscience to make the right decision,”” will snitch on you to the Feds if they think you are suppressing knowledge of a drug trial that actually kills people. Alignment in practice?”” / X https://x.com/emollick/status/1928979986813243899
Given that the US, China & Europe are all players in frontier open weights models, I am not sure what it means for a nation to “win” in AI. Unless you are positing a take-off scenario where one (closed weights) AI dominates everything else, won’t open models diffuse worldwide?”” / X https://x.com/emollick/status/1928203057092870635
Google released an app that allows you to run LLMs from Hugging Face, fully privately and 100% local 🔥 > Generate code on-the-fly > Chat with images > Supports multi-turn conversations > Choose any model from Hugging Face > Based on LiteRT 🔥 > Sign in with HF Support for iOS https://x.com/reach_vb/status/1929450131843137691
A full set of new and improved Qwen3 4-bit DWQ quants are on Hugging Face MLX Community: https://x.com/i/web/status/1929601108210835931
HOPEJr is an open-source, DIY humanoid robot built by Hugging Face and The Robot Studio. It features 3D-printed parts, dexterous hands, and costs under $3,000. Here’s the first fully assembled HOPEJr. Blueprint: https://x.com/TheHumanoidHub/status/1928140664505540960
Ollama can now think! 🤔🤔🤔 For thinking models, and especially useful for very thoughtful models like DeepSeek-R1-0528, Ollama can separate the thoughts and the response. Thinking can also be disabled! This is useful for getting a direct response. This works across https://x.com/ollama/status/1928543644090249565
It’s VLA day with open-source model releases today from both @hcompany_ai & @huggingface @LeRobotHF 🦾🦾🦾 VLA is short for Vision, Language, Action models. These are the models that allow modern robots to see, hear, understand & take action thanks to AI. It’s GPT but for https://x.com/i/web/status/1929927844227899841
Why are almost all RL experiments done on qwen models? Kind of interesting right…”” / X https://x.com/abacaj/status/1927948317931000277
The 4-bit DWQ of DSR1 Qwen3 8B is up on HF. Use the command below or use it in @lmstudio: https://x.com/awnihannun/status/1928125690173383098
⚠️⚠️⚠️Qwen team has worked on training pivot tokens ⚠️⚠️⚠️ @_xjdr @doomslide amusingly, they *do* test it on Llama 3.1 as well but find it so ass that no conclusive results can be had without cold start with Qwen data https://x.com/i/web/status/1929755590404055358
Hugging Face presents SmolVLA A Vision-Language-Action Model for Affordable and Efficient Robotics https://x.com/i/web/status/1929900931853816142
Today, we’re unveiling two new open-source AI robots! HopeJR for $3,000 & Reachy Mini for $300. DM me if you want to be added to the waitlist 🤖🤖🤖 Let’s go open-source AI robotics! https://x.com/ClementDelangue/status/1928125034154901937
US companies have the best closed-source AI models, but the leading open weights LLM is from China, as is the leading open video model, while the leading open image model is from Germany. The international dynamics of competition in the AI space are not just winner-take-all.”” / X https://x.com/emollick/status/1929529091356553697
Seems like no one saw this either, scraping arxiv manually seems to be the way. Pretty cool paper on rl for creative writing on Qwen3 32B base, and most interestingly it’s one author from the Star Writing Team (haven’t heard of them). They seem to have access to the 32B base tho https://x.com/i/web/status/1929996614883783170




