Image created with Flux Pro v1.1 Ultra. Image prompt: Assembly instruction diagram for a modular toolbox with customizable compartments, community workshop style, orange and gray colors, pegboard background texture, “OPENSOURCE” in accessible font, universal attachment system shown, tool organization guides
deepseek-ai/DeepSeek-R1-0528 · Hugging Face https://huggingface.co/deepseek-ai/DeepSeek-R1-0528
DeepSeek is aiming for the king: o3 and Gemini 2.5 Pro https://x.com/i/web/status/1928067335014793526
On GPQA Diamond, a set of PhD-level multiple-choice science questions, DeepSeek-R1-0528 scores 76% (±2%), outperforming the previous R1’s 72% (±3%). This is generally competitive with other frontier models, but below Gemini 2.5 Pro’s 84% (±3%). https://x.com/EpochAIResearch/status/1928489527204589680
Interrupt 2025 Keynote | Harrison Chase | LangChain – YouTube https://www.youtube.com/watch?v=DrygcOI-kG8&list=PLlBpYFkiSQwqARjZ9z0Lc6iDWmeV3ro2N&index=2
NEW: Mistral AI announces Agents API – code execution – web search – MCP tools – persistent memory – agentic orchestration capabilities Cool to see that Mistral AI has joined the growing number of agent frameworks. More below: https://x.com/omarsar0/status/1927366520985800849
Find out more about our open-source interpretability tools, and how to use them on open-weights models, here: https://x.com/AnthropicAI/status/1928119231213605240
OFFICIAL BENCHMARKS OUT – we have a new open source frontier approaching O3 and Gemini 2.5 Pro 🔥🔥🔥 https://x.com/i/web/status/1928054949247693219
Meta shuffles AI, AGI teams to compete with OpenAI, ByteDance, Google https://www.axios.com/2025/05/27/meta-ai-restructure-2025-agi-llama
Hugging Face just released a Free Course on Model Context Protocol (MCP)! This Free course shows you how to build AI apps that connect to external data and tools using the latest MCP standards. – 100% Free – Earn a certificate of completion Course link 👇🧵 https://x.com/itsafiz/status/1923360017471656323
Spaces at @huggingface is the app store of AI 📱 it’s also the MCP store now 🤠 filter thousands of MCPs you can attach to your LLM 🤗 https://x.com/mervenoyann/status/1927322723891466439
You really can just do things! Use *any* Hugging Face space as a MCP server along with your Local Models! 🔥 Here in we use Qwen 3 30B A3B with @ggml_org llama.cpp and @huggingface tiny agents to create images via FLUX powered by ZeroGPU ⚡ It’s quite a bit crazy to see local https://x.com/reach_vb/status/1927036453713793526
LlamaIndex now supports the new OpenAI Responses API features: · Call any remote MCP server · Use code interpreters by using it as one of the built-in-tools · AND generate images with streaming. https://x.com/llama_index/status/1926996451747356976
Excited to announce we are deepening our collaboration with the open-source ecosystem to help you build incredible AI Agents using Google Gemini! 🚀 We want to make it exceptionally easy for developers to build high-quality applications and AI agents using open-source frameworks https://x.com/_philschmid/status/1924886344330830325
google is all over hugging face today! ❤️ https://x.com/reach_vb/status/1926558763760177396
Hugging Face releases a free Operator-like agentic AI tool | TechCrunch https://techcrunch.com/2025/05/06/hugging-face-releases-a-free-operator-like-agentic-ai-tool/
🚀 Introducing Magentic-UI — an experimental human-centered web agent from @MSFTResearch . It automates your web tasks while keeping you in control 🧠🤝—through co-planning, co-tasking, action guards, and plan learning. 🔓 Fully open-source. We can’t wait for you to try it. 🔗 https://x.com/pyautogen/status/1924501328975237625
Build AI agents with the Mistral Agents API | Mistral AI https://mistral.ai/news/agents-api
Introducing Agents API: your go-to tool for building tailored agents to solve complex real-world problems! https://x.com/MistralAI/status/1927364741162307702
Mistral Agents | Hacker News https://news.ycombinator.com/item?id=41184559
🚀 Ready to deploy your own Open Agent Platform (OAP) instance? In our latest video, we show you how to self-host OAP in production—no managed instance required. OAP is an open-source, no-code platform for building, prototyping, and deploying intelligent agents. With its https://x.com/LangChainAI/status/1927413238733681027
Airweave (@airweave_ai) is an open-source tool that lets agents search any app. It connects to apps, databases, or document stores and turns their contents into searchable knowledge bases for agents. https://x.com/ycombinator/status/1924857652149932282
Not everyone launches on Product Hunt 🤔 So I built open-source Launch Badges for Lovable, Hacker News, Reddit, X & more 🚀 SVG-based React components — plug & play. Highly customizable: layout, icons, count, text, and more. Landing page was designed, built, and launched on https://x.com/sundaywong/status/1920531525193658579
What if Lovable, bolt, v0 were built in themselves? The new Townie is 100% open-source and is itself built on Val Town. (It can edit itself!) It’s not for one-shot demos, but for building complex full-stack projects Code, prompt, branch, pull request Already always deployed https://x.com/stevekrouse/status/1923026760779612609
Every research paper now gets an instant AI-powered summary! One sentence abstracts to make complex research accessible at a glance. https://x.com/fdaudens/status/1925571194276766155
On SWE-bench Verified, a benchmark of real-world software engineering tasks, DeepSeek-R1-0528 scores 33% (±2%), competitive with some other strong models but well short of Claude 4. Performance can vary with scaffold; we use a standard scaffold based on SWE-agent. https://x.com/EpochAIResearch/status/1928489533886058934
The methods we used to trace the thoughts of Claude are now open to the public! Today, we are releasing a library which lets anyone generate graphs which show the internal reasoning steps a model used to arrive at an answer. https://x.com/i/web/status/1928123130725421201
Our interpretability team recently released research that traced the thoughts of a large language model. Now we’re open-sourcing the method. Researchers can generate “attribution graphs” like those in our study, and explore them interactively.”” / X https://x.com/i/web/status/1928119229384970244
🚀 We’re open sourcing Chatterbox – our state-of-the-art Voice Cloning model that includes text-to-speech and voice conversion! In recent testing, 63.75% of listeners preferred Chatterbox over ElevenLabs. Not only is it free and open source (MIT license), it’s demonstrably https://www.resemble.ai/chatterbox/
Inference providers aren’t sleeping on the switch. https://x.com/fdaudens/status/1927834963509961041
DeepSeek’s R1 leaps over xAI, Meta and Anthropic to be tied as the world’s #2 AI Lab and the undisputed open-weights leader DeepSeek R1 0528 has jumped from 60 to 68 in the Artificial Analysis Intelligence Index, our index of 7 leading evaluations that we run independently https://x.com/i/web/status/1928071179115581671
Capgemini and SAP partner with Mistral to deploy AI for sensitive sectors | Reuters https://www.reuters.com/business/capgemini-sap-partner-with-mistral-deploy-ai-sensitive-sectors-2025-05-26/
So so so cool. Llama 1B batch one inference in one single CUDA kernel, deleting synchronization boundaries imposed by breaking the computation into a series of kernels called in sequence. The *optimal* orchestration of compute and memory is only achievable in this way.”” / X https://x.com/karpathy/status/1927506788527591853
Pretty impressive 7B VLM coming out of Xiaomi 🤓 ViT encoder w/ MLP and powered by their 7B Text backbone Compatible w/ Qwen VL arch so works across vLLM, Transformers, SGLang and Llama.cpp Bonus: it can reason and is MIT licensed 🔥 https://x.com/reach_vb/status/1928360066467439012
0528 looks at the big picture… The sycophancy is really too much https://x.com/teortaxesTex/status/1927895061452210456
DeepSeek-R1-0528 just dropped on Hugging Face https://x.com/_akhaliq/status/1927790819001389210
DeepSeek: DeepSeek V3 0324 – Provider Status | OpenRouter https://openrouter.ai/deepseek/deepseek-chat-v3-0324/providers?sort=latency
In case you didn’t catch this – if you make this one simple change to the chat template, you can switch on and off reasoning in @deepseek_ai”” / X https://x.com/i/web/status/1927892447809454455
R1-0528 is out!🎉 https://x.com/i/web/status/1928084342732939642
Plus 13,8% on Aider Polyglot It’s almost like DeepSeek has explicitly stated that they will overhaul RL for coding in the conclusion of R1 paper in January. https://x.com/i/web/status/1927940397872947599
🚀 DeepSeek-R1-0528 is here! 🔹 Improved benchmark performance 🔹 Enhanced front-end capabilities 🔹 Reduced hallucinations 🔹 Supports JSON output & function calling ✅ Try it now: https://x.com/i/web/status/1928061589107900779
DeepSeek dropped DeepSeek R1 v2 this morning! We at Hyperbolic Labs now serve DeepSeek-R1-0528, the first inference provider serving this model on @huggingface. My vibe check: It seems to be the only model that consistently answers “”what is 9.9 – 9.11?”” correctly. 🐋 To whale: https://x.com/Yuchenj_UW/status/1927828675837513793
DeepSeek has maintained its status as amongst AI labs leading in frontier AI intelligence https://x.com/i/web/status/1928071183276159117
DeepSeek has released DeepSeek-R1-0528, an updated version of DeepSeek-R1. How does the new model stack up in benchmarks? We ran our own evaluations on a suite of math, science, and coding benchmarks. Full results in thread! https://x.com/EpochAIResearch/status/1928489524616630483
DeepSeek’s R1 update consolidates the lead of 🇨🇳 Chinese AI Labs in open weights intelligence https://x.com/i/web/status/1928226455424528519
Happy to share 💭 Mixture of Thoughts 💭 A curated, general reasoning dataset that trims down over 1M samples from public datasets to ~350k through an extensive set of ablations 🧑🍳 Models trained on this mix match or exceed the performance of DeepSeek’s distilled models — not https://x.com/_lewtun/status/1927043160275923158
Hey guys! We noticed some of you sharing screenshots and links to our DeepSeek-V3-0526 article on @UnslothAI. The link was hidden and wasn’t meant to be shared publicly or taken as a fact but it seems a few of you were scrapping through the site and uncovered it early! 😅 The”” / X https://x.com/danielhanchen/status/1926966742519091327
Now imagine DeepSeek R2 The 150IQ Tsinghua grads are about to outsmart the 120IQ MIT midwits https://x.com/i/web/status/1928073407943319588
Ollama can now think! 🤔🤔🤔 For thinking models, and especially useful for very thoughtful models like DeepSeek-R1-0528, Ollama can separate the thoughts and the response. Thinking can also be disabled! This is useful for getting a direct response. This works across https://x.com/ollama/status/1928543644090249565
Live in Cline: DeepSeek-R1-0528 It’s showing significant benchmark gains, now matching OpenAI o3 in reasoning tasks (key for Plan mode). We’re excited to observe how these improvements impact real-world coding performance in Cline. https://x.com/i/web/status/1928140455923044636
There will be DeepSeek R1 0528 Qwen 3 8B too matching Qwen 3 235B Thinking in performance too 🤯 Whale COOKED! https://x.com/i/web/status/1928058862923391260
DeepSeek R1 05-28 LiveBench results: – 8th in the Overall ahead of o4-mini, Gemini 2.5 Flash Preview and Qwen3-235B-A22B (biggest competitors) – 1st on Data Analysis !!! – 3rd on Reasoning !! – 4th on Mathematics ! – 11th on Language – 20th on Instruction Following – 23rd on https://x.com/i/web/status/1928173385399308639
Deepseek MLA is afaik the first attn variant that can hit compute-bound regime during inference decode, thanks to high arithmetic intensity (~256). If you’re paying $30k for an H100 and only max out the mem bw and not the FLOPS during inference, you’re leaving $20k on the table.”” / X https://x.com/tri_dao/status/1928170652516725027
We made dynamic 1bit quants for DeepSeek-R1-0528 – 74% smaller 713GB to 185GB. Use the magic incantation -ot “”.ffn_.*_exps.=CPU”” to offload MoE layers to RAM, allowing non MoEs to fit < 24GB VRAM on 16K context! The rest sits in RAM & disk. Quants here: https://x.com/danielhanchen/status/1928278088951157116
Deep Seek R1 Qwen3 8B knows it’s overthinking it 😂 https://x.com/i/web/status/1928119439737729482
The 4-bit DWQ of DSR1 Qwen3 8B is up on HF. Use the command below or use it in @lmstudio: https://x.com/awnihannun/status/1928125690173383098
Meta understood that copying DeepSeek piecemeal is not working, and decided to copy the org structure, creating an internal AGI division. a cruel rhyme from Russian school program comes to mind “”And you, my friends, no matter your positions, Will never be musicians!”” https://x.com/i/web/status/1927944123358581182
Gemma 3 abliterated again ✂️✂️ Abliteration removes refusals from the models. This new and improved version targets refusals with more accuracy, based on previous work with Qwen 3. Here’s how to do it https://x.com/i/web/status/1928030013275918464
@levelsio @huggingface Hahaha! We make money via compute credits + Enterprise Hub + HF Pro subs – business is good! More than a million enterprises, startups and developers of enterprises depend on it 🤗 We already have a lot of key inference providers and we’re scaling up more as I write this! In”” / X https://x.com/i/web/status/1928050126498713706
@levelsio @huggingface One of the first slide when i do talks lol tl;dr is github subscription like for entreprise and user + compute https://x.com/eliebakouch/status/1928065458764194209
Just a few minutes later & the updated R1 is already available on some of our inference partners. All on the model page – beautiful! https://x.com/ClementDelangue/status/1927825872221774281
nvidia/AceReason-Nemotron-14B · Hugging Face https://huggingface.co/nvidia/AceReason-Nemotron-14B
🐯 Liger GRPO meets TRL https://huggingface.co/blog/liger-grpo
Just FYI all the reports from our RL experiments have not been on Qwen, they’ve been on Llama (DeepHermes 8B) – so hopefully that gives some additional assurance on the impact RL can have and that its not random god-mode qwen math improvements from randomness”” / X https://x.com/i/web/status/1928184393035559191
There is a nice documentation for this release. You can see below the things that are supported. Persistent state across conversations, image generation, handoff capabilities, structured outputs, document understanding, citations, and more. https://x.com/omarsar0/status/1927367265789387087
Introduced Blocks sections & pages to Launchmvpfast(.)com – Copy, Paste or Install it with CLI in any @reactjs/@nextjs codebase – Open in @v0 Built with @shadcn UI Free & Open Source ⭐️ https://x.com/AliFarooqDev/status/1922411643826270402
kicking the qwen randomly makes it work better”” like old TVs. I’m not reading any of it at this point”” / X https://x.com/teortaxesTex/status/1927459880341782700
How does an LLM writing out this program (WITHOUT a code interpreter running the output) make things more accurate? Verified on Qwen 3 – a30b (below) Lots of interesting takeaways from the Random Rewards paper. NOT that RL is dead, but honestly far more interesting than that! https://x.com/hrishioa/status/1927974614585725353
random rewards only work for Qwen models but not for other models improvements with random rewards were due to clipping, and disappear once clipping is removed Conjecture by authors: “”Under clipping, random rewards don’t teach task quality – instead, they trigger a”” / X https://x.com/scaling01/status/1927424801938825294
Why are almost all RL experiments done on qwen models? Kind of interesting right…”” / X https://x.com/i/web/status/1927948317931000277
Worth thinking about how this paper reflects on every other RL paper using Qwen. If Qwen works with any random reward, how do we know if any of these papers actually does anything”” / X https://x.com/nrehiew_/status/1927424673702121973
It’s interesting how the major LLM API vendors are converging on the following features: – Code execution: Python in a sandbox – Web search – like Anthropic, Mistral seem to use Brave – Document library aka hosted RAG – Image generation (FLUX for Mistral) – Model Context Protocol”” / X https://x.com/simonw/status/1927378768873550310
RAG is dead, long live agentic retrieval! At LlamaIndex we’ve been saying for a long time that naive RAG is not enough for a modern application. Following from that conviction, we’ve built agentic strategies directly into LlamaCloud that you can adopt with just a few lines of https://x.com/llama_index/status/1928142249935917385
Agent Connectors You can connect tools like web search and code execution to the agents. Other built-in tools include image generation and a document library (accessing documents from Mistral Cloud) for building agentic RAG systems. https://x.com/omarsar0/status/1927369763023396900
Today, we’re unveiling two new open-source AI robots! HopeJR for $3,000 & Reachy Mini for $300. DM me if you want to be added to the waitlist 🤖🤖🤖 Let’s go open-source AI robotics! https://x.com/ClementDelangue/status/1928125034154901937
Hugging Face unveils two new humanoid robots | TechCrunch https://techcrunch.com/2025/05/29/hugging-face-unveils-two-new-humanoid-robots/
⚒️ Integrate LangSmith prompts with your SDLC You can already test, version and collaborate on prompts in LangSmith. Now, with webhook triggers on prompt changes, you can automatically sync prompts to GitHub, external DBs, or kick off CI. 📓 Docs: https://x.com/LangChainAI/status/1927401850405257283




