Image created with gemini-3.1-flash-image-preview with claude-opus-4.7. Image prompt: Using the provided trail landscape and trail sign reference images, keep the authentic Sonoran Desert setting with rocky singletrack, saguaro, scrub, and bright Arizona sky, but place a brown wooden post sign reading ‘OPEN SOURCE’ in bold ranger-style type with fictional trail entries like ‘→ Fork Ridge 2.1 mi’ and ‘← Pull Request Pass 0.8 mi’ and ‘→ Upstream Spring 3.4 mi’, and next to the post add a small weathered wooden trail register box with its lid propped open showing a logbook full of signatures, with two hikers in the middle distance pausing to sign in. Maintain photorealistic midday desert documentary lighting, weathered sign textures, and the exact sign construction and typography of the reference.
I have been testing DeepSeek-V4-Pro with the Pi coding agent. I am mindblown by how well it works out of the box. A few notes: I spent a few hours building an LLM wiki with an agent powered entirely by DeepSeek-V4-Pro on @FireworksAI_HQ inference. This is the first time I
https://x.com/omarsar0/status/2050009901234282649
Two weeks after release, Hy3 preview is #1 on @OpenRouter’s weekly leaderboard with 3.66T tokens processed, up 298% week-over-week. #1 in overall usage, tool calls, and coding. 15.4% market share across all providers.🏆 Top apps running Hy3 preview: Hermes Agent, Claude Code,
https://x.com/TencentHunyuan/status/2051978552900538403
We’re donating Petri, our open-source alignment tool, to @meridianlabs_ai, so its development can continue independently. Working with Meridian Labs, we’ve also released a major update that improves the adaptability, realism, and depth of Petri’s tests.
https://x.com/AnthropicAI/status/2052494460966019137
FT: DeepSeek is in talks for its first fundraising round at a valuation of around $45 billion. FT: China’s largest state-backed semiconductor investment fund, which has invested in YMTC and CXMT, is in talks to lead DeepSeek’s fundraising round.
https://x.com/jukan05/status/2051904572038455634
Introducing nanowhale 🐳! A tiny DeepSeek model fully pretrained by an agent. Inspired by @karpathy’s nanochat, we gave ml-intern the task of training a tiny MoE with all the architectural advancements of DeepSeek v4. To test it end-to-end, it trained a 100M-parameter MoE
https://x.com/cmpatino_/status/2051343930373837125
There’s a serious gap in multimodal models – they work with images, but still reason in language, which isn’t that precise for visual stuff. @deepseek_ai just dropped an idea to solve this: let the model literally point to exact locations in the image while it thinks. They call
https://x.com/TheTuringPost/status/2050597658423927134
Welcome to DS4, a specialized inference engine for DeepSeek v4 Flash.
https://t.co/UrUJz5I2R1 This project would have been impossible without the existence of llama.cpp and GGML and the work of @ggerganov and all the other contributors. Thanks!
https://x.com/antirez/status/2052405820235678175
will be very funny if DeepSeek-Vision crushes V4-Pro on ARC-AGI-2 solely because it has some semblance of spatial reasoning
https://x.com/teortaxesTex/status/2049947128189923625
👀 DeepSeek briefly released (then deleted) a vision model tech report — what did it reveal? 👀 Zhihu contributor 刘聪NLP breaks it down: Core idea: 👉 A new multimodal reasoning framework that embeds spatial pointers (boxes & points) directly into the chain-of-thought • The
https://x.com/ZhihuFrontier/status/2050238000433659958
DeepSeek V4–almost on the frontier, a fraction of the price
https://simonwillison.net/2026/Apr/24/deepseek-v4/
🎙️Hugging Face’s Clem Delangue: Stop Comparing Engines to Cars
https://www.turingpost.com/p/clem-delangue-hugging-face-ai-builders
I wrote Deep Learning with Python to be the definitive guide to how deep learning works and how to best make use of it. Tens of thousands of people got their career start via this book. 120,000 copies sold, and downloaded by millions more. And now it’s free to read online:
https://x.com/fchollet/status/2051370269445615965
somebody made a huggingface model visualizer!! just plug in the url and explore at any granularity
https://x.com/andrew_n_carr/status/2051102625613897887
vLLM V0 to V1: Correctness Before Corrections in RL
https://huggingface.co/blog/ServiceNow-AI/correctness-before-corrections
Kimi Chatbot Maker Moonshot AI Valued at $20 Billion in Meituan-Led Round
https://finance.yahoo.com/sectors/technology/articles/kimi-chatbot-maker-moonshot-ai-032926690.html
DeepSeek v4 pro (max) scores 48.9% on WeirdML, improving on v4 pro (high) at 46.5%, but still well behind Kimi-k2.6 and GLM-5.1 at 56% and 57%, let alone the closed frontier. These runs, like the previous ones, were through Fireworks AI.
https://x.com/htihle/status/2052042076196335658
🎯 Orchestration War Room: una capa visual para el orquestador de tareas de Hermes. Abstrae la dificultad: tú contratas perfiles expertos, pides una tarea, y ves el progreso en tiempo real. En el hilo enlace al repositorio y puntos clave. Y aquí el vídeo en acción. 🧵
https://x.com/naroh/status/2050998576486973759
Comparing deepseek-tui, opencode and Hermes side by side with the same tasks to V4-Pro, I am fairly confident that Hermes is the best agent among them right now. Highest success rate, fastest, and cheapest. A bit of a shame because the other two appeal to me aesthetically.
https://x.com/teortaxesTex/status/2051549309707928028
Introducing Hermes Agent v 0.13.0 – Multi-Agent orchestration through the Kanban system – Enforced goal completion with /goal – Big optimizations for disk usage – Much more extensibility, custom LLM Providers, custom gateway channels, and much more
https://x.com/Teknium/status/2052495174404874714
it’s very important to build agents that maximize cache hits. That’s the main axis of cost reduction with V4. Here, with OpenCode, I’ve had 91.6% cache hit. It it were the (typical of eg Hermes) 96%, the cost would’ve been ≈30% lower.
https://x.com/teortaxesTex/status/2051525774851682409
Lightpanda is now a browser backend in Hermes by @NousResearch. Open source autonomous agent. Open source browser built for machines. It had to be done. Set Lightpanda as default with automatic Chrome fallback.
https://x.com/lightpanda_io/status/2052369346928758861
opencode-GUI seems a bit buggy, it just burned $0,15 without building anything lol. Testing cli… $0.25, fail. Tui built something… which didn’t work. After a kick, fixed splendidly. ≈84% cache hit, $0.17. Hermes did well first try, auto-debugged, 95% cache hit, $0.12 or so.
https://x.com/teortaxesTex/status/2051551506134896976
Our first dive into Multi-Agent Coordination and Cooperation is here, with Hermes Agent Kanban Orchestrate tasks across multiple agent profiles and dependencies easily and visually. Achieve more. See the docs here:
https://x.com/Teknium/status/2051001156005151226
People love to ask what they can use Hermes Agent for. Hermes scraped the internet for answers and added them to our docs as inspiration. And if you’ve found an interesting use case, you can submit your own!
https://x.com/NousResearch/status/2052140057222369541
study @Teknium: >me asking him the best way to host Hermes on windows >him explaining that WSL2 is the preferred way right now >him sending a previous NousResearch documentation about the set up >him deciding that it is too sparse and reworking the documentation >1 hour
https://x.com/witcheer/status/2052033039379673374
Traditional cron jobs are great for silent tasks on a machine, and Hermes Agent cronjobs are great for extending that to your agent, but why not utilize the gateway and hermes’ cron to access things that don’t need to cost an agent’s time across any messenger service you have
https://x.com/Teknium/status/2052219963591762194
Trinity-Large-Thinking, @arcee_ai’s latest model, is now free on Nous Portal for the next week Sign up for Nous Portal to use it in your Hermes Agent today
https://x.com/NousResearch/status/2051321586980880506
Video content creation sounds simple, but what if you don’t have time to: • Write the script, • Prepare the visuals, • Generate the voiceover, • Create the subtitles, • And finally render the video? This is why we built Noustiny on top of @NousResearch Hermes Agent by
https://x.com/UfukDegen/status/2051088239579345329
You can now use `hermes profile create <name> –no-skills` to create a new agent with no built in skills whatsoever, start with a blank slate, fresh canvas!
https://x.com/Teknium/status/2052351650279645590
All three leading open weights models were released last week. Progress continues for open weights models alongside proprietary ones, with the gap to GPT-5.5, the leading proprietary model, sitting at 6 points on the Artificial Analysis Intelligence Index @Kimi_Moonshot’s Kimi
https://x.com/ArtificialAnlys/status/2050096370200281539
concern trolling Xi about dangers of open source and the need to unite in decelism, while you’re throttling China’s AI progress, is unserious. He’s justified to say «well, you’re a big guy, you’ve got those scary things like Mythos, our smol bean labs won’t add any risk»
https://x.com/teortaxesTex/status/2052045988936683674
For better or worse, regulation for closed-source models served by a few (quite large) companies is easy. It is not as easy to imagine how you regulate open-source models that can be served by a range of decentralized players. Suspect that will become a big policy discussion soon
https://x.com/emollick/status/2049880544477913271
Have open source models closed the gap with proprietary ones? We’ve tracked three years of Arena data across three arenas. The short answer: mostly yes. In Text Arena, the proprietary winner had a +250 Arena lead. By early 2025, it had fallen to low double digits, and at its
https://x.com/arena/status/2052455463573426452
Join us for in-person workshops to develop problems for FrontierMath: Open Problems! We are seeking highly interesting unsolved problems from research mathematics whose solutions can be verified programmatically. These are hard to find. Come take a crack at it! Link below.
https://x.com/EpochAIResearch/status/2051682607918886971
SGLang grew from an open-source inference project into RadixArk this year. Today we officially announce our $100M Seed, led by Accel and co-led by Spark Capital. The numbers are in the announcement (25K+ stars, 400K+ GPUs in deployment). I want to say something else. Frontier
https://x.com/GenAI_is_real/status/2051703162722263180
So happy about this release. Zyphra is still sandbagging a little – doing small models – but here, everything else comes together. Their SoTA architecture (“”DSMoE-MLA++””), high-end RL and test-time scaling. Frontier open lab, talented and brave. Best of the West. 80B upcoming.
https://x.com/teortaxesTex/status/2052106600882528326
This is a good explanation of why the gap between open and closed models is larger than it appears in benchmarks. I would add in that current open models are also more fragile than closed: they handle out-of-distribution problems far less well & have lower emergent capabilities.
https://x.com/emollick/status/2050904152511848871
Today @sparkcapital is co-leading the $100m seed in @radixark with @Accel and partnering alongside many of our other friends across the venture community: A bet on @ying11231 , @BanghuaZ, and the @lmsysorg , @sgl_project community, and on the idea that open infrastructure is
https://x.com/Arpan_Shah_/status/2051651802484150278
We are releasing ZAYA1-8B open-weights under Apache 2.0, free to try on Zyphra Cloud today. Blog:
https://t.co/oXAIrAED8j Technical report:
https://t.co/hwjcBlnYuI Weights:
https://t.co/EqhmQetjg2 Available on Zyphra Cloud:
https://x.com/ZyphraAI/status/2052103646712828119
Zyphra under 1B active parameters, AMD-Trained, big evals, look strong? Zyphra says its new ZAYA1-8B model delivers unusually high reasoning power for its size, using under 1 billion (!) active parameters while competing with much larger open-weight and proprietary systems on
https://x.com/kimmonismus/status/2052346978240205249
💫Very happy to release NeuralBench, to benchmark Neuro AI models and datasets in the open! 🧵Thread, 💻Code, 📝White Paper below:
https://x.com/JeanRemiKing/status/2052034314120896582
“The numbers of people who are going to be able to become AI builders is going to explode!” @ClementDelangue is so bullish on open source that he thinks it will soon bring millions of new builders into AI. I was eager to talk to him and understand how he thinks about business
https://x.com/TheTuringPost/status/2051834531057938799
A fully open source mocap system that works with cheap webcams: The FreeMoCap Project A free-and-open-source, hardware-and-software-agnostic, minimal-cost, research-grade, motion capture system and platform for decentralized scientific research, education, and training:
https://x.com/IlirAliu_/status/2050484464220827774
Multipath Reliable Connection (MRC): a new open networking protocol for large AI training clusters, deployed in production on our largest training clusters.
https://x.com/gdb/status/2052059553542328829
Today we’re releasing ZAYA1-8B, a reasoning MoE trained on @AMD and optimized for intelligence density. With <1B active params, it outperforms open-weight models many times its size on math and reasoning, closing in on DeepSeek-V3.2 and GPT-5-High with test-time compute. 🧵
https://x.com/ZyphraAI/status/2052103618145501459
Hello again, everyone! Our latest Qwopus3.6-35B-A3B-v1 is now live, and it is once again breathtaking! Full HF space benchmark showcase and write-up is in the comments, so you can make conclusions for yourself! GGUF Link below! We had some issues to fix in our training with
https://x.com/KyleHessling1/status/2052064943999267212
.@huggingface’s agentic robotics app store for Reachy Mini is a big step toward more accessible physical AI. 🙌 Excited to see NVIDIA Isaac GR00T N integrated with Hugging Face LeRobot, helping developers post-train, evaluate, and deploy robot foundation models with open
https://x.com/NVIDIARobotics/status/2052446013949149649





Leave a Reply