Image created with gemini-3.1-flash-image-preview with claude-opus-4.7. Image prompt: A centered Byzantine gold reliquary shrine with hinged doors flung open to reveal an exposed clockwork golden bird of visible gears, springs, and etched schematics, while stylized flat-icon mosaic hands emerge from the surrounding tesserae field offering small gilded cogs and filigree feathers, warm candlelit glow on burnished gold with imperial purple and Tyrian crimson accents, symmetrical iconic composition on a gold-ground mosaic with visible grout, the words ‘OPEN SOURCE’ set in heavy ivory-gold Trajan capitals across the lower third, 16:9 full-bleed.
ByteDance just open-sourced one of the most capable multimodal models out there. BAGEL does image generation, editing, style transfer, and visual understanding – all in a single 7B parameter model. Apache 2.0 licensed! One model. No switching between specialized tools. Amazing
https://x.com/kimmonismus/status/2060050186076815792
Let that sink in for a moment. DeepSeek v4 pro 75% discount. Permanent! In: $0.43 Out: $0.87 If you read the DeepSeek v4 tech paper you know that this model is insanely good when it comes to efficiency. Only 27% compute and only 10% cache compares to v3.2. SemiAnalysis wrote
https://x.com/kimmonismus/status/2057868472965640194
DeepSeek has made its temporary 75% price cut on the first-party V4 Pro API permanent, putting V4 Pro on the Pareto frontier of Intelligence Index vs Cost to Run Intelligence Index alongside V4 Flash @deepseek_ai’s first-party V4 Pro API is now $0.435/1M input, $0.87/1M output,
https://x.com/ArtificialAnlys/status/2058021452465799403
DeepSeek is the only lab that is still trying to make intelligence too cheap to meter
https://x.com/scaling01/status/2057835507858518178
DeepSeek just made its 75% price cut on V4-Pro permanent. Xiaomi’s MiMo slashed V2.5 pricing by up to 99%, effective today. Most coverage frames this as a price war. The more interesting part is the engineering that makes these numbers sustainable. DeepSeek’s V4 paper describes
https://x.com/kimmonismus/status/2059578380329394292
DeepSeek made its 75% discount permanent. The AI price war just escalated.
https://thenextweb.com/news/deepseek-v4-pro-75-percent-price-cut-permanent
We are making our discount permanent! 🎉 Enjoy building with DeepSeek-V4-Pro and bring your innovative ideas to life! 🚀
https://x.com/deepseek_ai/status/2057854261699195173
today was a massive day for protein engineering. esmfold2 dropped–next gen of the esm series, fully open on @huggingscience. 1.1 billion predicted structures, 6.8 billion sequences. 800m more entries than the alphafold db, and reportedly edging out alphafold3 on protein
https://x.com/cgeorgiaw/status/2059694583856927201
Reasonix — DeepSeek-native AI coding agent for your terminal
https://esengine.github.io/DeepSeek-Reasonix/
RF-DETR just landed to @huggingface transformers 🥵🔥 sota real-time detection & segmentation models by @roboflow 💜 > play with our real-time demo > fine-tune the models on your use case with our tutorials (takes a toaster’s VRAM) > or just hand them to your agents 😄
https://x.com/mervenoyann/status/2059647988373373253
SHIPPED. Mistral Vibe is now the AI agent for long-horizon productivity and coding, and the home for Work mode, Code mode, the CLI, and a brand new VS Code extension. Let’s go… 🧵
https://x.com/mistralvibe/status/2059984963932499973
I’ve started experimenting with gBrain + Hermes Agent it’s a shared memory layer that sits underneath my Hermes Agent company. every specialist reads from the same brain before they do anything the architecture I’m currently testing: > inputs flow in: my ideas, strategy
https://x.com/shannholmberg/status/2057821004676956586
Excited to dive into this – an open source agent designed with memory/continual learning in mind
https://x.com/hwchase17/status/2059487107144655356
Fast, faster, Qwen. 🚀 Thrilled to see Qwen3.5 reaching a record-breaking 580 tps for agentic workloads on the TokenSpeed engine! This milestone wouldn’t be possible without our incredible partners. Huge thanks to @lightseekorg, @NVIDIAAI, the Mooncake team, and @tri_dao for
https://x.com/Alibaba_Qwen/status/2059674574397313277
Proud to see Qwen3.7‑Max debut at #4 in Code Arena, marking a significant milestone for Qwen in agentic web development. #AlibabaAI #Qwen
https://x.com/AlibabaGroup/status/2059317802935423028
Today’s best coding models from Qwen, DeepSeek, Minimax etc are trained on 1000s of concurrent RL environments to simulate SWE tasks in real git repos. Very excited to open source a tool that unlocks this capability for the whole AI community: point Repo2RLEnv at any GitHub
https://x.com/_lewtun/status/2059995216937886088
Opus 4.8 is now supported in Hermes Agent ^_^
https://x.com/Teknium/status/2060054418821906652
Qwen3.7 Max (20250517) debuts at #4 in Code Arena: Frontend – the top-ranked Chinese lab on the board, surpassing GLM-5.1 and is now on par with Claude Opus 4.6 on agentic web development tasks. Huge congrats to @Alibaba_Qwen on this achievement!
https://x.com/arena/status/2059297720079393107
Most models are only evaluated on a fraction of the benchmarks out there. ArtifactLinker, our new system, predicts which ones would set a new state-of-the-art on benchmarks hosted on @HuggingFace, then runs the evaluation to verify. 🧵
https://x.com/allen_ai/status/2057838486204326078
Today we’re announcing our $113M Series B led by @CapitalGVC. Over the last 6 months, weekly volume on OpenRouter grew from 5T to 25T tokens as AI rapidly shifts from experimentation into production. We’re excited for what comes next.
https://x.com/OpenRouter/status/2059277623629664758
bytedance-research/Lance · Hugging Face
https://huggingface.co/bytedance-research/Lance
DeepSeek Is Building a Harness Team to Rival Claude Code — Why It Matters 🚀 🌟 Insights from Zhihu contributor 刘杨 This may matter more than launching another new model. Because it means DeepSeek has finally realized one thing: a strong model alone is not enough. Claude has
https://x.com/ZhihuFrontier/status/2059180748637376843
Wow. A massive 75% discount from DeepSeek. Either they’ve done some serious inference optimizations, or Huawei chips are just that much cheaper? More open-source AI models, better token economy.
https://x.com/Yuchenj_UW/status/2057855546460676410
Shipping a Trillion Parameters With a Hub Bucket: Delta Weight Sync in TRL
https://huggingface.co/blog/delta-weight-sync
The HF science team just made async RL weight sync ~100x cheaper on bandwidth, and you don’t need a shared cluster anymore. The problem: every RL step, the trainer typically has to sync fresh weights to the inference engine. for a 7B in bf16 that’s ~14GB. for a frontier 1T fp8
https://x.com/ClementDelangue/status/2059989047947260203
China Limits Overseas Travel for AI Talent at DeepSeek, Alibaba, Private Firms – Bloomberg
https://www.bloomberg.com/news/articles/2026-05-26/china-expands-travel-curbs-to-top-ai-talent-at-private-firms
Mistral to explore designing own chips, CEO Arthur Mensch says
https://www.cnbc.com/2026/05/28/mistral-arthur-mensch-design-chips-ai-data-centers.html
We’re taking on the hardest problems in the real world 🏗️🚚 🛫⚛️ Today at The AI Now Summit, held at the Louvre, we announced AI solutions for aerospace, automotive, energy, and physics. Deployed in production at @Airbus , @BMW, @EDFofficiel , and more. More below:
https://x.com/MistralAI/status/2059951137839616110
Official @NVIDIAAI GLM5.1-NVFP4 spotted on @huggingface 🤩
https://x.com/mr_r0b0t/status/2059973066436853769
How far behind are open models? — LessWrong
https://www.lesswrong.com/posts/rJcCrXyEsJKmmDpWG/how-far-behind-are-open-models
Narrative violation: Open model use in Factory has more than 3x’d in the last month relative to closed models. By both total consumption and event count. Will be interesting to see what open vs closed token-share looks like at EOY
https://x.com/matanSF/status/2060005777348112734
Where is the open source frontier model company in the USA?
https://x.com/Jason/status/2060079403212980402
@Jason the two US companies that are most seriously pushing open models above 100B params are NVIDIA and Arcee
https://x.com/willccbb/status/2060122252931412034
.@Microsoft has just open-sourced 2 useful tools: ▪️ RAMPART – a framework for stress-testing AI agents with repeatable attack and safety scenarios directly in CI. ▪️ Clarity – helps teams design the right system before they build it, saving results in a readable markdown
https://x.com/TheTuringPost/status/2057268273952264279
Today we’re announcing ESMFold2, an open scientific engine to power prediction, design, and discovery across protein biology. The new model delivers state of the art performance on protein interactions, especially antibodies, a critical modality for therapeutics. We have
https://x.com/alexrives/status/2059611151860683097
📢 New @heyjasper release ! 📢 MONET 🌸 : An Apache2.0 deduped and recaptioned dataset of 105M samples unlocking reproducible text-to-image research. Nano T2I 🖌️ : A codebase to train your own T2I model 🤗 @huggingface:
https://t.co/RrEDnjSzVg 💻:
https://t.co/gSgVcLXGO9 Very
https://x.com/CChadebec/status/2059983277306351674
Perplexity Is Open-Sourcing Bumblebee
https://www.perplexity.ai/hub/blog/perplexity-is-open-sourcing-bumblebee
Qwen3.7-Max: How Good Is It? 🚀 🌟Evaluation & insights from Zhihu contributor toyama nao TL;DR: A fresh start — but with real acceleration ⚡ Alibaba’s Tongyi team has been working on trillion-parameter Qwen Max models for over half a year. Earlier Max versions felt oddly
https://x.com/ZhihuFrontier/status/2057772126162354660
@huggingface’s @LeRobotHF team announced LeRobot Humanoid, an open-source bipedal robot platform built for roughly $2,500 using mostly 3D-printed and off-the-shelf parts: The release provides complete hardware designs, runtime software, identification tools, simulation
https://x.com/IlirAliu_/status/2057522561790149048
Hugging Face has released LeRobot Humanoid, a full-stack, open research robot. – Costing around $2,500, it is built using 3D-printable parts, off-the-shelf components, and affordable electronics. – The first release includes hardware files, assembly documentation, design
https://x.com/TheHumanoidHub/status/2058205087600984482
This #CVPR2026 paper from our research team is trending #1 on @HuggingFace 🤗 Meet LocateAnything: a vision-language detection model that rethinks bounding box prediction. For AI agents and robots, “seeing” is only useful if a model can pinpoint where something is fast enough to
https://x.com/NVIDIAAI/status/2060058563544801787
12 AI Co-Scientists of 2026 Open-source: ▪️ ERA – builds scientific simulations and software for biology, forecasting, and more ▪️ DISCO – designs proteins and enzymes from scratch ▪️ kUPS – fast molecular simulation engine ▪️ Axplorer by @axiommathai – solved trillion-scale
https://x.com/TheTuringPost/status/2058618540685742450
Thanks to Julien Grok Build v0.1 now has its appropriate 256K context length in Hermes – sorry bout that!
https://x.com/Teknium/status/2057930638632812642





Leave a Reply