Image created with gemini-3.1-flash-image-preview with claude-opus-4.7. Image prompt: A centered Byzantine gold reliquary shrine with hinged doors flung open to reveal an exposed clockwork golden bird of visible gears, springs, and etched schematics, while stylized flat-icon mosaic hands emerge from the surrounding tesserae field offering small gilded cogs and filigree feathers, warm candlelit glow on burnished gold with imperial purple and Tyrian crimson accents, symmetrical iconic composition on a gold-ground mosaic with visible grout, the words ‘OPEN SOURCE’ set in heavy ivory-gold Trajan capitals across the lower third, 16:9 full-bleed.

ByteDance just open-sourced one of the most capable multimodal models out there. BAGEL does image generation, editing, style transfer, and visual understanding – all in a single 7B parameter model. Apache 2.0 licensed! One model. No switching between specialized tools. Amazing
https://x.com/kimmonismus/status/2060050186076815792

Let that sink in for a moment. DeepSeek v4 pro 75% discount. Permanent! In: $0.43 Out: $0.87 If you read the DeepSeek v4 tech paper you know that this model is insanely good when it comes to efficiency. Only 27% compute and only 10% cache compares to v3.2. SemiAnalysis wrote
https://x.com/kimmonismus/status/2057868472965640194

DeepSeek has made its temporary 75% price cut on the first-party V4 Pro API permanent, putting V4 Pro on the Pareto frontier of Intelligence Index vs Cost to Run Intelligence Index alongside V4 Flash @deepseek_ai’s first-party V4 Pro API is now $0.435/1M input, $0.87/1M output,
https://x.com/ArtificialAnlys/status/2058021452465799403

DeepSeek is the only lab that is still trying to make intelligence too cheap to meter
https://x.com/scaling01/status/2057835507858518178

DeepSeek just made its 75% price cut on V4-Pro permanent. Xiaomi’s MiMo slashed V2.5 pricing by up to 99%, effective today. Most coverage frames this as a price war. The more interesting part is the engineering that makes these numbers sustainable. DeepSeek’s V4 paper describes
https://x.com/kimmonismus/status/2059578380329394292

DeepSeek made its 75% discount permanent. The AI price war just escalated.
https://thenextweb.com/news/deepseek-v4-pro-75-percent-price-cut-permanent

We are making our discount permanent! 🎉 Enjoy building with DeepSeek-V4-Pro and bring your innovative ideas to life! 🚀
https://x.com/deepseek_ai/status/2057854261699195173

today was a massive day for protein engineering. esmfold2 dropped–next gen of the esm series, fully open on @huggingscience. 1.1 billion predicted structures, 6.8 billion sequences. 800m more entries than the alphafold db, and reportedly edging out alphafold3 on protein
https://x.com/cgeorgiaw/status/2059694583856927201

Reasonix — DeepSeek-native AI coding agent for your terminal
https://esengine.github.io/DeepSeek-Reasonix/

RF-DETR just landed to @huggingface transformers 🥵🔥 sota real-time detection & segmentation models by @roboflow 💜 > play with our real-time demo > fine-tune the models on your use case with our tutorials (takes a toaster’s VRAM) > or just hand them to your agents 😄
https://x.com/mervenoyann/status/2059647988373373253

SHIPPED. Mistral Vibe is now the AI agent for long-horizon productivity and coding, and the home for Work mode, Code mode, the CLI, and a brand new VS Code extension. Let’s go… 🧵
https://x.com/mistralvibe/status/2059984963932499973

I’ve started experimenting with gBrain + Hermes Agent it’s a shared memory layer that sits underneath my Hermes Agent company. every specialist reads from the same brain before they do anything the architecture I’m currently testing: > inputs flow in: my ideas, strategy
https://x.com/shannholmberg/status/2057821004676956586

Excited to dive into this – an open source agent designed with memory/continual learning in mind
https://x.com/hwchase17/status/2059487107144655356

Fast, faster, Qwen. 🚀 Thrilled to see Qwen3.5 reaching a record-breaking 580 tps for agentic workloads on the TokenSpeed engine! This milestone wouldn’t be possible without our incredible partners. Huge thanks to @lightseekorg, @NVIDIAAI, the Mooncake team, and @tri_dao for
https://x.com/Alibaba_Qwen/status/2059674574397313277

Proud to see Qwen3.7‑Max debut at #4 in Code Arena, marking a significant milestone for Qwen in agentic web development. #AlibabaAI #Qwen
https://x.com/AlibabaGroup/status/2059317802935423028

Today’s best coding models from Qwen, DeepSeek, Minimax etc are trained on 1000s of concurrent RL environments to simulate SWE tasks in real git repos. Very excited to open source a tool that unlocks this capability for the whole AI community: point Repo2RLEnv at any GitHub
https://x.com/_lewtun/status/2059995216937886088

Opus 4.8 is now supported in Hermes Agent ^_^
https://x.com/Teknium/status/2060054418821906652

Qwen3.7 Max (20250517) debuts at #4 in Code Arena: Frontend – the top-ranked Chinese lab on the board, surpassing GLM-5.1 and is now on par with Claude Opus 4.6 on agentic web development tasks. Huge congrats to @Alibaba_Qwen on this achievement!
https://x.com/arena/status/2059297720079393107

Most models are only evaluated on a fraction of the benchmarks out there. ArtifactLinker, our new system, predicts which ones would set a new state-of-the-art on benchmarks hosted on @HuggingFace, then runs the evaluation to verify. 🧵
https://x.com/allen_ai/status/2057838486204326078

Today we’re announcing our $113M Series B led by @CapitalGVC. Over the last 6 months, weekly volume on OpenRouter grew from 5T to 25T tokens as AI rapidly shifts from experimentation into production. We’re excited for what comes next.
https://x.com/OpenRouter/status/2059277623629664758

bytedance-research/Lance · Hugging Face
https://huggingface.co/bytedance-research/Lance

DeepSeek Is Building a Harness Team to Rival Claude Code — Why It Matters 🚀 🌟 Insights from Zhihu contributor 刘杨 This may matter more than launching another new model. Because it means DeepSeek has finally realized one thing: a strong model alone is not enough. Claude has
https://x.com/ZhihuFrontier/status/2059180748637376843

Wow. A massive 75% discount from DeepSeek. Either they’ve done some serious inference optimizations, or Huawei chips are just that much cheaper? More open-source AI models, better token economy.
https://x.com/Yuchenj_UW/status/2057855546460676410

Shipping a Trillion Parameters With a Hub Bucket: Delta Weight Sync in TRL
https://huggingface.co/blog/delta-weight-sync

The HF science team just made async RL weight sync ~100x cheaper on bandwidth, and you don’t need a shared cluster anymore. The problem: every RL step, the trainer typically has to sync fresh weights to the inference engine. for a 7B in bf16 that’s ~14GB. for a frontier 1T fp8
https://x.com/ClementDelangue/status/2059989047947260203

China Limits Overseas Travel for AI Talent at DeepSeek, Alibaba, Private Firms – Bloomberg
https://www.bloomberg.com/news/articles/2026-05-26/china-expands-travel-curbs-to-top-ai-talent-at-private-firms

Mistral to explore designing own chips, CEO Arthur Mensch says
https://www.cnbc.com/2026/05/28/mistral-arthur-mensch-design-chips-ai-data-centers.html

We’re taking on the hardest problems in the real world 🏗️🚚 🛫⚛️ Today at The AI Now Summit, held at the Louvre, we announced AI solutions for aerospace, automotive, energy, and physics. Deployed in production at @Airbus , @BMW, @EDFofficiel , and more. More below:
https://x.com/MistralAI/status/2059951137839616110

Official @NVIDIAAI GLM5.1-NVFP4 spotted on @huggingface 🤩
https://x.com/mr_r0b0t/status/2059973066436853769

How far behind are open models? — LessWrong
https://www.lesswrong.com/posts/rJcCrXyEsJKmmDpWG/how-far-behind-are-open-models

Narrative violation: Open model use in Factory has more than 3x’d in the last month relative to closed models. By both total consumption and event count. Will be interesting to see what open vs closed token-share looks like at EOY
https://x.com/matanSF/status/2060005777348112734

Where is the open source frontier model company in the USA?
https://x.com/Jason/status/2060079403212980402

@Jason the two US companies that are most seriously pushing open models above 100B params are NVIDIA and Arcee
https://x.com/willccbb/status/2060122252931412034

.@Microsoft has just open-sourced 2 useful tools: ▪️ RAMPART – a framework for stress-testing AI agents with repeatable attack and safety scenarios directly in CI. ▪️ Clarity – helps teams design the right system before they build it, saving results in a readable markdown
https://x.com/TheTuringPost/status/2057268273952264279

Today we’re announcing ESMFold2, an open scientific engine to power prediction, design, and discovery across protein biology. The new model delivers state of the art performance on protein interactions, especially antibodies, a critical modality for therapeutics. We have
https://x.com/alexrives/status/2059611151860683097

📢 New @heyjasper release ! 📢 MONET 🌸 : An Apache2.0 deduped and recaptioned dataset of 105M samples unlocking reproducible text-to-image research. Nano T2I 🖌️ : A codebase to train your own T2I model 🤗 @huggingface:
https://t.co/RrEDnjSzVg 💻:
https://t.co/gSgVcLXGO9 Very
https://x.com/CChadebec/status/2059983277306351674

Perplexity Is Open-Sourcing Bumblebee
https://www.perplexity.ai/hub/blog/perplexity-is-open-sourcing-bumblebee

Qwen3.7-Max: How Good Is It? 🚀 🌟Evaluation & insights from Zhihu contributor toyama nao TL;DR: A fresh start — but with real acceleration ⚡ Alibaba’s Tongyi team has been working on trillion-parameter Qwen Max models for over half a year. Earlier Max versions felt oddly
https://x.com/ZhihuFrontier/status/2057772126162354660

@huggingface’s @LeRobotHF team announced LeRobot Humanoid, an open-source bipedal robot platform built for roughly $2,500 using mostly 3D-printed and off-the-shelf parts: The release provides complete hardware designs, runtime software, identification tools, simulation
https://x.com/IlirAliu_/status/2057522561790149048

Hugging Face has released LeRobot Humanoid, a full-stack, open research robot. – Costing around $2,500, it is built using 3D-printable parts, off-the-shelf components, and affordable electronics. – The first release includes hardware files, assembly documentation, design
https://x.com/TheHumanoidHub/status/2058205087600984482

This #CVPR2026 paper from our research team is trending #1 on @HuggingFace 🤗 Meet LocateAnything: a vision-language detection model that rethinks bounding box prediction. For AI agents and robots, “seeing” is only useful if a model can pinpoint where something is fast enough to
https://x.com/NVIDIAAI/status/2060058563544801787

12 AI Co-Scientists of 2026 Open-source: ▪️ ERA – builds scientific simulations and software for biology, forecasting, and more ▪️ DISCO – designs proteins and enzymes from scratch ▪️ kUPS – fast molecular simulation engine ▪️ Axplorer by @axiommathai – solved trillion-scale
https://x.com/TheTuringPost/status/2058618540685742450

Thanks to Julien Grok Build v0.1 now has its appropriate 256K context length in Hermes – sorry bout that!
https://x.com/Teknium/status/2057930638632812642

Leave a Reply

Trending

Discover more from Ethan B. Holland

Subscribe now to keep reading and get access to the full archive.

Continue reading