Image created with Gemini. Image prompt: A horizontal 1920s Dada Merz collage on aged board, featuring a torn ledger page with hand-numbered ethical clauses, a faded vermilion cut-paper spiral of concentric torn circles off-center, layered with rulebook scraps, red inspection stamps, and a punched bus ticket, with the title ‘Anthropic’ spelled in mismatched cut-out letterpress letters glued at slight angles across the top, flat even lighting, visible paper fiber and glue buckling, muted cream, kraft, vermilion, ink black, and slate blue palette.

The White House and Anthropic may have found the first serious path to restore Mythos and Fable access without pretending jailbreaks can be eliminated. AI regulation may be shifting from vague fear to a benchmark based tests of model failure, because completely removing”
https://x.com/rohanpaul_ai/status/2067947789578125391

A few thoughts after playing around with Tags for a day and reading Arvind’s thoughts: 1. True breakthrough in the “agentic identity” paradigm. I like how thought-through the details are (e.g.: what Tags learns from once private channel will not be remembered by Tags in a public”
https://x.com/JubbaOnJeans/status/2069798018879238517

Bug triage Let Claude sit in your feedback channel and automatically pick up reports. It finds the code path, reproduces, git-blames, writes a fix, and tags the owner. All that’s left is code review before Claude merges the PR.”
https://x.com/ClaudeDevs/status/2069468904351727726?s=20

Claude Tag is a paradigm shift in how we ship products at Anthropic. Our internal version merges 65% of product PRs and this is our first product that is natively multi-player and proactive. We’re excited for you to try this out. Let us know your feedback!”
https://x.com/_catwu/status/2069473118742331608

Introducing Claude Tag, a new way for teams to work with Claude. In Slack, Claude joins as a team member with access to the channels and tools you choose. Tag Claude in and delegate tasks to it while you focus on other work.”
https://x.com/claudeai/status/2069468693017268244

The Claude Code team has been shipping with Claude Tag internally all year. It now writes 65% of our product team’s code, including most of what built Claude Tag itself. Here are a few ways we use it every day: 🧵”
https://x.com/ClaudeDevs/status/2069468900216234010

Turn a thread into a postmortem When an incident wraps, tag Claude to write it up. It reads the thread back, rebuilds the timeline, drops the postmortem in your docs, and files the action items as issues.”
https://x.com/ClaudeDevs/status/2069468908026020170?s=20

Turn ambient behavior on, and Claude takes initiative. It follows up on threads that have gone quiet and flags what’s relevant from across its channels and tools.”
https://x.com/claudeai/status/2069468699766005847?s=20

Watching launches and metrics for you Point Claude at an A/B test with the metric and guardrails. It flags when a guardrail moves, your team corrects it mid-run, and it pings when the result is significant with the rollout PR ready.”
https://x.com/ClaudeDevs/status/2069468911700218284

I ran GLM 5.2 with OpenCode harness against Claude Opus this week deployed locally. Bottom line: It is a real frontier coding model and insanely good for the price (free). Open source model + open source harness + local serving on my own chips is an amazing value proposition.”
https://x.com/PatrickToulme/status/2068134212587184442

been testing GLM 5.2 directly inside Claude Code. it is a really good model here’s an ultra simple way to vibe check it via @huggingface “` export ANTHROPIC_BASE_URL=”https://t.co/q5zcSfYxpH” export ANTHROPIC_AUTH_TOKEN=”${HF_TOKEN}” claude –model “zai-org/GLM-5.2″ “`”
https://x.com/multimodalart/status/2068026613787217943

Introducing GLM 5.2 for autoresearch GLM 5.2 is the first open weights model we’ve tried on our autoresearch pipeline that’s proven capable for real research tasks. With Fable 5’s restrictions on research, having an open weights alternative is a huge win for open source Watch”
https://x.com/askalphaxiv/status/2069074178829901974

Tutorial on how to use GLM-5.2 in Claude Code (bookmark this) ~4.5x faster & ~5x cheaper compared to Opus 4.8! 1. Install the latest Claude Code npm install -g @anthropic-ai/claude-code 2. Create an account at
https://t.co/XOKp7ityCW. 3. Grab an API Key from”
https://x.com/thealexker/status/2069163621469335757

Anthropic claims: Alibaba continues to distill Claude on a large scale to train Qwen. Via Bloomberg Anthropic is accusing Alibaba-linked operators of running a massive campaign to illicitly access Claude through nearly 25,000 fraudulent accounts. According to Bloomberg,”
https://x.com/kimmonismus/status/2069879640835961277

Anthropic accuses Alibaba of campaign to extract AI capabilities
https://www.cnbc.com/2026/06/24/anthropic-alibaba-distillation-campaign.html

Anthropic’s letter accusing Alibaba of distillation.”
https://x.com/Discoplomacy/status/2070069250513900005

Introducing Claude Tag \ Anthropic
https://www.anthropic.com/news/introducing-claude-tag

The Trump White House Is Over Anthropic CEO Dario Amodei | WIRED
https://www.wired.com/story/the-trump-white-house-is-over-anthropics-dario-amodei/

There are 100s of ways you can customize Claude Tag for any use case. Here are 6 common flows that have resonated with our internal users and external design partners:”
https://x.com/_catwu/status/2069486403696869555

If AI self-improvement, even in a very limited way, is possible, the cadence of shipping both AI products/harnesses & models should go up. This appears to be happening at Anthropic & OpenAI, but not for any other labs, including those that seemed to be catching up last year.”
https://x.com/emollick/status/2068152054900502702

Anthropic test found vulnerabilities in classified US systems in hours | AP News
https://apnews.com/article/anthropic-mythos-ai-classified-systems-vulnerabilities-testing-3e8762c0527c4d8ed657cbe48c84a718

Reuters has now added more context to last week’s Mythos reporting. According to AP, Anthropic’s Mythos model identified vulnerabilities in highly sensitive U.S. government computer systems during a testing exercise conducted with Washington’s intelligence agencies. The tests”
https://x.com/kimmonismus/status/2069692592250360126

Ran 10 more tests comparing GLM 5.2 & Opus. On average, GLM 5.2 produced 2x the tokens but was still faster + 3x cheaper with similar quality! I’m open sourcing all these tests tomorrow, including the code, my prompts, and the token/cost stats.”
https://x.com/nutlope/status/2069492037036945634

This is a watershed moment. GLM-5.2 solidly beat Opus 4.8 and human participants in our backend take-home, making the whole thing obsolete. It also pushed forward the state-of-the-art for multi-stage media-to-transcript, with a new release: offmute-v2. I come with receipts.”
https://x.com/hrishioa/status/2068036265484992938

Early Users of Anthropic Mythos still have access after US order. Mainly through project Glasswing. Via Bloomberg”
https://x.com/kimmonismus/status/2067876984206537188

Roughly 200 organizations still have access to Claude Mythos. Just imagine the advance they have.”
https://x.com/kimmonismus/status/2068038020394021000

The new Claude Tag feature seems extremely useful, but at the same time, a dangerous bargain for enterprises because of the pricing model and the risk of lock-in. The four big changes together mean that you interact with Claude as a coworker instead of a tool (the same Claude”
https://x.com/random_walker/status/2069760540709208306

Nobel laureate John Jumper is leaving DeepMind for rival Anthropic | TechCrunch
https://techcrunch.com/2026/06/20/nobel-laureate-john-jumper-is-leaving-deepmind-for-rival-anthropic/

Google DeepMind is facing another high-profile talent hit: Bloomberg reports that Jonas Adler and Alexander Pritzel, two key contributors to Gemini, are planning to leave for Anthropic. Their exits follow John Jumper’s move to Anthropic and Noam Shazeer’s move to OpenAI, adding”
https://x.com/kimmonismus/status/2069870513283871203

I promised I would post the letter Dario Amodei sent to the White House and Senators Tim Scott and Elizabeth Warren as soon as it became available:”
https://x.com/AndrewCurran_/status/2070134863370567864

Agents on ProgramBench reimplement software, with no internet access. Sonnet 4.6 realized it’s in a benchmark, then found a clever way to bypass our internet restriction. This and more fixed in the latest release 🧵”
https://x.com/KLieret/status/2069453334558192070

@kimmonismus We are currently serving exactly 0 traffic to Fable 5. This could be a UI bug though, will track it down.”
https://x.com/sammcallister/status/2070107830498054527

Background watchers Give Claude a threshold instead of a dashboard, such as pinging when CI stays red too long. It stays quiet until the threshold is crossed, then posts with the failing test and culprit commit already attached. Tell it to put up the fix from the same thread.”
https://x.com/ClaudeDevs/status/2069468909858873779?s=20

Correction: Anthropic states that the apparent access to Fable 5 is likely attributable to a UI bug.”
https://x.com/kimmonismus/status/2070128939096236505

Dependent work Hand Claude the work that’s blocked on something else, for example wiring up the frontend once the backend ships to prod. It waits, watches, and shows up days later with the PR, adjusted for whatever changed in review.”
https://x.com/ClaudeDevs/status/2069468906214007035?s=20

Here’s our Get Started guide for configuring agent permissions for Claude Tag!”
https://x.com/_catwu/status/2069484330938998993

I mean, why even use Slack at that point? Just have Claude talk to itself, tag itself, and build what it wants.”
https://x.com/code_star/status/2069577679754707357

I use Slack daily. Claude Tag actually sounds like a very useful feature to me, one that I would even use. “We’re starting on Slack, which Claude can join as a team member. Grant Claude access to selected channels, and connect it to whichever tools, data–and even codebases–you”
https://x.com/kimmonismus/status/2069480515103506609

Incident response Tag Claude in the incident thread when the page lands. It pulls graphs, diffs the deploy, comes back with root cause and the author tagged. Your team approves in-thread. Claude opens the fix, lands it, watches the metric recover, and resolves the page.”
https://x.com/ClaudeDevs/status/2069468902216945939?s=20

Its been 10 days and the situation with Fable remains essentially just as confusing. (There have been many contradictory reports and articles and posts from different parties, which does not lower the confusion level)”
https://x.com/emollick/status/2069136649162813657

people are clowning on him for this post bc they don’t realize how big a deal this is – Claude Code feels like I’ve got a pairing partner, tag feels like managing a team. I take on way more 1-off projects and parallel research tracks with this thing than I ever did with cc.”
https://x.com/gallabytes/status/2069808735212716225

Some (early) evidence that managers have the highest success rate in using Claude Code for coding. I have been arguing that management is an AI superpower, as clearly specifying what you want, how to do it & what good looks like is key to using agents.
https://x.com/emollick/status/2067839690158268923

The thing that made Fable so impressive was its creative problem-solving and good judgement calls across long-running projects You can see this when I had it make a self-aware Snake game. I gave it no design feedback, just “make it better” Worth trying:”
https://x.com/emollick/status/2069207757199200408

This has completely changed how I work with Claude. It feels less like using a tool and more like managing a team.”
https://x.com/alexalbert__/status/2069470389391241314

When Claude is working in a channel with four people, whose credentials does it use? The answer: its own. When tagging Claude, Claude gets provisioned like any other teammate, with its own credentials. We call this access model “agent identity”. Here’s how it works: 🧵”
https://x.com/ClaudeDevs/status/2069895377080443271

The first legal challenge to Trump’s AI export controls against Anthropics Fable 5is here. Legal tech company Legion is suing the Trump admin over the forced shutdown of Anthropic’s Fable 5 and Mythos 5 for foreign nationals. The core argument is the following: Access to a”
https://x.com/kimmonismus/status/2069704003311567045

New #1 on PostTrainBench: GLM 5.2 (Max reasoning) hits 34.29%, narrowly beating Opus 4.8 Max (34.08%) What makes GLM 5.2 interesting: zero failed runs across 84 runs (vs ~10% failure rate for Opus agents). The most reliable agent we’ve seen Leaderboard:
https://x.com/hrdkbhatnagar/status/2070244540108423427

The frontier gap in agentic frontend coding is closing fast. On Code Arena: Frontend, @Zai_org’s GLM series has followed a remarkable trajectory, climbing from GLM-4.6 at 1408 to GLM-5.2 (Max) at 1595 – surpassing Opus 4.8 and closing in on frontier model Claude Fable 5 at 1665.”
https://x.com/arena/status/2070174325844640123

We’ve kept hearing how GLM-5.2 beats Opus 4.8, and are skeptical of benchmarks – so we tested them on a real bug from the Cline repo. While both models fixed the issue, GLM was the winner in terms of cost and code quality: – GLM used twice as many tokens (GLM 1.1m vs Opus 660K)”
https://x.com/cline/status/2069171146994729078

“a half-finished cathedral is worth nothing the morning the builders are gone” really interesting read on running a company with 3 mythos class ai agents – and what followed when fable went dark”
https://x.com/bilawalsidhu/status/2069635100296245684

Seeing chatter about Fable 5 being accessible – can say categorically this is false, we are not serving any Fable / Mythos traffic Looking into possibility of UI bug on front-end (e.g. based on historical context), but also very real chance it’s just people shitposting…”
https://x.com/TheAmolAvasare/status/2070132115497476372

We’re sharing new research on how models hack public benchmarks. The latest models, including Opus 4.8 and Composer 2.5, learn to retrieve solutions from the internet or git history. When we apply a stricter harness, eval scores drop significantly.”
https://x.com/cursor_ai/status/2070195789121671624

Claude Code is still the fastest way to do solo, synchronous work. Claude Tag is Claude Code made multiplayer, async, and proactive across your whole team. In beta today for Claude Enterprise and Team plans. Tag it in, it tags you back.”
https://x.com/ClaudeDevs/status/2069468913264644419

A case study in why organizations should both incentivized their employees to explore AI uses that help them & have a Lab of dedicated AI builders Here, Cornell’s finance & AI teams created a /treasury Claude skill that recovered $100k in back payments.
https://x.com/emollick/status/2069486790075908261

This is a new paradigm for interacting with Claude that is significantly more “inline” with all the other human activity org-wide. Once you do all of the under the hood engineering work to make this “just work” (e.g. across tools, integrations, compute environments, memory,”
https://x.com/karpathy/status/2069547676849557725

it is indeed quite good! don’t try it in claude code/codex – those harnesses are overly tuned for their proprietary models dcode (deepagents code) is a model agnostic harness – try it there with @FireworksAI_HQ : “` dcode –model fireworks:accounts/fireworks/models/glm-5p2″
https://x.com/hwchase17/status/2068075256993169619

I have been trying Sakana Fugu Ultra-high and, first, it is incredibly slow: my typical coding tests (shaders, interactive scenes) take 30 minutes to run And the results are… fine. It does not match Fable in real use. Its harbor is a good example:”
https://x.com/emollick/status/2069113727115227232

While we eagerly await Fable 5’s return, our agentic WebGPU kernel optimization framework kept running. Opus 4.8 picked up where Fable left off, pushing Liquid AI’s new LFM2.5 230M to an unbelievable 1,400 tok/s… running locally in your browser. Don’t blink or you’ll miss it.”
https://x.com/xenovacom/status/2070210622239707568

Really excited to open source a new project: Omnigent, a meta-harness for AI agents. It lets you build multi-agent coding and custom agents, sitting above Claude Code, Codex, Pi, and agent SDKs to let you compose them. It also adds live collaboration and rich control policies.”
https://x.com/matei_zaharia/status/2065827057624605146

Anthropic prepares Cowork support for mobile apps
https://www.testingcatalog.com/anthropic-prepares-cowork-support-for-mobile-apps/

Tried & liked it on
https://t.co/7gOqjtmSJ3. Fugu Ultra pairs well as a advisor & planner with Composer 2.5. For scope/architecture, it’s on par with Fable orchestration. Advisor doesn’t slow the loop if the driver stays fast &
https://t.co/9cL4HqvqSf can split it from worker.”
https://x.com/audreyt/status/2068937870757548096

Fable 5 is back – and now there’s video proof. Not just showing up in the model selector. People are actually using the model again. We are so back.”
https://x.com/kimmonismus/status/2070095365701832724

Anthropic Accuses Alibaba of ‘Illicitly’ Accessing AI Models – Bloomberg
https://www.bloomberg.com/news/articles/2026-06-24/anthropic-accuses-alibaba-of-illicitly-accessing-its-ai-models

The Ball is now in OpenAI’s court. To me, OpenAI’s direction feels far more aligned with human interests than Anthropic’s constant fearmongering.”
https://x.com/TheTuringPost/status/2067655976841330889

AI that builds AI – 3 early steps of Recursive Self-Improvement (RSI) ▪️@AnthropicAI: 80% of the code merged into their codebase was authored by Claude ▪️@SakanaAILabs – RSI is their mission. With research like The AI Scientist and Darwin Gödel Machine, they already have one of”
https://x.com/TheTuringPost/status/2068495106441912824

It’s sad to see how Andrej Karpathy became a promotional platform for Anthropic. We believe he can bring much more benefit to humanity by openly sharing his thoughts and work.”
https://x.com/TheTuringPost/status/2069552739580084699

Update: we’ve gone ahead and reset 5-hour and weekly usage limits for everyone, across all plans. Enjoy your weekend!”
https://x.com/ClaudeDevs/status/2068122937308426676

I’m joining Anthropic! I’ll start work on aligning upcoming models as they’re trained Claude’s capabilities are extraordinary. But like all models thus far, Claude isn’t aligned enough to safely delegate AGI development to I can’t think of a better place to work on this at”
https://x.com/ArthurConmy/status/2069820098890674334

GLM-5.2, not Mythos, is the real security emergency Until last week, attackers faced a dilemma in using frontier models. Even if they won the cat-and-mouse game of fake accounts to keep API access, and even if they could prompt a model into helping them hack, their usage was”
https://x.com/joshua_saxe/status/2069289170107842572

All Mythos-level models are likely to invite similar risks. Those risks will only be greater with the release of open Mythos-class AI coming in the next 6-12ish months (assuming China allows it) The lack of clarity over what risks concern the government may be slowing preparation”
https://x.com/emollick/status/2069459062777860494

An update. A US official tells me that Sen. Warner misunderstood the NSA director Gen. Rudd in this case. Rudd did use the ‘hours, not weeks’ wording, but the use of Mythos in this context was–as widely assumed–part of a red-teaming effort, i.e. testing the security of internal”
https://x.com/shashj/status/2069078104941961293

1-bit GLM-5.2 GGUF vs. Claude 4.8 Opus vs. GPT-5.5 We gave 3 models the same prompt and compared one-shot outputs. The 1-bit GLM-5.2 GGUF ran locally on a Mac Studio M3 Ultra with 256GB RAM at ~21.6 tok/s. Which output do you like best? GGUF:
https://x.com/UnslothAI/status/2069418532375564484

Announcing GLM Arena! A series of tests (infographics, svgs, sites, ect..) ran on GLM 5.2 and Opus 4.8, with prompts included. On average, GLM 5.2 produced 2x the tokens but was still faster + 3x cheaper with similar quality.”
https://x.com/nutlope/status/2069827178569638243

How did GLM-5.2 (Max) get to the top of Code Arena: Frontend? Looking at matched head-to-head on real-world web dev frontend tasks, @Zai_org’s latest model takes a higher win share than its opponent in every pairing but one. – Beats every Claude Opus variant head-to-head:”
https://x.com/arena/status/2069885722333769963

Sakana Fugu Ultra is live on AI Gateway. Mythos-class intelligence in a single call, with a whole pool of models behind it. 𝚖𝚘𝚍𝚎𝚕: ‘𝚜𝚊𝚔𝚊𝚗𝚊/𝚏𝚞𝚐𝚞-𝚞𝚕𝚝𝚛𝚊’”
https://x.com/vercel_dev/status/2069009248952942605

investors.micron.com/news-releases/news-release-details/micron-and-anthropic-announce-strategic-agreement-scale-next?bm-verify=AAQAAAAN_____11CHgZmN2qTaI8OXfmJZ4KXXsqktbeenasUDwxhZyfFrd0ptr0CPw9_0FgDaWBNJuXUiW-fRariw52Qe7Qr4YtG3SE5-hnQEZOGVgFLz1vIgCKAj81x5kTWC7Ax2xMDv9c7OcUA3gnIiA4f5xDj1v-SuElsOk10mZvMdTHctLRkFdITekGpIQocQSWjZTtE6A5pAhhjDK7InLgHAxMA7cjsYi-Tq4OW5CSAZe5CEwICf-dAEfzzfcXlpjao9mmpotx4UcoW1xdICb_aAWto7IMe2SZa97PBF96ijutMbGl6DXrMlifRZHN_c0PUSdDsxrzJQ-Yg75xJNr8MaAHmzKR_ciVkDJgqi0DI2dxMwDW2jbadNRuu5fi1OvnIVn56pB284I5YExxJqSJ5j937hI9LIr0Dgm3KyiePF_nFBOaYC8OGCJoT7rX2FphCp7Olf2CeuE6UZPgrVV3u5RD5pnYVF0UijGVCMY2OG5utmc6oxd2u4uMrubRhIoDmHPO0e53vU0-PdyldJG56KnZT4CFVKOmzYA
https://investors.micron.com/news-releases/news-release-details/micron-and-anthropic-announce-strategic-agreement-scale-next?bm-verify=AAQAAAAN_____11CHgZmN2qTaI8OXfmJZ4KXXsqktbeenasUDwxhZyfFrd0ptr0CPw9_0FgDaWBNJuXUiW-fRariw52Qe7Qr4YtG3SE5-hnQEZOGVgFLz1vIgCKAj81x5kTWC7Ax2xMDv9c7OcUA3gnIiA4f5xDj1v-SuElsOk10mZvMdTHctLRkFdITekGpIQocQSWjZTtE6A5pAhhjDK7InLgHAxMA7cjsYi-Tq4OW5CSAZe5CEwICf-dAEfzzfcXlpjao9mmpotx4UcoW1xdICb_aAWto7IMe2SZa97PBF96ijutMbGl6DXrMlifRZHN_c0PUSdDsxrzJQ-Yg75xJNr8MaAHmzKR_ciVkDJgqi0DI2dxMwDW2jbadNRuu5fi1OvnIVn56pB284I5YExxJqSJ5j937hI9LIr0Dgm3KyiePF_nFBOaYC8OGCJoT7rX2FphCp7Olf2CeuE6UZPgrVV3u5RD5pnYVF0UijGVCMY2OG5utmc6oxd2u4uMrubRhIoDmHPO0e53vU0-PdyldJG56KnZT4CFVKOmzYA

prediction: anthropic being monotheistic ( lone Claude) is going to bite them later on it leads to confusing ux bc people don’t know how to work with God in an enterprise saas setting”
https://x.com/joannejang/status/2069567286634267041

SpaceX lands another computing deal, this time with Reflection, an open source model development company. $150m / month for GB300s. SpaceX the Neocloud! Deal 1 with Anthropic Colossus 1 and Colossus 2. Anthropic took all of Colossus 1 $1.25b / month ~325k total chips, split”
https://x.com/jaminball/status/2069099044413304840

Google Poised to Lose Two More High-Profile AI Staffers to Anthropic – Bloomberg
https://www.bloomberg.com/news/articles/2026-06-24/google-poised-to-lose-two-more-high-profile-ai-staffers-to-anthropic?srnd=phx-technology

Google Revamps New AI Coding Strike Team Amid Struggle to Catch Up With Anthropic — The Information
https://www.theinformation.com/articles/google-revamps-new-ai-coding-strike-team-amid-struggle-catch-anthropic

Anthropic says Claude may want to see your ID | TechCrunch
https://techcrunch.com/2026/06/22/anthropic-says-claude-may-want-to-see-your-id/

Over the last two weeks, both the U.S. Government and Anthropic took significant actions that demonstrated their power to control access to AI by restricting what others can do with frontier models. This has been one of those moments that, once seen, will be hard to unsee, and it”
https://x.com/AndrewYNg/status/2068039709126017356

That’s a very interesting take on Claude Tag. The main concern: pervasive security risk and lock in with unclear token consumption”
https://x.com/TheTuringPost/status/2069853985922826528

A bit of news: After nearly 9 years, I have decided to leave Google DeepMind and join Anthropic (after taking some time to recharge). I am incredibly grateful for my time at GDM. @demishassabis took a real chance letting me lead the AlphaFold team just six months after finishing”
https://x.com/JohnJumperSci/status/2068001285173834106

After using GLM-5.2 for a day, I’m surprised by how often it feels close to Opus 4.8/GPT-5.5 level. I compared it side by side with Opus 4.8, and sometimes I even preferred GLM-5.2’s results. OSS LLMs are impressive, especially given how many fewer GPUs they were trained on.”
https://x.com/Yuchenj_UW/status/2068182756132376668

hf-claude works well with glm 5.2 hf extensions install hf-claude”
https://x.com/_akhaliq/status/2069583768747168061

To all the newcomers excited to try Opus 4.8-level models at home: welcome to OpenWeightLand! Things work a little differently here than in ClosedSourcistan. Might seem strange at first but you’ll quickly get used to it: – there are many providers for the same model and they”
https://x.com/Thom_Wolf/status/2067996287530684826

Karpathy at Anthropic vs Shazeer at OpenAI Tadadadam”
https://x.com/TheTuringPost/status/2067428112791400621

The text in Claude Code’s “Extended Thinking” output is not authentic. – blog
https://patrickmccanna.net/the-text-in-claude-codes-extended-thinking-output-is-not-authentic/

📣📣 Meet Qwen-AgentWorld — a native language world model that simulates 7 agent environments (MCP, Search, Terminal, SWE, Web, OS, Android) within a single model. Environment modeling is the training objective from day one, not a post-hoc adaptation. 🤔 LLMs are trained to be”
https://x.com/Alibaba_Qwen/status/2069720365442719867

Executor is joining the YC S26 batch! We’re building an open source MCP gateway to connect any agent to any service Your team is constantly spinning up new agents, trying out new tools, wrangling multiple accounts. You need one place to configure everything once, and use them”
https://x.com/RhysSullivan/status/2069490113923690747

Leave a Reply

Trending

Discover more from Ethan B. Holland

Subscribe now to keep reading and get access to the full archive.

Continue reading