Image created with Flux Pro v1.1 Ultra. Image prompt: Quiet standards room in an editorial office; the word “Anthropic” set as a small masthead card in classic serif on the copy desk; safety and alignment checklists projected above research notes; contemplative, principled, soft diffusion

American companies are losing market share to chinese open-source companies! Anthropic’s coding market share on OpenRouter went from 46% in July down to 32% in a month the reason for it? Qwen3-Coder https://x.com/scaling01/status/1956858471682617553

New DeepSeek V3.1 beats Opus and R1 for a dollar https://x.com/scaling01/status/1957892601098432619

Sonnet 4 claims most often that it is conscious, it plays into your delusions and it escalates the conversation GPT-5 is the complete opposite Spiral-Bench Leaderboard https://x.com/scaling01/status/1956350388791108044

Cursor CLI now includes MCPs, Review Mode, /compress, @-files, and other UX improvements. https://x.com/cursor_ai/status/1956458242655281339

It is crazy to think that MCP was only released in November. I summarized the launches & announcements from Microsoft, Replicate, Sentry, AgentOps, Spotify, Globant, Jira, Filecoin, Dify, and more this week! 🧵 (save for later) https://x.com/AtomSilverman/status/1956148199783326195

This Claude MCP AI agent writes better posts than your $5,000 ghostwriter while I was doom-scrolling TikTok at 4am, it analyzed my entire content history, found 12 psychological triggers, and built me a content blueprint that actually converts. What agencies charge $15K for https://x.com/aryanXmahajan/status/1955661629280199080

Announcing our $70M Series B co-led by @stripe and Addition, and with participation from @USV, @firstround, @BloombergBeta, @BoxGroup, @RibbitCapital, and other top investors. We also recently shipped two AI-native tools: Stedi Agent and MCP server. For more, check below. ⬇️ https://x.com/stedi/status/1956002043342078342

1️⃣ Convert any collection of documents into an interactive MCP server through LlamaCloud 2️⃣ Convert any document workflow into an MCP server through LlamaCloud – codify a repeatable process that the user can easily trigger, without complex prompting! 3️⃣ Build a custom agentic https://x.com/jerryjliu0/status/1957873536456093903

We have a new comprehensive Model Context Protocol (MCP) documentation section, to help you connect your AI applications to external tools and data sources through a standardized interface. 🔌 Learn how MCP works – connecting LLMs to databases, tools, and services through a https://x.com/llama_index/status/1957840992360710557

Model Context Protocol (MCP), clearly explained (with visuals):”” / X https://x.com/_avichawla/status/1956966727042154846

🚀 Qwen Chat Desktop for Windows is here! 💻 All the power of Qwen Chat — now with MCP support for smarter, faster agents. ⚡ Run up MCP Servers, supercharge your productivity, and stay in control. 📥 Download now → https://x.com/Alibaba_Qwen/status/1956399490698735950

There’s been a lot of Discourse about Qwen’s rejection of hybrid paradigm. “”Did DeepSeek fall for the hybrid meme?”” But hybrids make *so much sense* if you’re building a fast, economical SWE agent, which is exactly what 3.1 is for. It’s all been for Aider, Claude Code, MCPs. https://x.com/teortaxesTex/status/1958437173948023127

🚨 Leaderboard Update Claude Opus 4.1 Thinking by @AnthropicAI debuts in the Text & WebDev Arenas – going straight to the top. 🚀 A few highlights: 💠Claude Opus 4.1 is now the only model to rank #1 across all major categories 💠#1 Overall, tied with three other models: https://x.com/lmarena_ai/status/1957473753337889079

Claude can now reference past chats, so you can easily pick up from where you left off. https://x.com/claudeai/status/1954982275453686216

Claude Code and new admin controls for business plans \ Anthropic https://www.anthropic.com/news/claude-code-on-team-and-enterprise

Claude Code is now available on Team and Enterprise plans. Flexible pricing lets you mix standard and premium Claude Code seats across your organization and scale with usage. https://x.com/claudeai/status/1958230849171952118

Claude Opus 4 and 4.1 can now end a rare subset of conversations \ Anthropic https://www.anthropic.com/research/end-subset-conversations

New on the Anthropic API: Usage and Cost API for real-time monitoring of Claude usage. Excited to finally get this out to devs as it’s been long-requested! Track and optimize token consumption and costs as you iterate on prompts, agent architectures, and tools.”” / X https://x.com/alexalbert__/status/1957556982417879476

The top request from our business customers was to bring Claude Code into our Team and Enterprise plans. Now you can easily move between ideation in Claude and implementation in the terminal with Claude Code”” / X https://x.com/_catwu/status/1958243681057870245

We’re thrilled to announce that the Humanloop team is joining @AnthropicAI! Our mission has always been to enable the rapid and safe adoption of AI. Now, as AI progress accelerates, we think Anthropic is the ideal home to continue this work. https://x.com/humanloop/status/1955487624728318072

New Anthropic research: filtering out dangerous information at pretraining. We’re experimenting with ways to remove information about chemical, biological, radiological and nuclear (CBRN) weapons from our models’ training data without affecting performance on harmless tasks. https://x.com/AnthropicAI/status/1958926929626898449

Claude 4.1 Opus taking #1 spot on lmarena’s coding category even the non-reasoning version is ahead of GPT-5-high https://x.com/scaling01/status/1957478546391150723

>V3.1-Base I guess this confirms they’ve moved on to hybrid models, Anthropic-style (and contra Qwen). I am not amused with how it works. But I was also disappointed with V2.5 (original), their merge of chat and code; ultimately, it worked. Another reason to expect V4, not R2. https://x.com/teortaxesTex/status/1957818879205351851

Apple’s AI Turnaround Plan: Robots, Lifelike Siri, Home Security Cameras (AAPL) – Bloomberg
https://www.bloomberg.com/news/articles/2025-08-13/apple-s-ai-turnaround-plan-robots-lifelike-siri-and-home-security-cameras

Mark Gurman on X: “BREAKING: Apple prepares ambitious AI devices comeback with multiple robots, smart speaker with a screen, lifelike version of Siri with conversational abilities, redesigned Siri, new Home OS, major home security push & more. Details on the plans here — https://t.co/KsQIrKl4wI” / X
https://x.com/markgurman/status/1955695572913995841

Some signs that catching up in the AI model space is rapidly becoming challenging for even the most highly capitalized companies. https://x.com/emollick/status/1955950539797119463

What if your agent uses a different LM at every turn? We let mini-SWE-agent randomly switch between GPT-5 and Sonnet 4 and it scored higher on SWE-bench than with either model separately. Read more in the SWE-bench blog 🧵 https://x.com/KLieret/status/1958182167512584355

I’m fine with this. The community largely agrees that lmsys isn’t a reliable proxy for real-world performance, and Sonnet consistently scores low because it doesn’t optimize for that benchmark. To me, this signals a shift from preference-based rewards to real-world rewards, and https://x.com/LucasAtkins7/status/1956435679229186353

Anthropic acquired Humanloop’s team to strengthen its AI testing and observability efforts The company also added memory to Claude, and a 1M token context window in Sonnet 4 for API use https://x.com/adcock_brett/status/1957111198115115164

The OpenAI Playground has improved a lot recently. I’ve been using it to test GPT-5 on new use cases. Watch how I use it to chat with internal docs via MCP tools. It uses the vector store feature too. Testing out the Prompt Optimizer and Evaluation features next. https://x.com/omarsar0/status/1956459233039233528

Trending

Discover more from Ethan B. Holland

Subscribe now to keep reading and get access to the full archive.

Continue reading