Image created with Gemini. Image prompt: A flat acrylic illustration of a person standing beside a turquoise swimming pool with hand-drawn white ripple squiggles, the entire scene fragmented into a grid of slightly offset and overlapping rectangular photo panels that show the same view from different angles and moments, hard-edged shapes with terracotta pink deck and lawn green grass under a hot yellow sun, no shadows, flat saturated colors, tablet-drawn line with visible digital stroke, high horizon and generous white sky.
Anthropic Embeds Engineers in the NSA to Deploy Mythos
https://www.implicator.ai/anthropic-embeds-engineers-in-the-nsa-to-deploy-mythos-for-offensive-cyber/
Anthropic Walks Back Policy That Could Have ‘Sabotaged’ AI Researchers Using Claude | WIRED
https://www.wired.com/story/anthropic-responds-to-backlash-on-claudes-secret-sabotage-on-ai-research/
Anthropic’s new Fable 5 safeguards are fascinating. When the model is used for frontier LLM development, it apparently does not simply refuse or warn the user. Instead, it quietly limits its own effectiveness through techniques like prompt modification, steering vectors, and
https://x.com/kimmonismus/status/2064417460715962479
I believe what Anthropic is doing, gating the ability to do certain harmless things like LLM research, and with incredibly sensitive filters that even medical questions are often blocked, is *deeply* wrong. They got open research, the Transformer, GPT2, …
https://x.com/antirez/status/2064766431531532588
I don’t really want to have to go to bat against Anthropic, but they’ve just been unnecessarily antagonistic to all of China, then not so subtly to open weight models, and now more broadly open AI research. What’s next on the list?
https://x.com/natolambert/status/2064412173527556298
I got a good nights sleep and I’m still just as angry about Anthropic’s choices. I enjoy working in AI so much and to have my access to the cutting edge models for my work rugpulled in an under the table fashion is appalling. I expected to be restricted eventually, but not
https://x.com/natolambert/status/2064699044145095104
I think it is really worth reading this piece on RSI at Anthropic. There is a bit of navel-gazing, some marketing, and a lot of very sincere beliefs about what Anthropic thinks is likely in the near future of AI that you probably want to be aware of.
https://x.com/emollick/status/2062582362194460698
If Claude Fable stops helping you, you’ll never know — Jonathon Ready
https://jonready.com/blog/posts/claude-fable5-is-allowed-to-sabotage-your-app-if-youre-a-competitor.html
In good faith and with no judgment (mistakes happen), I truly hope that Anthropic will hear the feedback and change course on this. Anthropic is a company that has been raising awareness about AI manipulation which is a very important topic! You don’t want to go down as the
https://x.com/ClementDelangue/status/2064673792303955985
Makes me wonder how long this has already been going on without users being notified. I had been feeling like codex was running circles around Claude code for months now. Now I wonder if Claude code was just self nerfed. Regardless of the intentions behind this, this is a bad
https://x.com/code_star/status/2064464447662707180
Microsoft AI head calls out Anthropic for acting like Claude is conscious | The Verge
https://www.theverge.com/tech/947197/microsoft-ai-mustafa-suleyman-anthropic-claude-conscious
My last observation re: Anthropic’s secret sabotage safety policy, is that it undermines actually good safety policy. How? 1. First, it is very plausible to describe this as anti-competitive behavior (even if you are maximally sympathetic to Anthropic here you must admit this),
https://x.com/deanwball/status/2064665679307985244
Policy on the AI Exponential \ Anthropic
https://www.anthropic.com/policy-on-the-ai-exponential
Re the Fable ML sandbagging, the model’s AI research capabilities were probably at least partly trained on Anthropic employees diffing atop proprietary algos and infra. So the IP leak is somewhat like a researcher who knows Anthropic’s stack getting poached to another lab.
https://x.com/dwarkesh_sp/status/2064826554442719502
SITUATION UPDATE: Anthropic is reversing its Fable 5 policy of covertly degrading performance for competing AI researchers, per Wired.
https://x.com/MTSlive/status/2064922000020398331
That was quick: Anthropic reversed a controversial policy that would have secretly degraded Claude Fable 5 for users doing frontier AI research after backlash from researchers who saw it as covert sabotage of competing AI development.
https://x.com/kimmonismus/status/2065003618710008084
Things I really dislike about Fable: 1. Anthropic collects my prompt history, stores it, and does whatever they want with it for 30 days. No opt-out 2. They can nerf their most expensive model without telling me, billing me the same amount, wasting my time. Whenever they want
https://x.com/GergelyOrosz/status/2064618497150210391
Very pleased to hear Anthropic have walked back this policy
https://x.com/simonw/status/2064918665859080392
When AI builds itself \ Anthropic
https://www.anthropic.com/institute/recursive-self-improvement
When Fable 5 is used for frontier LLM development, it does not notify the user and instead limits the model’s capabilities through methods such as prompt modification, steering vectors, and PEFT. Anthropic estimated that this would affect approximately 0.03% of traffic.
https://x.com/Hangsiin/status/2064397550434816088
Synthetic performers’ in ads must be identified as AI as new New York law takes effect | AP News
https://apnews.com/article/new-york-ai-law-hochul-synthetic-performers-e433625bfb61c8abeab0d619869192ed
Built to benefit everyone: our plan | OpenAI
https://openai.com/index/built-to-benefit-everyone-our-plan/
Confidential submission of draft S-1 to the SEC | OpenAI
https://openai.com/index/openai-submits-confidential-s-1/
Exclusive: OpenAI Preps New AI Model, Expects To Go Public ‘Within the Next Year’ — The Information
https://www.theinformation.com/briefings/exclusive-openai-preps-new-ai-model-expects-go-public-within-next-year
Here is our current plan for OpenAI:
https://x.com/sama/status/2064088940932641225
Introducing the OpenAI Economic Research Exchange | OpenAI
https://openai.com/index/economic-research-exchange/
Lockdown Mode | OpenAI Help Center
https://help.openai.com/en/articles/20001061-lockdown-mode
Lockdown mode is now available in ChatGPT. We rolled this out for organizations a few months ago, and now it’s available for all users on all plans. Lockdown Mode is designed to help prevent the final stage of data exfiltration from a prompt injection attack by limiting
https://x.com/cryps1s/status/2062923575049531422
much better ChatGPT memory:
https://x.com/gdb/status/2062608071411540196
OpenAI files confidential SEC S-1 paperwork for IPO | Fortune
https://fortune.com/2026/06/09/openai-files-confidential-s-1-sec-ipo/
The goals we’re working towards at OpenAI, to achieve the OpenAI mission and expand human agency as AI progresses:
https://x.com/gdb/status/2064093657888960998
Trump administration, OpenAI discussing possible government stake
https://www.cnbc.com/2026/06/05/trump-open-ai-altman-stake.html
We’ve been researching new ways for ChatGPT memory to carry context across conversations and keep it useful over time. Today, that work is rolling out as a more capable memory system in ChatGPT.
https://x.com/OpenAI/status/2062567556524003631
Art Directors Guild Slams Martin Scorsese for AI Partnership
https://variety.com/2026/film/news/art-directors-guild-statement-martin-scorsese-ai-1236770996/
Today I’m publishing a new essay, Policy on the AI Exponential. AI is progressing extremely fast–much faster than the policy process was built to handle. The essay lays out where I think the technology is now, and the action needed to close the gap:
https://x.com/DarioAmodei/status/2064781775247950326
Mythos 5 speeds up its own training by over 69 times a human expert needs 4-8h eq. for a mere 4x speedup
https://x.com/scaling01/status/2064392809293939119
AI is advancing at a pace our policymaking institutions were never built for–and the gap between the two is becoming the central challenge of the technology. In his latest essay, our CEO Dario Amodei lays out how to close it. We’re launching three new initiatives to support the
https://x.com/AnthropicAI/status/2064783418844762489
Cybersecurity and biosecurity requests may auto-reroute to Opus 4.8 (shown in the UI, billed at Opus prices). Docs + prompting guide:
https://x.com/ClaudeDevs/status/2064394931033248226
Dario Amodei — Policy on the AI Exponential
https://darioamodei.com/post/policy-on-the-ai-exponential
Fable 5 lies 96% of the time. We were surprised by it’s skill… 🧵
https://x.com/kradleai/status/2064907897373642912
First Aronofsky, then Scorsese, and now Gareth Edwards. Generative media has been polarizing – but is the Overton window shifting before our eyes?
https://x.com/bilawalsidhu/status/2062358867510440354
I am curious if the new Fable safegaurds would trigger accidentally on meaningfully important work as a false positive that has a real life consequence. Just realizing this is eroding trust and please change this to just down right blocking which is a fine position to take.
https://x.com/_arohan_/status/2064644778147643401
I think we’ve reached the point where normal people can’t really determine whether new models are better than previous ones. Like Fable doesn’t seem that much better to me, but every 150 IQ person I know is like “wow the singularity came sooner than I thought”.
https://x.com/citrini/status/2064480613852201336
I’ve been testing something after @OliviaHelenS noticed you can’t even say “”Hi”” to Fable if you’re a biologist. I checked, and several of us are able to interact with Fable in Incognito Mode, but not in normal mode. This didn’t happen to our non-biologist friends.
https://x.com/cremieuxrecueil/status/2064449457869984035
One interesting pattern with Fable 5 is that it will often say things that are gibberish when I use it for coding. Things like “”The morning’s slim-scan fix cured the scan hang””, “”this is a latent-drift API-shape wrinkle””, etc. When I ask why it does this, Fable explains that it
https://x.com/tamaybes/status/2065147305494450248
One thing I mentioned only in passing in my Fable post is that, for long running tasks, Fable starts to develop its own dialect as its many agents and tasks reinforce themselves and make Claudish language ever more Claudish. You need to ask it to report out in plain English.
https://x.com/emollick/status/2064542441848422611
The right term for this is supply chain attack. Depending on your use case, the Fable model will be a malware.
https://x.com/deliprao/status/2064485687374569897
why not just refuse the prompt? why so sneaky?? @AnthropicAI
https://x.com/DBahdanau/status/2064692204287799728
anthropic doesn’t owe anyone “”frontier capabilities””. none of the labs do. they are all simply selling a product, or a story, that people pay for. that aside, the more telling bit is how far anthropic is willing to go to secure a narrative around “”capability slowdown””, post a
https://x.com/suchenzang/status/2064452548753559644
Anthropic just delegated a lot of European companies to the permanent underclass if Anthropic saves data for Claude Mythos and Fable 5 for 30 days, then all companies that require zero data retention simply can’t use them
https://x.com/scaling01/status/2064685085379477742
As my entire feed is criticizing Anthropic, I think that the team there genuinely believes what they’re saying. It’s not a marketing/anticompetitive tactic. They genuinely believe these models are dangerous and that AI research should be slowed down.
https://x.com/finbarrtimbers/status/2064427031543341450
asked claude fable 5 to optimise the inference on my local gemma 4 12b setup and just got a call from FBI
https://x.com/dejavucoder/status/2064420742129967331
Both Anthropic and OpenAI mention the possibilities of slowing AI development in their latest “”what comes next”” in AI posts, but say they need to be an action coordinated across the entire world using as-yet-unidentified methods.
https://x.com/emollick/status/2064158792145609114
Claude Fable 5 and new safety fables – by Nathan Lambert
https://www.interconnects.ai/p/claude-fable-5-and-new-ai-safety
Data retention practices for Mythos-class models | Claude Help Center
https://support.claude.com/en/articles/15425996-data-retention-practices-for-mythos-class-models
Exclusive | OpenAI Considers Drastic Price Cuts, Anticipating War for Users With Anthropic – WSJ
https://www.wsj.com/tech/ai/openai-considers-drastic-price-cuts-anticipating-war-for-users-with-anthropic-9b8c178e
Huge: OpenAI is considering drastically lowering the prices it charges users as it seeks to win customers from its rival Anthropic. The company is weighing significant cuts to what it charges for tokens, the unit of measurement AI firms use to bill for their products, according
https://x.com/kimmonismus/status/2065043333941207160
Measuring LLMs’ impact on N-day exploits \ Anthropic
https://www.anthropic.com/research/n-days
Microsoft restricts Claude Fable for employees over data retention concerns | The Verge
https://www.theverge.com/report/947575/microsoft-claude-fable-5-restricted-internally
Reports claim Claude’s API may have returned another user’s inference output during today’s outage. Anthropic’s status page confirms elevated errors affecting Claude API, Claude Code, Claude. ai and Claude Cowork but Anthropic has not confirmed a customer data leak yet. That
https://x.com/kimmonismus/status/2062997809067139468
The word “cancer” is flagged as a biosecurity risk by Claude Fable 5! I also tried to code a website on cancer mutations & Fable 5 was immediately removed from my list! @AnthropicAI will probably soon ban me for such dangerous prompts! FYI @karpathy “little trigger happy Fable”
https://x.com/DeryaTR_/status/2064414826122866707
Was using Fable 5 to write inference code Anthropic flagged it as frontier AI research steering vector kicked in and it started importing ONNX 🤨
https://x.com/vikhyatk/status/2064515989795127744
Was using Fable 5 to write my world model training code. Anthropic flagged it as frontier AI research. The steering vector kicked in and it started implementing JEPA 🤨
https://x.com/MattVMacfarlane/status/2064440740483403829
We’ve also added refusal-fallback middleware to the Python, TypeScript, Go, Java, and C# SDKs for client-side retries. Middleware is useful for Claude API providers without support for server-side fallbacks. It detects the refusal, retries on Claude Opus 4.8, and keeps using it
https://x.com/ClaudeDevs/status/2064428351029449214
New blog post: on the million-x sample efficiency gap between AIs and humans, and whether it matters: “”The reason it is relatively easy for open source and previous laggards to catch up to within months of the frontier is that data is the real driver of progress. And data can
https://x.com/dwarkesh_sp/status/2064047458708308024
Palantir’s Karp says businesses are ‘unhappy’ with frontier AI labs
https://www.cnbc.com/2026/06/10/palantir-karp-enterprise-ai.html
Bro, Fable 5 won’t even answer “What does the heart do?” We’ve reached the point where a middle-school biology question can’t pass the safeguard.
https://x.com/Yuchenj_UW/status/2064524668208545955
AI access programs that allow third parties working on safety/security/resilience to access the most capable AIs (without their work being blocked by safeguards) seem very important going forward. Robustness to misuse should be doable with strong KYC and monitoring.
https://x.com/RyanPGreenblatt/status/2065182720133841069
I mostly agree with this, but it does seem like a bad and trust-damaging move to degrade performance on AI R&D tasks silently, rather than handling like other topics of concern (warning box + bumping the chat down to a less capable model)
https://x.com/hlntnr/status/2064733332882026565
I roughly agree. I think it’s reasonable/good for AI companies to block frontier AI R&D: AI R&D seems way riskier than nuclear or bioweapon R&D. But doing this with silent sandbagging is bad. Ideally, usage by actors that match in safety/security/governance would be allowed.
https://x.com/RyanPGreenblatt/status/2064948033423598035
NEW: malware developers added nuclear & biological weapons text to to their spyware. Goal? To trigger LLM safety refusals… so that their spyware wouldn’t be analyzed by an AI security scanner. Cleanest practical example I can think of for why over-indexing on first order
https://x.com/jsrailton/status/2064661778978533571
The Definitive Guide to AI Security: Market Map & Archetypes
https://zenity.io/resources/ebooks/the-definitive-guide-to-ai-security
DeepMind cofounder Shane Legg thinks that search is essential for a model to be genuinely creative. Pre-trained base models can do incredible things. But Shane thinks this is just a matter of them mixing together existing concepts from their training data. If he’s right, coming
https://x.com/dwarkesh_sp/status/2064827030273945676
imagine telling your customers there’s a small chance you’ll randomly decide they’re using your product wrong and you won’t tell them but will secretly silently sabotage their work
https://x.com/ericzelikman/status/2064442174373314701
EU Orders Meta To Stop Blocking Rival AI Chatbots On WhatsApp
https://www.engadget.com/2191213/eu-orders-meta-to-stop-blocking-rival-ai-chatbots-on-whatsapp/
Sovereign AI for all.
https://x.com/cohere/status/2064414912768618898
Project Glasswing: Securing critical software for the AI era | Hacker News
https://news.ycombinator.com/item?id=47679121
Human beings whose emotional centres are damaged, even if their intelligence is still intact, have terrible decision-making skills. Whatever role emotions are playing in humans, it’s necessary for agency. Ilya speculates that the equivalent for AIs is something to do with value
https://x.com/dwarkesh_sp/status/2062564414142972041
Imagine building a computer and not allowing its use in CS research. Thats some dystopian shit.
https://x.com/martin_casado/status/2064727048460058937
They didn’t mean pause AI research, they meant pause *your* AI research
https://x.com/bayeslord/status/2064437399292203401
A real problem with feeling the acceleration viscerally is that current models are really good and it is hard to feel the vibe difference on most individual tasks with new models, even as AIs continue to increase in ability by large amounts (which they actually are doing).
https://x.com/emollick/status/2062573152547279042
About 24 hours it seems. I’m glad to see them course correct on this as well. Model moderation and safeguards have always been a feature of deployed frontier models, but obfuscation without warning is a violations of the contract between the user and the provider.
https://x.com/code_star/status/2064931207310118940
AI companies say their models are getting better at finding software vulnerabilities. Is that bearing out in public data? Introducing our Cyber Vulnerabilities explorer, which visualizes Common Vulnerabilities and Exposures (CVE) reported to the CVE Program since 2022.
https://x.com/EpochAIResearch/status/2063027791638237487
ai gateway cost controls, now live. what i’m particularly excited about is what’s next: integrating with Cloudflare Access so you can see usage and set limits based on IdP resources (user, group, service, etc). unique to us because we have an entire ZT+ dev platform
https://x.com/michellechen/status/2062894017545720129
AI Gateway now supports spend limits. Stop runaway costs by setting budgets based on actual dollar usage per model or user.
https://x.com/CFchangelog/status/2062762883222483347
Bugbot is now over 3x faster, 22% cheaper, and finds 10% more bugs · Cursor
https://cursor.com/blog/bugbot-updates-june-2026
friends don’t let friends use vector search without guardrails
https://x.com/rishdotblog/status/2065026144903315545
I’m not even allowed to greet Fable.
https://x.com/OliviaHelenS/status/2064445405102784568?s=20
Labs starting to pull up the ladders on the ability to diffuse AI was inevitable. Doing it without telling the user is misaligned.
https://x.com/natolambert/status/2064404993193754830
OK first use of Mythos and it blocks any engineering, even the simplest. Why even release this lmao
https://x.com/Teknium/status/2064462936677203983
One reason you want AIs to be better writers is that there is a lot of writing even in software, and it is incredibly painful to hit a menu which is filled with Claudisms or ChatGPTish phrases. A report is not “”what leaves the room”” & analyses are not “”every number makes a mark””
https://x.com/emollick/status/2063368660798898284
Our statement on the UK government’s demand that all content on all devices sold or used in the country be scanned, on the presumption of nudity, using a dystopian combination of age verification and content scanning. This proposal will not safeguard children. It endangers us
https://x.com/signalapp/status/2064069692168519931
Releasing a model this capable comes with risks. Without safeguards, Fable 5’s capabilities in areas like cybersecurity could be misused to cause serious damage. Queries on a narrow range of topics will instead receive a response from our next-most-capable model, Opus 4.8.
https://x.com/claudeai/status/2064394155258765783
The best thing you could be spending time on this week is how you can avoid model lock-in. Strategize on how you can leverage different types of models. Not a choice IMO. There is no stronger data point than what we got this week on where things are headed. You want
https://x.com/omarsar0/status/2064753171214299209
the heart attack continues w/ Fable 5, we checked there were no reward hacks
https://x.com/karinanguyen/status/2065198770292146280
The Matrix idea of keeping humans as batteries is obviously weird… we would be more useful as dice. LLMs default to very similar kinds of arguments & structure, and even different LLMs seem to collapse to similar concepts. Humans provide a lot more variation in their own work.
https://x.com/emollick/status/2064109842390729123
This is very bad. Silent handicaps should not be a thing in a paid product
https://x.com/nrehiew_/status/2064400440264179923
VA launches clinical trial using hallucinogen to treat PTSD and alcohol addiction in veterans | Stars and Stripes
https://www.stripes.com/veterans/2026-06-02/ptsd-clinical-trial-mdma-va-21852986.html





Leave a Reply