“We’re open-sourcing our benchmark so that others can build on our work. Making systems perfectly robust might not be possible. Rapid response is a more tractable option. Paper:
The case for targeted regulation \ Anthropic
“Anthropic released Claude 3.5 Haiku on their API, Amazon Bedrock and Google Cloud’s Vertex AI The new model outperforms GPT-4o on SWE-bench Verified and even surpasses Claude 3 Opus on many benchmarks But comes with increased pricing
“guys who use chatgpt over claude might as well have green text msgs too” / X
x.com/AnthropicAI/status/1857108263042502701
Dario Amodei: Anthropic CEO on Claude, AGI & the Future of AI & Humanity | Lex Fridman Podcast #452 – YouTube
Anthropic hires its first “AI welfare” researcher – Ars Technica
Improve your prompts in the developer console \ Anthropic
“New research: Jailbreak Rapid Response. Ensuring perfect jailbreak robustness is hard. We propose an alternative: adaptive techniques that rapidly block new classes of jailbreak as they’re detected. Read our paper with @MATSprogram:
“🤯 Mind-blown! Just built a complete flashcard web app in less than 30 seconds using @Qwen’s new Coder demo! Like Claude’s artifacts but open source. One prompt = full web app with cards flipping. Try it:
“Because of its compatibility with OpenAI and Anthropic APIs, here is how easy it is to create your own Grok Engineer with @xai. Watch Grok generate code and folders in one shot. I updated the repo so you can try it out! https://x.com/skirano/status/1855727722196324424





Leave a Reply