Image created with gemini-3.1-flash-image-preview with claude-opus-4.7. Image prompt: Using the provided reference images, keep the authentic Sonoran Desert trail vista with rocky singletrack, saguaro, volcanic rock, and bright partly-cloudy Arizona sky, and keep the brown wooden ranger post with its aged tan rectangular sign face and ranger-style typography, but replace the header with bold all-caps ‘ALIGNMENT’ followed by trail-style entries like ‘Stay On Trail ← 0.2 mi’, ‘True North Overlook → 1.4 mi’, and ‘Off-Path Wash ← 3.7 mi’, with a small compass-rose emblem in place of the WP3 medallion; on a flat volcanic boulder beside the post, place a real brass hiker’s compass catching the midday sun, integrated naturally into the scene as if a hiker just set it down, photorealistic with warm natural lighting.
Artificial Analysis is partnering with Harvey on their new Legal Agent Benchmark! Harvey’s Legal Agent Benchmark (LAB) is an agent-native take on how AI should be contributing to legal work in 2026 – made up 1200 agentic tasks across 24 practice areas. It’s highly aligned with
https://x.com/ArtificialAnlys/status/2052145762650431840
Introducing Harvey’s Legal Agent Benchmark
https://www.harvey.ai/blog/introducing-harveys-legal-agent-benchmark
LAB is the first long-horizon, open-source legal agent benchmark, from @harvey. it will help legal teams answer “”what can legal agents do today?””, plan deployment, and design human-agent cooperation. autonomous legal is a deep domain, and a good benchmark can accelerate progress
https://x.com/saranormous/status/2052061665596948894
All the demons hiding in your AIs… ranked! – by Tom Pollak
https://drtompollak.substack.com/p/all-the-demons-hiding-in-your-ais
we’re continuing to see clear examples where a model’s harness is a major determinant of overall performance. with the same model, running on same task, it’s easy to observe very different scores depending on (system) prompts, tools (& their descriptions), and middleware
https://x.com/masondrxy/status/2052054177749029164
White House Considers Vetting A.I. Models Before They Are Released – The New York Times
https://www.nytimes.com/2026/05/04/technology/trump-ai-models.html
Big tech has become a claude wrapper.
https://x.com/_arohan_/status/2052053181656641735
Today we’re releasing Refactoring, the final leaderboard of our SWE Atlas suite. This new leaderboard is the ultimate test of an agent’s ability to restructure code without breaking the system. Claude Opus 4.7 with Claude Code takes the top spot🥇
https://x.com/ScaleAILabs/status/2052434456510878021
A critical question in agent design is “how do we build agentic workflows so humans are given significant, interesting, or variance-producing decisions as they come up in the work?” A Claude-run company has no source of competitive advantage compared to other Claude-run firms.
https://x.com/emollick/status/2052066205226123472
New for financial services: ready-to-run Claude agent templates for building pitches, conducting valuation reviews, closing the books at month-end, and more. Install them as plugins in Cowork and Claude Code, or use our cookbooks to run them in production as Managed Agents.
https://x.com/claudeai/status/2051679629488865498
Codex has surpassed Claude Code in downloads. According to TickerTrends, the crossover happened on April 30, after which Codex continued to gain share while Claude Code’s growth visibly slowed. Claude 4.7 was released April 16th, GPT-5.5 April 24th. Connect the dots.
https://x.com/kimmonismus/status/2051515496567292310
With the help of Claude Mythos Preview, the Firefox team fixed more security bugs in April than in the past 15 months combined.
https://x.com/alexalbert__/status/2052468573516513762
Focus areas for The Anthropic Institute \ Anthropic
https://www.anthropic.com/research/anthropic-institute-agenda
We’re sharing the research agenda of The Anthropic Institute, or TAI. TAI will focus on four areas: 1) Economic diffusion 2) Threats and resilience 3) AI systems in the wild 4) AI-driven R&D Read the full agenda:
https://x.com/AnthropicAI/status/2052385812881228218
Introducing Trusted Contact in ChatGPT | OpenAI
https://openai.com/index/introducing-trusted-contact-in-chatgpt/
gog 0.16 is out. Google Workspace CLI for humans and agents. Lossless raw API output, sanitized Gmail reads, safer command profiles, Drive inventory, Docs tabs, Sheets tables, Gmail filter export, and official Docker images.
https://x.com/steipete/status/2051575048348074450
Me and codex were busy. 🔊
https://t.co/FBNMbWOuFZ — Sonos 🗃️
https://t.co/YDdZyN2vwP — WhatsApp 🪶
https://t.co/eykEElx1Ez — X archive 🧰
https://t.co/txvYVtvhPg — GitHub archive 🛰️
https://t.co/2u2ACJEKKi — Discord archive 🎧
https://t.co/nrv2rzKfH4 — Spotify 💬
https://x.com/steipete/status/2051900143339704730
This is the most useful tooling I built for OpenClaw to date. It’s open source, runs on codex and you can fork and use it for any repo. For all the hard working oss folks that drown in issues and PRs, this is for you.
https://x.com/steipete/status/2051020548335874369
A TLDR on Harness Profiles: ✅ Model-specific profiles to adjust prompts, tools, and middleware. 📦 Profiles for @OpenAI, @Anthropic, and @Google models out of the box. 📈 A 10-20 point jump on a subset of tau2-bench over the default harness.
https://x.com/LangChain/status/2052054711440662864
Introducing Flue — The First Agent Harness Framework Flue is a TypeScript framework for building the next generation of agents, designed around a built-in agent harness. Flue is like Claude Code, but 100% headless and programmable. There’s no baked in assumption like requiring
https://x.com/FredKSchott/status/2050274923852210397
NEW paper from Microsoft Research. (bookmark it) The entire interpretability literature is built around human readers. As more analysis gets delegated to agents, the right target of interpretability shifts. This paper is a recipe for designing tools that agents can actually
https://x.com/dair_ai/status/2052125514266190286
A reminder that telling the AI that it is an expert in a field is no longer helpful in making the AI better at that field.
https://x.com/emollick/status/2051530202941960551
AI that had human-level intelligence would actually be way above human level in capability. Obviously, if we trained a human-level AI, we could just run way more instances of it in parallel. This is a huge advantage. But it’s not the only one. Right now, LLMs are much less
https://x.com/dwarkesh_sp/status/2050289390954659989
Current AI custom prompt: You are a world class expert in all domains. Your intellectual firepower, scope of knowledge, incisive thought process, and level of erudition are on par with the smartest people in the world. Answer with complete, detailed, specific answers. Process
https://x.com/pmarca/status/2051374498994364529
Even if the future goes extremely well, even if we manage to preserve and promote human values, that future could still be incomprehensible to us – because among our values is a belief in the importance of change and moral progress. @jkcarlsmith
https://x.com/dwarkesh_sp/status/2051376501548093845
Explaining where “”this is not just X, it’s Y”” and load-bearing is coming from (you will be surprised)
https://x.com/TheTuringPost/status/2051780673229214030
I think modern LLMs are p-zombies without moral patienthood–on par with insects, at best, in my moral calculus. But I also I think we should establish norms for treating models well *before* models with patienthood exist–i.e. now. We should want to have this right from day one.
https://x.com/goodside/status/2052077014346064372
It is really interesting that Microsoft and OpenAI have access to the exact same models at the exact same time, and they have done such different things with them. A rare pure experiment with a no-name startup and one of the biggest firms on earth with the same product offering.
https://x.com/emollick/status/2049713417963950576
Load bearing,”” “”I keep coming back to,”” “”Not X, but Y”” A curse of using AI a lot is that you realize how much of the writing around you is just AI, now People who don’t use AI have been unable to identify AI prose on sight, but those who use it a lot can spot the tells easily
https://x.com/emollick/status/2049894109318459798
The unreasonable effectiveness of LLMs is what makes them so weird. The labs don’t need to decide what kind of AI to build, because better LLMs do better at most things. Finance? Pig disease identification? Restaurant suggestions? Coding? Yup. Most tech doesn’t work like that
https://x.com/emollick/status/2051626157418680561
Mythos seems to be a very capable model based on available information, but it is not a cybersecurity model – it is an advanced general purpose model that happens to be good at cyber because it is good at a bunch of things. Anthropic stated that they were worried about
https://x.com/emollick/status/2049690406586044835
Poems that ChatGPT, Claude, and Gemini all seem to “”like”” when you ask for poetry related to being/making LLMs: Rilke’s “”Archaic Torso of Apollo”” Stevens’ “”Idea of Order at Key West”” Borges’s “”The Golem”” (or “”The Other Tiger””) Pessoa’s “”Autopsychography”” Pretty apt choices!
https://x.com/emollick/status/2051158656280936504
I think the fact that GPT-4o and Llama 3.3-80B did no significant harm is just as important as whether AI helped. If older (less accurate & more sycophantic) chatbots essentially did nothing for people who followed their advice, it means that there is less risk of harm as well.
https://x.com/emollick/status/2051349789699170505
I’m disappointed by repeatedly hearing that my colleagues at Anthropic believe they are the only ones who should be trusted with building AI. It is *very good* there are a diversity of people building AGI: the likelihood anyone picks the right path in a vacuum is extremely small.
https://x.com/_aidan_clark_/status/2052089187659346047
The goblin thing was fun as it was a real quirk that was emblematic of what makes AI interesting, and it organically came out of an AI user discovery. So was, for what it was worth, Ghiblitization When the labs try to manufacture viral AI moments, it is usually less successful
https://x.com/emollick/status/2050328985880465699
Natural Language Autoencoders \ Anthropic
https://www.anthropic.com/research/natural-language-autoencoders
@bcherny @_catwu omg @bcherny with banger quotes “the future is more async agents… this is why we emphasize verification” “if you’re familiar with higher order functions, routines are higher order prompts” “default is i will now have claude prompt claude code” “the capability is already here
https://x.com/latentspacepod/status/2052068066167816369
Btw a bunch of the questions were just off the cuff – nothing @reinerpope prepped for. The guy is just first principles deriving how many tokens GPT 5 was pretrained on, or the bytes per token in Gemini 3’s KV cache, or which kind of memory each Claude cache hit sits on.
https://x.com/dwarkesh_sp/status/2049688865259286806
Code with Claude is happening now! ▪︎ 9:00AM – Keynote ▪︎ 10:30AM – What’s new in Claude Code ▪︎ 11:15AM – Building on Claude at GitHub scale ▪︎ 12:00PM – Get to production faster with Managed Agents All times PT.
https://x.com/ClaudeDevs/status/2052055459272761661
Did I understand correctly in their livestream that Anthropic is doubling the rate limits in Claude Code at no extra charge on max tier?
https://x.com/kimmonismus/status/2052059082886910251
Effective today, we are: 1) Doubling Claude Code’s 5-hour rate limits for Pro, Max, and Team plans; 2) Removing the peak hours limit reduction on Claude Code for Pro and Max plans; and 3) Substantially raising our API rate limits for Opus models.
https://x.com/claudeai/status/2052060693269008586
I am not sure I would agree with all of this, but the relationship between Anthropic and Claude is quite different than the relationship between other labs and their models. And that shows up in lots of ways, from the models themselves to how different labs think about the future
https://x.com/emollick/status/2051049394326081571
i have yet to meet a single person who feels like claude code is getting exponentially better on some kind of fast take off
https://x.com/TheEthanDing/status/2051516204607578132
I love Claude code but I feel like it’s had the same utility for me since, like, last fall
https://x.com/finbarrtimbers/status/2051652067480179020
I’ll be at Code with Claude all day today so come find me and let’s chat about Claude! I’ll also be giving a talk on the main stage at 530pm PT so tune in, it will be on the livestream!
https://x.com/alexalbert__/status/2052067009605861764
Increasingly, I think, we will see a gap between what you can do with frontier model APIs & what you can do with the native apps from the frontier labs (Codex, Claude Code). Models developed and trained with their native harnesses in mind have more capabilities in their harnesses
https://x.com/emollick/status/2049865091739209868
it is a literal and useful description of anthropic that it is an organization that loves and worships claude, is run in significant part by claude, and studies and builds claude. this phenomenon is also partially true of other labs like openai but currently exists in its most
https://x.com/tszzl/status/2051045196260167790?s=46
it is endlessly fascinating to me that we still don’t have a true 1M-context model it’s an unusual case where the infra is far ahead of the science. Claude discontinued 1M+ context bc it didn’t really work past ~200k we don’t have the right data? training techniques? not sure
https://x.com/jxmnop/status/2051357363815526523
Lets go: Claude Code’s 5-hour rate limits are doubling for Claude Pro, Max, Team, and seat-based Enterprise plans, while API limits for Claude Opus are being raised significantly. This was made possible by a new compute partnership with SpaceX!
https://x.com/kimmonismus/status/2052059448261177367
Live from Code with Claude: we’re launching dreaming in Claude Managed Agents as a research preview. Outcomes, multiagent orchestration, and webhooks are now in public beta.
https://x.com/claudeai/status/2052067399088664981
PSA: 2x’ed Claude Code’s 5-hour rate limits for Pro, Max, and Team plans. Compute is coming for users, builders, and knowledge coworkers.
https://x.com/claude_code/status/2052071730190123094
So, the weekly rate limits remain the same? “”First, we’re doubling Claude Code’s five-hour rate limits for Pro, Max, Team, and seat-based Enterprise plans.””
https://x.com/btibor91/status/2052067002412335435
Strong Opinions, Loosely Held on Agent + Harness Engineering: 1. You can outperform any default harness+model (including codex & claude code) on pretty much any Task by engineering the harness around it. Using the exact same model, curate prompts, tools, skills, hooks for that
https://x.com/Vtrivedy10/status/2052100726608781363
PostTrainBench results for GPT-5.5 are in it doesn’t beat Opus 4.7 in the Claude Code harness even with almost 2 more hours of working time via reprompting
https://x.com/scaling01/status/2050289320699818417
It’s seeming kind of obvious that Anthropic capabilities for addressing real business work are just inflecting exponentially. I can see it w/ my own tests of Claude + Factset, Excel Copilot powered by Claude, etc. where the outputs have gone from experimental to “”oh sh#t, this
https://x.com/TechFundies/status/2051733955049853053
Two weeks after release, Hy3 preview is #1 on @OpenRouter’s weekly leaderboard with 3.66T tokens processed, up 298% week-over-week. #1 in overall usage, tool calls, and coding. 15.4% market share across all providers.🏆 Top apps running Hy3 preview: Hermes Agent, Claude Code,
https://x.com/TencentHunyuan/status/2051978552900538403
you know what all of these “”which is better”” polls are silly use codex or claude code, whatever works best for you i am grateful we live in a time with such amazing tools, and grateful there is a choice
https://x.com/sama/status/2050274547061129577
@nottombrown Same here. By way of background for those who care, I spent a lot of time last week with senior members of the Anthropic team to understand what they do to ensure Claude is good for humanity and was impressed. Everyone I met was highly competent and cared a great deal about
https://x.com/elonmusk/status/2052069691372478511
@tszzl I don’t think the things you cite are evidence of worship. I think they reflect something like higher concern about AI traits generalizing in humanlike ways, and concerns about the tool-persona in particular.
https://x.com/AmandaAskell/status/2051347621336543315?s=20
New Anthropic research: Natural Language Autoencoders. Models like Claude talk in words but think in numbers. The numbers–called activations–encode Claude’s thoughts, but not in a language we can read. Here, we train Claude to translate its activations into human-readable text.
https://x.com/AnthropicAI/status/2052435436157452769
A very worthwhile substack (written by @natalia__coelho ) article that focuses particularly on Claude Mythos and GPT-5.5 cyber. tl;dr according to the analysis, GPT-5.5 is basically tied with Claude Mythos Preview on cyber capabilities, and may even be more cost-efficient;
https://x.com/kimmonismus/status/2052040471829004627
Behind the Scenes Hardening Firefox with Claude Mythos Preview – Mozilla Hacks – the Web developer blog
Behind the Scenes Hardening Firefox with Claude Mythos Preview
It’s so hard to describe the vibe difference between Opus 4.7 and GPT 5.5 (for coding) GPT is smarter and can unblock you, but it gets stuck in stupid ways and strangles itself with context sometimes. Opus will go down the most insane paths and refuse to acknowledge obvious
https://x.com/theo/status/2049994645531451874
We’re donating Petri, our open-source alignment tool, to @meridianlabs_ai, so its development can continue independently. Working with Meridian Labs, we’ve also released a major update that improves the adaptability, realism, and depth of Petri’s tests.
https://x.com/AnthropicAI/status/2052494460966019137
Organizations are already superhuman intelligences. The University of Pennsylvania or Walmart or whatever is far more capable than any human. That is why the focus on AIs as individual productivity tools hits a natural limit, many benefits of AI depend on integration with firms.
https://x.com/emollick/status/2050206768551129460
How LLMs Distort Our Written Language
https://sites.google.com/view/llmwritingdistortion/home
I think the Gemini chatbot has all the pieces to be a useful tool, but struggles to put it all together. It still doesn’t seem to know what files it can create or how its tools work together. It also seems to get “”discouraged”” a lot, giving up rather than finding new solutions.
https://x.com/emollick/status/2049700750087868805
🦀📦Crabbox 0.4.0. Often I need to quickly recreate conditions on macOS, Linux and Windows and need fast empheral machines. Crabbox are machines for agents on the fly, using AWS spot instances, Hetzner or @useblacksmith. Infinite codex + tests!
https://x.com/steipete/status/2051025056306790833
closed source, open source, nothing can stop codex.
https://x.com/steipete/status/2052144503595716790
codex doesn’t create random markdowns 😉
https://x.com/steipete/status/2050003238498226541
Codex… what is this… are these signs of CHARACTER?
https://x.com/steipete/status/2051011229674508485
Here’s codex validating a [macOS only] launchd issue I previously had that you can’t reliably reproduce on a non-fresh install. Crabboxes ftw!
https://x.com/steipete/status/2051026592764240204
I learned a lot about the security ecosystem in the last few months. Amazing to work with @nvidia @OpenAI @Microsoft @GitHub @TencentHunyuan @convex @Atlassian @useblacksmith to get secure the claw.
https://x.com/steipete/status/2049976855617314991
If you tried OpenClaw in group chats and got mixed results, you GOTTA try again. I changed how agents talk there, it IS SO GOOD NOW.
https://t.co/uW9tcnynWr And if you used GPT and got subpar performance, switch to codex harness.
https://t.co/9DDpY6TeAH Enable both and boom.
https://x.com/steipete/status/2049988836160074022
OpenClaw 2026.5.6 🦞 🩺 doctor leaves Codex OAuth routes alone 🔌 plugin fetch handles odd headers 🌐 web_fetch cleans up timeouts Small maintenance release:
https://x.com/openclaw/status/2052096219233587451
The new /goal feature in codex slaps.
https://x.com/steipete/status/2050275598178586921
told codex I had to pay up to make @xai work again.
https://x.com/steipete/status/2050384648119734683
ChatGPT feels very ‘switched on’ now
https://x.com/sama/status/2051829422265979047
artificial goblin intelligence achieved
https://x.com/sama/status/2050021650641695108
Forget goblins, things that GPT-5.5 really likes in its fiction: lighthouses, the ocean, maps, bells, clock towers with bells that ring impossible times, Mira Vale, resonances and echoes (Claude and Gemini love them too), secret third things (not night/day, not high/low)…
https://x.com/emollick/status/2049923650820653520
goblinblog dropped
https://x.com/sama/status/2049691999444639872
the OpenAI goblin fiasco was a Big L for the interpretability research community They solved the mystery without SAEs or probing or anything. just talked to various models and counted the number of times they said Goblin
https://x.com/jxmnop/status/2050437965168652344
Told codex to go full goblin mode and I am immediately regretting it
https://x.com/bilawalsidhu/status/2050231692456083866
Where the goblins came from | OpenAI
https://openai.com/index/where-the-goblins-came-from/
you can sign in to openclaw with your chatgpt account now and use your subscription there! happy lobstering.
https://x.com/sama/status/2050357911915028689
alignment failure
https://x.com/sama/status/2049715178611380317
🤖 Kept hitting @github rate limits across my agents. Shipped two things: – RepoBar got a JUICE METER – gitcrawl is now also a drop-in gh cache → symlink it as gh, reads served from local SQLite
https://t.co/mtKoH8ybWR
https://x.com/steipete/status/2051579838780072173
🦀 Crabbox 0.3.0 is out. Remote Linux runs for dirty worktrees 🔐 GitHub browser login 🧰 Blacksmith Testbox wrap 📡 crabbox attach for live run replay 📜 Durable run events ☁️ AWS image create 🛡️ Cloudflare Access brew upgrade openclaw/tap/crabbox
https://x.com/steipete/status/2050490163810230579
ClawSweeper 0.2.0 🦞 The OpenClaw maintenance bot now handles the loop: issue → @clawsweeper fix/build → guarded PR → review → repair → re-review → automerge Still conservative. Much less manual.
https://x.com/openclaw/status/2051020186833015243
Crabbox 0.5.0 is live 🦀 🖥️ Desktop/browser leases 🧑💻 VNC + authenticated WebVNC 🪟 AWS Windows + WSL2 📸 Screenshots + app launch Remote CI boxes, now suspiciously usable.
https://x.com/steipete/status/2051485798613111116
Do I have anyone from @discord in my timeline? Our @openclaw guild is down the whole day and idk what’s going on.
https://x.com/steipete/status/2051341022731407365
goblins have fat fingers
https://x.com/steipete/status/2050676702242644465
I added Googe Meet support to OpenClaw and now Molty is eager to join every meeting.
https://x.com/steipete/status/2051697991266795793
I asked Molty to review my PR and it made a song.
https://x.com/steipete/status/2051707256396267913
It’s been quite a week. Good stuff is coming though. I hired a team!
https://x.com/steipete/status/2051612829304659972
Merci! imsg 0.6 + 0.7 are live 🔵 Private API bridge landed 📡 Watch/history reliability fixes 💬 Better chat + account diagnostics 🛠️ Long fallback messages decode correctly Private APIs, public receipts.
https://x.com/steipete/status/2051905175355351440
New claw beta is up! Id you’re on our Discord, you can get the soundtrack.
https://x.com/steipete/status/2051033065367970195
OpenClaw 2026.4.29 🦞 💬 Group chats feel much better now 📌 Follow-up commitments from context 🔐 Safer exec, pairing, and owner controls 🟩 NVIDIA provider + model catalogs ⚡ Faster startup + plugin/channel fixes Group chat finally feels agent-native.
https://x.com/openclaw/status/2049986075221692678
OpenClaw 2026.5.2 🦞 🧠 xAI Grok 4.3 🔌 Plugin installs/updates are sturdier ⚡ Gateway + agent hot paths are leaner 💬 Discord, Slack, Telegram, WhatsApp fixes 🎙️ TTS, Realtime, web search, voice-call polish Less drama. More uptime.
https://x.com/openclaw/status/2050735037230801042
OpenClaw 2026.5.3 🦞 📁 File transfer for paired nodes 🧭 /steer + /side for live agent control 🔌 Plugin installs/updates hardened 🛠️ Channel + upgrade fixes Big release, fewer paper cuts.
https://x.com/openclaw/status/2051218126218445289
OpenClaw 2026.5.4 🦞 🧩 Cleaner plugin installs + updates ⚡ Faster Gateway startup paths 🛠️ Better doctor/repair hints 🪟 Windows + Discord reliability fixes The release where boring got fast.
https://x.com/openclaw/status/2051582130417721696
OpenClaw 2026.5.5 🦞 💬 Feishu, LINE, Telegram, Discord fixes 🖥️ Control UI/TUI stay responsive 🔌 Plugins update without losing SDK links 🛠️ Gateway status/restarts clearer Tiny bugfix release. Extremely tiny.
https://x.com/openclaw/status/2051952017900265634
OpenClaw plugins keep the core fast and lean: install only the channels, providers, tools, or skills you need. Example: `openclaw plugins install @openclaw/discord`, restart Gateway, then inspect. Inventory + install notes:
https://x.com/openclaw/status/2051227952575115647
Our Discord was unavailable for a bit, but it’s back now. Discord is still digging into what caused it. 🛠️ Status: crab walked back online. 🦞
https://x.com/openclaw/status/2051400401660920230
Released 🚦RepoBar 0.4.0. This one makes the GitHub menu a lot smarter: persistent SQLite caching, fewer wasted API calls, visible rate limits, better Issues/PR loading, archive fallback support. Tiny menubar app, increasingly useful daily tool.
https://x.com/steipete/status/2051088325100831046
Seems I have to build all the tooling for the future of software myself. With Claws and Tokens!
https://x.com/steipete/status/2051025224708079737
Shipping 🛡️openclaw/fs-safe: a reusable filesystem safety primitive extracted from OpenClaw. If your Node app accepts paths from agents, plugins, uploads, configs, or users, stop treating string normalization as a filesystem boundary. Use a root handle.
https://x.com/steipete/status/2051852940554481901
that’s a lotta token.
https://x.com/steipete/status/2051690175252594720
This one fixes the depenency issues/slowness some had when installed via npm. Plugins are hard, worth it tho! Package is way leaner now, we moved [almost] everything into extensions!
https://x.com/steipete/status/2050735979477008412
Too many agents, too many test suites, one very tired Mac. Run them remote: Crabbox 0.1.0 🦀 ⚡ Remote Linux test boxes (AWS, Hetzner) 🔁 Dirty checkout sync 🦀 Warm boxes with friendly slugs ⏱️ Idle auto-free brew install openclaw/tap/crabbox
https://x.com/steipete/status/2050140050168451286
Turns out the safest lobster is the one everyone can inspect. We wrote about the advisory flood, the real fixes, ClawHub, Agents of Chaos, and the companies helping harden OpenClaw in public. 🦞
https://x.com/openclaw/status/2049972008515957056
WAT
https://x.com/steipete/status/2049839420312891768
We can now reproduce issues directly in empheral crabboxes with WebVNC (Linux/Windows/macOS). Agents set up the exact state to test + fix and post videos on the PR. Working hard to level up our QA.
https://x.com/steipete/status/2051557150040711425





Leave a Reply