Image created with gemini-3.1-flash-image-preview with claude-opus-4.7. Image prompt: A Byzantine gold-ground mosaic apse featuring a single centered iconic six-petal aperture-flower of hammered gold with a deep imperial-purple pupil, encircled by a halo of fine silver circuit-rings and a slow gyre of tiny clockwork songbirds, tesserae texture with visible grout and warm candlelit glow, the words OPENAI set in bold ivory Trajan Roman capitals across the lower third, opulent burnished gold and Tyrian crimson palette, painterly tactile render, 16:9 full-bleed.
Another one: today we released Remote Computer Use in Codex! This means you can use all the apps on your Mac from Codex Mobile, even when your computer is at home and locked. It’s kinda magic.
https://x.com/AriX/status/2057645366640828660
so Codex on iPad acts like a Codex mobile phone, which gives you the full desktop UI/UX. meaning, you can use your iPad to control your mac mini at home and have full screen portable development, it’s really magical.
https://x.com/kevinrose/status/2059297989039128700
Economic Futures in the Age of AI
https://openaifoundation.org/news/economic-futures-in-the-age-of-ai
very belated but in retrospect i think @sama’s mythical “”build a business that gets better when models get better”” is basically what I called Agent Labs here. seeing a very direct correlation with model performance and agent lab revenue, discontinuity in Q4 2025 (clip from
https://x.com/swyx/status/2057119153337545096
For complicated agent work, it’s amazing how much GPT5.5 has improved. I found 5.2 to be very far behind Opus. Now using Opus 4.7 after 5.5 feels like a big step backwards. Gotta love this level of competion! Strong comeback for OpenAI.
https://x.com/dhh/status/2057906669158309913
Huge credit to the OAI team for solving the unit distance problem with 5.5 – it is now my go to example that models can in fact pull together disparate ideas into new discoveries. As with all 4 minute miles, we had to try and cross it too! Turns out mythos solves it with a cute,
https://x.com/_sholtodouglas/status/2059303540150137244
37 minutes to eat my entire token budget with a task that used <1% on Codex, and the output so far is somehow worse than Codex. Maybe I do not like ultracode.
https://x.com/cremieuxrecueil/status/2060161310302630154
At @ThriveHoldings, we built a product with @OpenAI to automate tax prep for the 30+ accounting firms we own across the country. This season, it processed 7k+ returns. But what I think is more interesting is that the product meaningfully self-improved as accountants used it.
https://x.com/samaysham/status/2059649942910562378
Codex & Pool
https://x.com/Dimillian/status/2058181477943083193
Codex just launched one of the coolest features – Appshots. by pressing both CMD keyboard buttons, context of whatever app you’re on gets sent to Codex, including screenshots and text. you might be thinking “”isn’t this the same as screenshotting and pasting it””, and no it
https://x.com/kr0der/status/2057728283010306227
codex thursday no. 6: – appshots – /goal – remote computer use while locked – advanced annotation mode – plugin sharing – improved analytics – small stuff everywhere
https://x.com/ajambrosino/status/2057716220963803577
codex… made a smiley? 🙂
https://x.com/steipete/status/2058332234247987379
Gone shipping 🚢
https://x.com/ChatGPTapp/status/2057178633249394819
GPT 5.5 found a 27-year-old RCE introduced in April of 1999. I’ve triple-checked the flow and commit history, it’s real. Can’t wait to responsibly disclose!
https://x.com/PhiloGroves/status/2059661579466006608
GPT-5.5 in Codex helps @databricks parse complex customer documents more reliably.
https://x.com/OpenAIDevs/status/2059353117934899289
GPT-5.5 Pro is a very solid fact checker. I can throw entire chapters at it and it will hunt down every key reference accurately. The only real annoyance is that it loves nuance, so returns a lot of “the general idea is right, but you are not taking into account tiny detail X”
https://x.com/emollick/status/2058331615525232988
grok-imagine-image-quality lands at #5 on both the Artificial Analysis Text to Image and Image Editing leaderboards, the leading model outside of OpenAI and Google and at a much lower price! grok-imagine-image-quality is @xAI’s latest image model and a higher quality variant of
https://x.com/ArtificialAnlys/status/2060073523528581249
I built an autotriage skill for codex that has a set of guidelines + reads VISION.md from my repos, so issues/prs that have a clear way of – fit vision of the project – being inferrable in code with high confidence – clear fix – can be live tested Are now worked on autonomously.
https://x.com/steipete/status/2058240758801420530
I found this Wired article on AI fact-checking frustrating. It could have been about why we continue to need human fact checkers (talk to people, use judgement, resolve conflict). Instead it is full of old info & stuff about free models GPT-5.5 Pro checked it (& I checked GPT)
https://x.com/emollick/status/2059277092315869650
i had codex audit my entire macbook to see how much space we can save and it’s found 500 GB to save, AWESOME prompt was: “”do a FULL read only analysis on my Macbook to help me optimize storage”” note: why tf is there a codex-tui.log file that is 116gb ??????? WHAT ????
https://x.com/KingBootoshi/status/2058943762207006959
I redid the multi-digit multiplication experiment, now with gpt-5.5. With medium reasoning and 7 samples each cell, it pretty much aced the test with 99.46% accuracy. The model had no tools to call and had to rely on its reasoning. Can it go further? (1/4)
https://x.com/cozyblazex/status/2057739317649588558
I’m late to the party, but cmux is great.
https://t.co/8uuStvqwcm current split: codex mac app: knowledege work, learning, reading cmux + codex cli: coding
https://x.com/steipete/status/2058093406874689770
If you know one thing about every right now, it’s that we’re heavily Codex pilled. So we wrote a guide on how to use Codex for knowledge work as well as we do. You dont want to miss this one…
https://x.com/bran_don_gell/status/2059385209401798684
It took me like 2 months, but I’ve grown to love gpt-5.5. You have to prompt entirely different and put some time into your agents[.]md. Now that I’m over the hump, I can’t really use any other model for code.
https://x.com/theo/status/2059372156753219938
It’s Codex Thursday, and yes, we have updates for you. First up: Appshots, a new way to bring the context of what you’re working on into Codex. On your Mac, press Command-Command to attach your app window to a Codex thread. Codex gets both a screenshot and text from the window,
https://x.com/OpenAIDevs/status/2057530207976989179
Modded-NanoGPT optimization result #15: A Newton-Muon based setup has achieved a step count which is below Muon and slightly above NorMuon. This result was submitted by Zhehang Du, one of the Newton-Muon authors.
https://x.com/kellerjordan0/status/2059353883881976044
OpenAI is offering $2M in tokens to every YC company in the spring and summer batches. We extended the summer deadline to May 25 so more founders can get in on it.
https://x.com/ycombinator/status/2057555656673210639
Over the last 4 days I’ve probably spent 3+ hours trying to work around weird bugs and limitations of the remote features in the Codex app. I’m trying to give them all the feedback they need to fix it. Did not expect the experience to be 10x smoother in our open source app 🙃
https://x.com/theo/status/2057961165175873930
Over the weekend, I asked Codex to analyze my Slack message history and recommend a better way to organize my growing number of channels. Then I had Codex reorganize and categorize my Slack sidebar with computer use while I worked on something else. I now have an automation for
https://x.com/derrickcchoi/status/2059277053925478714
realisation: I haven’t opened an IDE in more than a month, diff view + file viewer in the codex app is more than sufficient one app to rule them all!
https://x.com/reach_vb/status/2057830243201622368
Seems GPT-5.2 reaches expert level in peer review: 45 scientists took 469 hours evaluating human & AI reviews on 82 papers. “”Surprisingly, current AI reviewers are competitive even with the top-rated reviewers in Nature’s official peer review…”” though not without weaknesses.
https://x.com/emollick/status/2057528309727088907
still getting token-mogged by GPT-5.5
https://x.com/scaling01/status/2060080401947746483
Still limited by compute, so I built a thing that runs codex in the cloud, powered by @Cloudflare firecracker boxes (and since that’s not beefy enough for larger projects, tests are run via crabbox) Uses Ghostty ofc, via WebAssembly. Codex replicated itself, basically.
https://x.com/steipete/status/2058248513662697622
the model alone is no longer the product
https://x.com/gdb/status/2057670776803996110
The wild part of Codex sub-agents isn’t that one AI can use Chrome. It’s watching a single prompt turn into seven browser sessions running at the same time. Flights, cars, Airbnbs, hikes, forms, checkout pages. Still rough around the edges but still feels like the future
https://x.com/georgepickett/status/2059709672765173950
Thread by @ChatGPTapp on Thread Reader App – Thread Reader App
https://threadreaderapp.com/thread/2057908052968521902.html
To simplify our Codex compute fleet management, we will be sunsetting GPT-5.2 and GPT-5.3-Codex in Codex on June 2nd when logged in with your ChatGPT account. For free plans, GPT-5.5 will be the default frontier model to build and work with going forward. These models will
https://x.com/thsottiaux/status/2059650685948551384
To simplify our Codex compute fleet management, we will be sunsetting GPT-5.2 and GPT-5.3-Codex in Codex on June 2nd when logged in with your ChatGPT account. For free plans, GPT-5.5 will be the default frontier model to build and work with going forward. These models will
https://x.com/thsottiaux/status/2059650685948551384?s=20
try Appshots in the Codex app:
https://x.com/gdb/status/2057802037757157838
trying to remember what it was like to code before codex
https://x.com/gdb/status/2057704270531903811
UPDATE: Came up with an even better version of this prompt after the feedback Ask Codex to look across your sessions, Memories, and Chronicle, identify patterns, reuse what already exists, and only create the smallest useful skill, subagent, or automation. “”Look back over my
https://x.com/reach_vb/status/2058538305872949490
We’ve expanded the Admin API to help enterprises manage OpenAI projects programmatically. New support includes spend alerts, model allowlists, data retention controls, hosted tool controls, and more granular cost visibility for capabilities like file search and web search. 🔗
https://x.com/OpenAIDevs/status/2059703665276145920
Workload Identity Federation brings cloud-based identity to the OpenAI API platform. Teams can manage access through IAM workflows while reducing the need to distribute permanent API keys across services. 🔗
https://x.com/OpenAIDevs/status/2059703600662925635
ぼくの着想の限界=Codexの限界。 それくらいまーじでCodexでなんでもできる。 これアリエクで買ったやっすいMP3プレイヤー。 でもBluetoothの音飛びと操作性が悪くて放置してたんですよ。 だけど昨日急にシャワーしている時にエウレカして、
https://x.com/bunkaich/status/2059178996126900703
Almost everyone is building agent harness systems the wrong way. The default move: pick LangChain or LangGraph or the OpenAI Agents SDK, accept the loop, the tools, the memory, the orchestration, the policy engine, the credential store, the budget tracker, all of it, as one
https://x.com/ghumare64/status/2060072412868235587
Cloudsail: Instant Sandboxes for Coding Agents Create a new Cloudflare Sandbox for each task with a shell, Codex and GitHub access. Tokens are never exposed to the sandbox. Update your deps far away from your laptop. npm install -g cloudsail cs
https://x.com/cnakazawa/status/2057823910574588238
Codex computer use entirely driving iphone simulator to bug bash a feature it just built
https://x.com/JustinBleuel/status/2058228412158758950
Macro Evals for Agentic Systems
https://developers.openai.com/cookbook/examples/partners/macro_evals_for_agentic_systems/macro_evals_for_agentic_systems
Private MCP servers 🤝 OpenAI products Your team can keep MCP servers inside your network while ChatGPT, Codex, and the Responses API connect through outbound-only HTTPS. 🔗
https://x.com/OpenAIDevs/status/2059703536825565499
Secure MCP Tunnel | OpenAI API
https://developers.openai.com/api/docs/guides/secure-mcp-tunnels
OpenAI kicked off the AI compute buildout in 2023. But today it uses ~10% of the world’s compute, and the top labs together are probably under half. In this week’s newsletter, @justjoshinyou13 discusses how much that share may change, and when it could hit a ceiling. 🧵
https://x.com/EpochAIResearch/status/2057499893854536185
⚙️ Behind the build of self-improving tax agents with Codex We co-built Tax AI with @ThriveHoldings around tax prep workflows so when reviewers fix any errors, Codex can trace the failure, improve the system, and test the change before it ships.
https://x.com/OpenAIDevs/status/2059638868983562640
You can now transcribe meetings in real time using Codex and ask Codex questions about meetings as they’re happening! I updated my new Codex Meeting Recorder skill to use GPT Realtime Whisper. Tell Codex to use the skill, and it will start transcription and show it in the
https://x.com/_simonsmith/status/2059626873479422250
Exciting news, MAI-Image-2.5 (Preview) from @MicrosoftAI debuts at #3 in the Text-to-Image Arena with a score of 1,254 — a +72 point improvement over MAI-Image-2. A top 5 arena previously held only by @GoogleDeepMind and @OpenAI has a new lab in the mix. Congrats to the
https://x.com/arena/status/2059346024632820146
We’re adopting the Linux Foundation’s OpenMDW framework across our open model families. This helps make open model licensing simpler and more consistent at scale. A single legal framework across models, code, documentation, and data helps reduce friction for developers and
https://x.com/NVIDIAAI/status/2060035668655677804





Leave a Reply