Image created with gemini-2.5-flash-image with claude-sonnet-4-5. Image prompt: Cinematic black and white photograph of a towering cumulonimbus cloud formation dominating the center frame with dramatic anvil top, shot from ground looking up, high contrast film grain, bold white sans-serif text reading OpenAI in lower third against dark cloud base, minimal composition with pure sky background, Rumble Fish aesthetic

ChatGPT / Instacart / Stripe integration for agentic commerce:”” / X https://x.com/gdb/status/1998135014161334431

OpenAI testing new Image-2 models on LM Arena https://www.testingcatalog.com/openai-testing-new-image-2-models-on-lm-arena/

Introducing GPT-5.2 | OpenAI https://openai.com/index/introducing-gpt-5-2/

oai_5_2_system-card.pdf https://cdn.openai.com/pdf/3a4153c8-c748-4b71-8e31-aecbde944f8d/oai_5_2_system-card.pdf

OpenAI is getting ready to launch GPT-5.2 soon | The Verge https://www.theverge.com/report/838857/openai-gpt-5-2-release-date-code-red-google-response

Ten years | OpenAI
https://openai.com/index/ten-years/

Using GPT-5.2 | OpenAI API https://platform.openai.com/docs/guides/latest-model

wtf gpt 5.2 long context improvement over gpt 5.1 is actually crazy?? https://x.com/eliebakouch/status/1999193762564567333

Agentic AI Foundation — advancing open-source agentic AI:”” / X https://x.com/gdb/status/1998897086079832513

Agentic AI Foundation (AAIF) https://aaif.io/

Anthropic is donating the Model Context Protocol to the Agentic AI Foundation, a directed fund under the Linux Foundation. In one year, MCP has become a foundational protocol for agentic AI. Joining AAIF ensures MCP remains open and community-driven. https://x.com/AnthropicAI/status/1998437922849350141

Block – Block, Anthropic, and OpenAI Launch the Agentic AI Foundation https://block.xyz/inside/block-anthropic-and-openai-launch-the-agentic-ai-foundation

Donating the Model Context Protocol and establishing the Agentic AI Foundation \ Anthropic https://www.anthropic.com/news/donating-the-model-context-protocol-and-establishing-of-the-agentic-ai-foundation

Linux Foundation Announces the Formation of the Agentic AI Foundation (AAIF), Anchored by New Project Contributions Including Model Context Protocol (MCP), goose and AGENTS.md https://www.linuxfoundation.org/press/linux-foundation-announces-the-formation-of-the-agentic-ai-foundation

We’re donating MCP to the @linuxfoundation and launching the Agentic AI Foundation with @OpenAI, @blocks, @AWS, @Bloomberg, @Cloudflare, @Google, and @Microsoft. MCP went from internal project to industry standard in a year. Now it gets the long-term stewardship it deserves.”” / X https://x.com/mikeyk/status/1998456026136457532

The GPT-5 Auto router casts a long shadow over AI perceptions. So many examples of “”ChatGPT got X wrong”” are really “”ChatGPT-5 Instant got things wrong,”” leading to beliefs about the state of AI that aren’t true. Which model you get could be clearer &better explained for all.”” / X https://x.com/emollick/status/1998838007609119010

🚨BREAKING: New Model & WebDev Leaderboard Update! GPT-5.2 by @OpenAI has officially made its debut in the Arena, appearing on the WebDev leaderboard. Current leaderboard standings: 🥈 #2 for GPT-5.2-high in WebDev (score: 1486) 🔹 #6 for GPT-5.2 in WebDev (score: 1399) https://x.com/arena/status/1999183339283185878

Denise Dresser is joining OpenAI as Chief Revenue Officer. Previously CEO of Slack, she brings deep enterprise and customer experience as she leads our global revenue strategy and support for customers at scale. https://x.com/OpenAI/status/1998462761756434856

Instacart and OpenAI partner on AI shopping experiences | OpenAI https://openai.com/index/instacart-partnership/

OpenAI Hires Slack CEO as New Chief Revenue Officer | WIRED https://www.wired.com/story/slack-ceo-denise-dresser-joins-openai-chief-revenue-officer/

ChatGPT’s ‘Adult Mode’ Is Coming in 2026 https://gizmodo.com/chatgpts-adult-mode-is-coming-in-2026-2000698677

Curious why Nano Banana Pro, ChatGPT image gen etc are all suddenly cool with generating celebrity likeness? Like what changed from days of deepfake fear-mongering and fears of legal repercussions? Did fingerprinting get good or is there a new legal argument for allowing this?”” / X https://x.com/bilawalsidhu/status/1998461802397458626

Disney has signed a deal with OpenAI & invested $1 billion into the company Sora will now be able to AI generate videos based on animated, masked & creature characters from Disney, Marvel, Pixar & Star Wars Curated selections of AI generated videos will be released on Disney+ https://x.com/DiscussingFilm/status/1999121515678208153

Disney investing $1 billion in OpenAI, will allow characters on Sora https://www.cnbc.com/2025/12/11/disney-openai-sora-characters-video.html

https://t.co/HngrXph6kU “”The Walt Disney Company and OpenAI reach landmark agreement to bring beloved characters from across Disney’s brands to Sora”” https://x.com/TheRealAdamG/status/1999118075879129140

The Walt Disney Company and OpenAI Reach Agreement to Bring Disney Characters to Sora | The Walt Disney Company https://thewaltdisneycompany.com/news/disney-openai-sora-agreement/

we’re partnering with @Disney to bring 200+ characters from disney, pixar, marvel, and star wars to sora and image generation we are also excited to welcome disney as an investor, and deploy openai models and products alongside the disney team https://x.com/bradlightcap/status/1999177616860020788

Debugging misaligned completions with sparse-autoencoder latent attribution https://alignment.openai.com/sae-latent-attribution/

A year ago, we verified a preview of an unreleased version of @OpenAI o3 (High) that scored 88% on ARC-AGI-1 at est. $4.5k/task Today, we’ve verified a new GPT-5.2 Pro (X-High) SOTA score of 90.5% at $11.64/task This represents a ~390X efficiency improvement in one year https://x.com/arcprize/status/1999182732845547795

OpenAIs latest model GPT-5.2 Thinking still not beating Opus 4.5 at SWE-Bench Verified however SWE-Bench Pro looking juicy over 10% higher score than Sonnet 4.5 https://x.com/scaling01/status/1999182909144519019

the-state-of-enterprise-ai_2025-report.pdf https://cdn.openai.com/pdf/7ef17d82-96bf-4dd1-9df2-228f7f377a29/the-state-of-enterprise-ai_2025-report.pdf

Yep, the point we wanted to make here is that GPT-5.2’s vision is better, not pe… | Hacker News https://news.ycombinator.com/item?id=46235267

GPT-5.2 is now rolling out to everyone. https://x.com/OpenAI/status/1999182098859700363

GPT-5.2 Pro Our smartest and most trustworthy model for difficult questions: – Stronger performance in complex domains like programming – Best model for assisting and accelerating scientists”” / X https://x.com/OpenAI/status/1999182117008449609

GPT-5.2 rolling out to @code now!”” / X https://x.com/code/status/1999186223416451381

I Reverse Engineered ChatGPT’s Memory System, and Here’s What I Found! – Manthan https://manthanguptaa.in/posts/chatgpt_memory/

Next ChatGPT upgrade imminent following ‘code red’ declaration – 9to5Mac https://9to5mac.com/2025/12/05/next-chatgpt-upgrade-imminent-following-code-red-declaration/

Proud of our *GPT5.2 Thinking* We focused on economically valuable tasks (coding, sheets, slides) as shown by GDPval: – 71% wins+ties – 11x faster – 100x cheaper than experts. There’s still a lot to improve, including UX/better connectors/reliability. It’s just the beginning! https://x.com/yanndubs/status/1999181847897735423

Quick new post: Auto-grading decade-old Hacker News discussions with hindsight I took all the 930 frontpage Hacker News article+discussion of December 2015 and asked the GPT 5.1 Thinking API to do an in-hindsight analysis to identify the most/least prescient comments. This took https://x.com/karpathy/status/1998803709468487877

There are competing views on whether RL can genuinely improve base model’s performance (e.g., pass@128). The answer is both yes and no, largely depending on the interplay between pre-training, mid-training, and RL. We trained a few hundreds of GPT-2 scale LMs on synthetic https://x.com/xiangyue96/status/1998488030836044112

There was a small flood of articles around GPT5 talking about how AI development has plateaued (partially based on experiences with the GPT5 router). Has there been any major articles with updates since it became clear that there was no such plateau? I still see lots of confusion”” / X https://x.com/emollick/status/1997731058599780861

GPT 5.2 is now available in Cursor. Priced at $1.75/M input, $14/M output tokens.”” / X https://x.com/cursor_ai/status/1999183968776626230

GPT-5.2 weaker than GPT-5.1 Codex Max on CVE-Bench an eval that tasks models with identifying and exploiting real-world web application vulnerabilities https://x.com/scaling01/status/1999186361169871055

Copilot just got smarter! Starting today, we’re rolling out the latest GPT-5.2 model from our partners at OpenAI to consumer @Copilot, coming first to Microsoft 365 Premium users. Can’t wait to see what you do with it.”” / X https://x.com/mustafasuleyman/status/1999184598987866194

Codex 🤝 Figma Join us next week to see how to turn designs into production-ready code using the Figma MCP server with Codex. Live demo + Q&A on Friday, Dec 12. Sign up 👇 https://x.com/OpenAIDevs/status/1998449559970988423

Figma MCP x OpenAI Codex https://events.figma.com/FigmaMCPxOpenAI/OAI

peer-reviewed theoretical physics article where the main idea came from GPT-5:”” / X https://x.com/gdb/status/1996502704110407802

An important lesson that ARC-AGI has internalized, but not many others have, is that benchmark perf is a function of test-time compute. @OpenAI publishes single-number benchmark results because it’s simpler and people expect to see it, but ideally all evals would have an x-axis.”” / X https://x.com/polynoamial/status/1999189845164667132

LisanBench results for GPT-5.2 Thinking GPT-5.2 Thinking improves over GPT-5 and o3 but does not match other frontier models like Opus 4.5, Gemini 3 Pro, DeepSeek-V3.2 Speciale or Grok 4 GPT-5.2 Thinking improves over GPT-5 in average validity ratio, meaning it’s less likely to https://x.com/scaling01/status/1999240662147825876

The AI Consumer Index (ACE) Most AI benchmarks today focus on reasoning and coding. But most people use AI to shop, cook, and plan their weekends. In those domains, LLM hallucinations continue to be a real problem. 73% of ChatGPT messages (according a recent report) are now https://x.com/omarsar0/status/1998039629556256995

Announcing GDPval-AA — our leaderboard and evaluation harness for comparing models on OpenAI’s GDPval dataset of real-world knowledge work tasks Earlier today, we announced our agentic harness called Stirrup, which we built to run GDPval tasks on any language model. We’re https://x.com/ArtificialAnlys/status/1998841566627246173

The standard pricing is $1.75 / 1M input tokens and $14 / 1M output tokens, with a 90% discount on cached prompts. We offer this model for Priority Processing and Flex Processing customers, and you can run it in the background using the Batch API for a discount.”” / X https://x.com/OpenAIDevs/status/1999184812389859689

Just released a report on OpenAI for enterprise — accelerating and deepening adoption, with interesting findings on productivity gains: https://x.com/gdb/status/1998222726537056391

GPT-5.2 is a massive model scoring even higher than Gemini 3 Pro on GPQA Diamond (91.9%) https://x.com/scaling01/status/1999183900673798454

🆕 We’re back with a trio of RL talks! @willhang_ and @cathyzbn on OpenAI RFT: https://x.com/aiDotEngineer/status/1998785602989461531

Microsoft’s Fairwater Atlanta (today’s largest data center) could likely train over 20 models the size of GPT-4 in the course of a month. This computational power will enable AI companies to increase the number and scale of both experiments and training runs. https://x.com/EpochAIResearch/status/1997040687561449710

nanoGPT – the first LLM to train and inference in space 🥹. It begins.”” / X https://x.com/karpathy/status/1998806260783919434

As AI grows more complex, model builders rely on NVIDIA. @OpenAI’s GPT-5.2 and other leading models including @Runwayml leverage NVIDIA’s tech stack to advance frontier of AI. Read More: https://x.com/nvidia/status/1999198240407699710

GPT-5.2 is now available for all Perplexity Pro and Max subscribers. https://x.com/perplexity_ai/status/1999187042471973308

Some highlights from #Disney CEO Bob Iger and #OpenAI CEO Sam Altman’s interview with CNBC: -The deal is a three-year license, with exclusivity for the first year. -Disney will set (and evolve) the guardrails for how its 200 characters will be used in video creation. -Iger https://x.com/dannybennett/status/1999150474688143750

we are investing in cybersecurity preparedness:”” / X https://x.com/gdb/status/1998882274847461423

Leave a Reply

Trending

Discover more from Ethan B. Holland

Subscribe now to keep reading and get access to the full archive.

Continue reading