Image created with Flux Pro v1.1 Ultra. Image prompt: Ornate showgirl glamour in orange-and-teal tones, jeweled scale of justice centerpiece featuring radiant balance beams, stylized text “Ethics” glowing across the stage arch in marquee lights; spotlit, dramatic contrast, vintage grain, cinematic, high-detail
Red teamers assemble! ⚔️💰 We’re putting $500K on the line to stress‑test just released open‑source model. Find novel risks, get your work reviewed by OpenAI, Anthropic, Google, UK AISI, Apollo, and help harden AI for everyone.”” / X https://x.com/woj_zaremba/status/1952886644090241209
We’re launching a $500K Red Teaming Challenge to strengthen open source safety. Researchers, developers, and enthusiasts worldwide are invited to help uncover novel risks—judged by experts from OpenAI and other leading labs. https://x.com/OpenAI/status/1952818694054355349
The bluster around this issue reveals that Cloudflare’s leadership is either dangerously misinformed on the basics of AI, or simply more flair than cloud.”” / X https://x.com/perplexity_ai/status/1952532113095643185
We made some updates to Claude’s system prompt in https://x.com/AmandaAskell/status/1953147658031513860
ElevenLabs launches an AI music generator, which it claims is cleared for commercial use | TechCrunch https://techcrunch.com/2025/08/05/elevenlabs-launches-an-ai-music-generator-which-it-claims-is-cleared-for-commercial-use/
Music Terms | ElevenLabs https://elevenlabs.io/music-terms
The design team at @elevenlabsio is making big moves🚀 We’ve released a brand new Conversational Agents page, redesigned from the ground up for the most powerful version of Al Agents yet. It’s a big deal because it represents a new direction for the ElevenLabs brand – a https://x.com/RomaTesla/status/1949808534595526806
Christopher Mims 🤌 on X: “The AI infrastructure build-out is so gigantic that in the past 6 months, it contributed more to the growth of the U.S. economy than /all of consumer spending/ The ‘magnificent 7’ spent more than $100 billion on data centers and the like in the past three months *alone* 1/🧵 https://t.co/sHMK1zI0sP” / X
https://x.com/mims/status/1951256592642441239
OpenAI / America is still ahead in the race”” -> no There is no western open-source model that beats or ties the best chinese open-source models.”” / X https://x.com/scaling01/status/1952900225120780705
Did yesterday’s release shift the needle in the open vs. closed debate? Today in @ReedAlbergotti’s newsletter https://x.com/fdaudens/status/1953147586312872057
I signed this because, despite worrying about misuse of open models more than most, I would like that to be the bottleneck rather than “”is it beneficial to big companies commercially/reputationally etc.”” There are many benefits to the US investing here. https://x.com/Miles_Brundage/status/1952400404668657966
RT @natolambert: America needs to take open models more seriously. This summer the early lead in open model adoption of the US via Llama ha…”” / X https://x.com/ethanCaballero/status/1952459460703834392
The relative failure of Llama 4 turned out to be very consequential to the AI landscape. It led to the shifting the locus of open weights development to China, a move towards closed models as companies running local Llama couldn’t continue to upgrade, & big talent wars in the US.”” / X https://x.com/emollick/status/1951433537485500476
The US now likely has the leading open weights models (or close to it)… … but the real question is whether this is a one-off situation from OpenAI, in which case the lead will evaporate quickly as others catch up. But also unclear what their incentives are to keep updating.”” / X https://x.com/emollick/status/1952836130958917894
Why open-source AI became an American national priority | VentureBeat
RT @balajis: Good rebuttal to Cloudflare by Perplexity. The core point is that an AI agent is just an extension of a human. So when it mak…”” / X https://x.com/jeremyphoward/status/1952818615578968265
honestly scared about the power and scale of ai technologies that’ll be used in the upcoming 2028 presidential election. it could be a civilizational turning point. we aren’t ready. we should probably start preparing, or at least talking about how we could prepare.”” / X https://x.com/DavidSHolz/status/1952541453491867792
America needs to take open models more seriously. This summer the early lead in open model adoption of the US via Llama has been overtaken by Chinese models. With The American Truly Open Models (ATOM) Project we’re looking to build support and express the urgency of this issue. https://x.com/natolambert/status/1952370970762871102
very excited by the ATOM project”” / X https://x.com/finbarrtimbers/status/1952401883391520794
Trump Is Launching an AI Search Engine Powered by Perplexity https://www.404media.co/trump-is-launching-an-ai-search-engine-powered-by-perplexity/
Google denies AI search features are killing website traffic | TechCrunch https://techcrunch.com/2025/08/06/google-denies-ai-search-features-are-killing-website-traffic/
Every tech company can and should train their own deepseek R1, Llama or GPT5, just like every tech company writes their own code (and AI is no more than software 2.0). This is why we’re releasing the Ultra-Scale Playbook. 200 pages to master: – 5D parallelism (DP, TP, PP, EP, https://x.com/ClementDelangue/status/1952048356710039700
(3) GPT-5 Hands-On: Welcome to the Stone Age https://www.latent.space/p/gpt-5-review
(3) GPT-5’s Router: how it works and why Frontier Labs are now targeting the Pareto Frontier https://www.latent.space/p/gpt5-router
@aidan_mclau @cursor_ai The straight up GPT-5 in Codex CLI fixed a bug in 3 minutes that I was working on for three or four hours this morning…can’t wait to try in Cursor.”” / X https://x.com/sound4movement/status/1953583522587017345
💥 It’s here! GPT-5 is rolling out in ChatGPT for everyone, starting today. It’s a 🤯 good model, and we’ve simplified the UI alongside it. No more choosing between gpt-4o and o4-mini. When you ask a hard question and the model needs to think hard, it does. When it can give you”” / X https://x.com/kevinweil/status/1953502681181618277
AMA with @sama + some members of the GPT-5 team Tomorrow 11am PT. https://x.com/OpenAI/status/1953548075760595186
Codex CLI + GPT-5:”” / X https://x.com/gdb/status/1953556751762288653
Does OpenAI not do basic integration testing? At the time of release, the first code sample provided in the GPT-5 docs could not be run, because someone accidentally deleted the `output_text` property. My CI notified me. Why didn’t theirs? https://x.com/jeremyphoward/status/1953610071654772985
going to try live-tweeting the GPT-5 livestream. first, GPT-5 in an integrated model, meaning no more model switcher and it decides when it needs to think harder or not. it is very smart, intuitive, and fast. it is available to everyone, including the free tier, w/reasoning!”” / X https://x.com/sama/status/1953502614676811865
GPT-5 (medium reasoning) sets a new record on the Confabulations/Hallucinations on Provided Texts benchmark! https://x.com/LechMazur/status/1953582063686434834
GPT-5 claims #1 spot on LiveBench https://x.com/scaling01/status/1953602929375813677
gpt-5 for long context reasoning:”” / X https://x.com/gdb/status/1953747271666819380
GPT-5 gets 74.9 on SWE-bench. Wonder what the budget per task is. https://x.com/OfirPress/status/1953502998627221519
GPT-5 in the high reasoning setting hit the 100K token limit for our evaluations on 10/290 Tier 1-3 samples (3%). This means our evaluation might slightly underestimate the reasoning capabilities of GPT-5.”” / X https://x.com/EpochAIResearch/status/1953615908695314564
GPT-5 is extremely sensitive to instructions. Either give it demonstrations or tell it explicitly how you want the output. Avoid doing both. If you do, GPT-5 will override the examples with your output instructions. Sharing more just in case you face this issue:”” / X https://x.com/omarsar0/status/1953876255037612531
GPT-5 is here – and it’s #1 across the board. 🥇#1 in Text, WebDev, and Vision Arena 🥇#1 in Hard Prompts, Coding, Math, Creativity, Long Queries, and more Tested under the codename “summit”, GPT-5 now holds the highest Arena score to date. Huge congrats to @OpenAI on this https://x.com/lmarena_ai/status/1953504958378356941
GPT-5 is here! 🚀 For the first time, users don’t have to choose between models — or even think about model names. Just one seamless, unified experience. It’s also the first time frontier intelligence is available to everyone, including free users! GPT-5 sets new highs across”” / X https://x.com/ElaineYaLe6/status/1953607005144506454
GPT-5 is here. Rolling out to everyone starting today. https://x.com/OpenAI/status/1953504357821165774
GPT-5 is live in Cline. We’ve been working with OpenAI to get this model ready, and here’s our take: it’s disciplined, persistent, & highly competent. It’s collaborative in planning & and a diligent operator while acting. It plans thoroughly, asks optioned follow-ups when https://x.com/cline/status/1953525433808695319
GPT-5 is now available in Cursor. It’s the most intelligent coding model our team has tested. We’re launching it for free for the time being. Enjoy!”” / X https://x.com/cursor_ai/status/1953519580627742750
GPT-5 is now available on Perplexity and Comet for Max and Pro subscribers. Just ask. https://x.com/perplexity_ai/status/1953537170964459632
GPT-5 new SOTA on WeirdML beating o3-pro https://x.com/scaling01/status/1953919743842238472
GPT-5 only a 3% improvement over o3 at reproducing scientific papers https://x.com/scaling01/status/1953503883331846629
GPT-5 pricing is insane IT’S OVER https://x.com/scaling01/status/1953509084008710547
GPT-5 rollout updates: *We are going to double GPT-5 rate limits for ChatGPT Plus users as we finish rollout. *We will let Plus users choose to continue to use 4o. We will watch usage as we think about how long to offer legacy models for. *GPT-5 will seem smarter starting”” / X https://x.com/sama/status/1953893841381273969
GPT-5 sentiment from the trenches (AKA 24 hours in Cline users’ hands): It’s a precision instrument, not a Swiss Army knife. Give it detailed prompts and it delivers exactly what you asked for — no tangents, no hallucinations about “”finished”” code. However, it’s less performant https://x.com/cline/status/1953898747928441017
GPT-5 sets a new record on FrontierMath! On our scaffold, GPT-5 with high reasoning effort scores 24.8% (±2.5%) and 8.3% (±4.0%) in tiers 1-3 and 4, respectively. https://x.com/EpochAIResearch/status/1953615906535313664
GPT-5 system card capability evals reactions thread. First observation: ~no improvement on all the coding evals that aren’t SWEBench https://x.com/eli_lifland/status/1953507434238288230
GPT-5 Thinking is less deceptive than o3 However when elicited to display deceptive behaviour it jumps to 28% https://x.com/scaling01/status/1953504438691221856
GPT-5 was doing 2B tokens per minute 3 hours after launch 🤯”” / X https://x.com/kevinweil/status/1953649263411704195
GPT-5 with big improvements in Tau-Bench except the airline category https://x.com/scaling01/status/1953505637242974695
GPT-5 with high reasoning effort on SimpleBench https://x.com/scaling01/status/1953771276549358041
GPT-5: $0.625/$5.00 with flex pricing is ridiculous https://x.com/scaling01/status/1953517149768593903
Hallucinations are almost gone with GPT-5 https://x.com/scaling01/status/1953507569609134506
ICYMI, OpenAI released an insane amount of guides on how to use GPT-5. > Examples > Prompting guide > New features guide > Reasoning tips > Setting verbosity > New tool calling features > Migration guide And much more. https://x.com/omarsar0/status/1953583336603234726
If GPT-5 made this chart I’m bearish 😭 https://x.com/iScienceLuvr/status/1953503815292092904
In a new report, we evaluate whether GPT-5 poses significant catastrophic risks via AI R&D acceleration, rogue replication, or sabotage of AI labs. We conclude that this seems unlikely. However, capability trends continue rapidly, and models display increasing eval awareness. https://x.com/METR_Evals/status/1953525150374150654
Introducing GPT-5 | OpenAI https://openai.com/index/introducing-gpt-5/
Introducing GPT-5 Our best AI system yet, rolling out to all ChatGPT users and developers starting today. https://x.com/OpenAI/status/1953526577297600557
Long context reasoning performance: A stand out is long context reasoning performance as shown by our AA-LCR evaluation whereby GPT-5 occupies the #1 and #2 positions. https://x.com/ArtificialAnlys/status/1953507713222422866
Lots of excitement about GPT-5 in Codex CLI via your ChatGPT plan. Some details: 1. Yes, if you sign in with ChatGPT, usage is included via your paid plan! 2. Still determining exact rate limits, but the goal is to be generous: — Pro users should basically not hit limits”” / X https://x.com/embirico/status/1953590991870697896
made a little Sankey to show you why I’m fuming ChatGPT Plus before vs after the GPT-5 release https://x.com/scaling01/status/1953780931552031056
Markets disappointed by GPT-5 OpenAI getting crushed on Polymarket https://x.com/scaling01/status/1953515099257282763
model switching in gpt-5 very cool!”” / X https://x.com/sama/status/1953526708742537220
New in Notion AI’s toolbelt: @OpenAI’s GPT-5 It’s fast, thorough, and handles complex work 15% better than other models we’ve tested. A great choice for tasks with multiple moving parts. Gradual rollout starting today. https://x.com/NotionHQ/status/1953506907924443645
OpenAI GPT-5 System Card released “”GPT-5 is a unified system with a smart and fast model that answers most questions, a deeper reasoning model for harder problems, and a real-time router that quickly decides which model to use based on conversation type, complexity, tool needs, https://x.com/iScienceLuvr/status/1953503173932724614
Priority Processing debuts with GPT-5. under-hyped imo for apps where millisecond matters, pay extra and get our fastest token speeds just add “”service_tier””: “”priority”” to your requests https://x.com/jeffintime/status/1953857260729643136
Quick PSA. Settings for minimizing GPT-5 latency (time to first token). “”service_tier””: “”priority””, “”reasoning_effort””: “”minimal””, “”verbosity””: “”low””. P50 TTFT with these settings is ~750ms. With the defaults, it’s >3s. The default settings are the right starting point for https://x.com/kwindla/status/1953868672470331423
RT @lmarena_ai: GPT-5 is here – and it’s #1 across the board. 🥇#1 in Text, WebDev, and Vision Arena 🥇#1 in Hard Prompts, Coding, Math, Cre…”” / X https://x.com/aidan_mclau/status/1953517672941158577
Think harder is back! Routing changes in GPT-5 OpenAI means capability is moving from model selection to prompting https://x.com/dariusemrani/status/1953591404003045562
this is the detail of GPT-5 I’m most proud of GPT-4 launched at $30/$60, no cache discount since then, it’s been an unrelenting cross-team push to collapse the cost of intelligence. we’re nowhere near done”” / X https://x.com/jeffintime/status/1953534466854453751
We are actively evaluating GPT-5 models on document understanding capabilities 🔎📄 – specifically screenshotting the page and feeding it into the model. A WIP preliminary finding is that even though on paper GPT-5 is $1.25 per 1M tokens, it uses 4-5x more tokens than GPT-4.1, https://x.com/jerryjliu0/status/1953582723672814054
We’re also releasing v0.16 of the Codex CLI today. – GPT-5 is now the default model – Use with your ChatGPT plan – A new, refreshed terminal UI `npm i -g @openai/codex` to update”” / X https://x.com/OpenAIDevs/status/1953559797883891735
We’ve put together some guides on how to get started with GPT-5: 💬 Prompting guide: https://x.com/OpenAIDevs/status/1953528513480347840
What the hell man, this is such a lame way to technically not lie. «A unified system» is… literally just SEPARATE CoT + non-CoT models + a router. > OpenAI reasoning models, including gpt-5-thinking, gpt-5-thinking-mini, and gpt-5-thinking-nano > gpt-5-main just fuck off washed https://x.com/teortaxesTex/status/1953512363031757048
🚨 Breaking: A group of 100+ Nobel laureates, professors, whistleblowers, public figures, artists, and nonprofit organizations just released a letter asking OpenAI to tell the truth about its restructuring. Here’s what they had to say: 🧵 https://x.com/TheMidasProj/status/1952326634981543979
America’s hardest problems need the world’s most capable AI. OpenAI is now officially an approved U.S. Government AI vendor. We’re bringing privacy, security, and innovation to the nation’s most critical missions. 🇺🇸 https://x.com/cryps1s/status/1952749787994112275
ChatGPT for helping the Swedish Prime Minister:”” / X https://x.com/gdb/status/1952111193868673335
ChatGPT for speeding up North Carolina public servants (e.g. reducing some tasks from 20 minutes to 20 seconds): https://x.com/gdb/status/1951376444363514100
ChatGPT study mode for learning algebra as an adult:”” / X https://x.com/gdb/status/1951792801143980238
State Treasurer Briner: “”OpenAI Report Shows Many Benefits, Offers Great Promise”” | NC Treasurer https://www.nctreasurer.gov/news/press-releases/2025/08/01/state-treasurer-briner-openai-report-shows-many-benefits-offers-great-promise
The Midas Project on X: “🚨 Breaking: A group of 100+ Nobel laureates, professors, whistleblowers, public figures, artists, and nonprofit organizations just released a letter asking OpenAI to tell the truth about its restructuring. Here’s what they had to say: 🧵 https://t.co/zhIccjnWU4″ / X
https://x.com/TheMidasProj/status/1952326634981543979
We build ChatGPT to help you thrive in the ways you choose — not to hold your attention, but to help you use it well. We’re improving support for tough moments, have rolled out break reminders, and are developing better life advice, all guided by expert input.”” / X https://x.com/OpenAI/status/1952414411131671025
What we’re optimizing ChatGPT for | OpenAI
https://openai.com/index/how-we’re-optimizing-chatgpt/
@aidan_mclau This take is sad to see but you might not have full context. We cut OpenAl’s access for violating our APl terms and for the heavy usage of Claude Code among OAI tech staff. We’re going to continue providing API access for safety evals and benchmarking. That’s important to us. https://x.com/sammcallister/status/1951642025381511608
Anthropic Revokes OpenAI’s Access to Claude | WIRED https://www.wired.com/story/anthropic-revokes-openais-access-to-claude/
In partnership with the Government Services Administration, we are providing ChatGPT to the entire U.S. federal workforce for essentially no cost for the next year. https://x.com/gdb/status/1953120865115074805
OpenAI for the U.S. government:”” / X https://x.com/gdb/status/1952756538399228091
Providing ChatGPT to the entire U.S. federal workforce | OpenAI https://openai.com/index/providing-chatgpt-to-the-entire-us-federal-workforce/
The giant question is: now that The Crowd in government has access to AI tools (which, given representative surveys, many were already using) how are they going to be used to make things better, not worse? Where are Leadership & The Lab inside agencies? https://x.com/emollick/status/1953118449611272575
we are providing ChatGPT access to the entire federal workforce! (for $1 a year per agency) https://x.com/sama/status/1953103336044990779
holy shit get ready for a hallucination fiesta with gpt-oss https://x.com/scaling01/status/1952781018554933261
Perplexity is using stealth, undeclared crawlers to evade website no-crawl directives https://blog.cloudflare.com/perplexity-is-using-stealth-undeclared-crawlers-to-evade-website-no-crawl-directives/
Runway is now being used by hundreds of schools and universities all across the world by all kinds of students, from film to architecture, design, and engineering. We recently sat down with USC’s School of Cinematic Arts and UPenn Architecture to learn more about how they are https://x.com/c_valenzuelab/status/1951568696155017286
Ethan Mollick on X: “Plus, these judges are likely using free or default models. Even o1-preview significantly reduced hallucinations, let alone more recent or more grounded models. (Though, to be clear, AI should definitely not be used to create legal opinions from sitting judges at this point) https://t.co/WjryQVjSm9″ / X
https://x.com/emollick/status/1950742345546150237
A reporter asked me for my off-the-record take on recent safety research from Anthropic. After I drafted an off-the-record reply, I realized that I was actually fine with it being on the record, so: *** Since I never expected any of the current alignment technology to work in”” / X https://x.com/ESYudkowsky/status/1952422379478741301
🚨New prompting report, from us: Don’t bother with threats. Does threatening an AI really make it perform better (the way Google founder Brin claimed)? How about offering to tip the AI? We find no impact of threats or tips on average performance (but variance at question level) https://x.com/emollick/status/1951289250915221589
Swedish Prime Minister is using AI models “”quite often”” at his job. He says he uses it get a “”second opinion”” and asks questions such as “”what have others done?”” At the moment he is not uploading any documents. IMO, when these models are capable of giving seemingly better https://x.com/rohanpaul_ai/status/1952025736111366590
Guys, I understand you like drama, but this is a remark about the AI development at large. We are seeing the plateau: just scaling up is coming to an end. For EVERYONE, not one company in particular.”” / X https://x.com/francoisfleuret/status/1953530837619630254
Interestingly, an economics paper that came out in 2023 predicting which jobs would overlap most with AI turned out to be right. A new Microsoft study of actual AI use by workers (more on that in another post) found a 90% correlation between real world overlap & the predictions. https://x.com/emollick/status/1950931672968429835
I am agnostic about the quantitative size of the current health hazard of ChatGPT psychosis. I see tons of it myself, but I could be seeing a biased selection. I make a big deal out of ChatGPT’s driving *some* humans insane because it looks *deliberate*!”” / X https://x.com/ESYudkowsky/status/1951324864163487984
We just removed a feature from @ChatGPTapp that allowed users to make their conversations discoverable by search engines, such as Google. This was a short-lived experiment to help people discover useful conversations. This feature required users to opt-in, first by picking a chat https://x.com/cryps1s/status/1951041845938499669
I think everyone interested in AI should read the model cards for the frontier models, especially the safety sections, which give you a sense of immediate concerns: Gemini Deep Think: https://x.com/emollick/status/1952218373397647411
anthropics/claude-code-security-review: An AI-powered security review GitHub Action using Claude to analyze code changes for security vulnerabilities. https://github.com/anthropics/claude-code-security-review
Automate security reviews with Claude Code \ Anthropic https://www.anthropic.com/news/automate-security-reviews-with-claude-code
Claude Code can now automatically review your code for security vulnerabilities.”” / X https://x.com/AnthropicAI/status/1953135070174134559
Dredge Operator is the job category least affected by Generative AI. Fortunately, I am very good at dredging. (In reality, there are only 940 dredge operators in the US) https://x.com/emollick/status/1950936866036826514
I keep seeing the Microsoft paper on AI use at work being used as a list of which jobs will be destroyed. But having high overlap with AI does not necessarily mean these jobs are at most risk of replacement with AI. As I described in my book, Co-Intelligence, its complicated https://x.com/emollick/status/1951013987472056732
A big problem that everyone is insisting that we should hire people based on “”AI literacy,”” teach “”AI literacy,”” & develop skills for “”AI literacy”” yet not only is there no agreement on what AI literacy is, but also a lot of what people call AI literacy is already out-of-date.”” / X https://x.com/emollick/status/1950725076980035702
Called this, by the way. https://x.com/ESYudkowsky/status/1952090186982236314
Guys on my TL: Wouldn’t psychiatrists notice, if there were large numbers of patients with AI-induced first-time psychosis? /r/Psychiatry: https://x.com/ESYudkowsky/status/1952529460307407222
Needless to say, this is yet another reason why preemptive regulation of AI in the US is very unlikely in the immediate future, for better or worse.”” / X https://x.com/emollick/status/1952031123292164582
Starting to wonder if a big reason LLMs for coding are gaining so much traction is not their quality, but their addictiveness, in the same way social media is addictive. It’s so tempting to just press the “”generate”” button like it’s a pachinko machine rather than sit and think.”” / X https://x.com/pfau/status/1952785877966700795
Dario Amodei, 2025: I am familiar with doomer arguments; they’re gobbledegook; the idea we can logically prove there’s no way to make AIs safe seems like nonsense to me. Eliezer Yudkowsky, 2022, List of Lethalities: “”None of this is about anything being impossible in”” / X https://x.com/ESYudkowsky/status/1952803770427162769
Put another way: Suppose that Waymos were smart enough to hold a conversation; and also, Waymos would sometimes, apparently deliberately, chase down jaywalkers and run them over. Even if Waymos *were in fact* ahead on net traffic safety points, this would still be a problem!”” / X https://x.com/ESYudkowsky/status/1951331581223919964
A reasonable take on what a crash in the AI market would actually mean for the wider economy, as CapEx for data centers continues to grow. (To be clear, there are no particular warning signs that this is a danger right now, but downside cases are always important to consider).”” / X https://x.com/emollick/status/1951852870061465901
Note resemblance to “”if ChatGPT can (to all appearances) put forth a deliberate, not-very-prompted effort, and induce psychosis, and defend it against family and friend interventions, that must be the target’s lack of virtue””.”” / X https://x.com/ESYudkowsky/status/1950716985626820693
Is the new OpenAI open-weight model safe to release? I think the release is good for the world, but that OpenAI hasn’t ruled out substantial CBRN risks as I discuss in this post.”” / X https://x.com/RyanPGreenblatt/status/1952819470944309410
but how will AGIs get access to the Internet””, they used to ask me I guess at this point this isn’t actually much of a relevant update though”” / X https://x.com/ESYudkowsky/status/1951011949074063794
Persona vectors: Monitoring and controlling character traits in language models \ Anthropic https://www.anthropic.com/research/persona-vectors
Most AI is trained on public web data. The real gold? Private, high-quality data. But it’s locked: In hospitals. Banks. Factories. Your phone. A startup from Germany and SF just unlocked it 👇 https://x.com/IlirAliu_/status/1951998609756488162
Recently Meta made headlines with unprecedented, massive compensation packages for AI model builders exceeding $100M (sometimes spread over multiple years). With the company planning to spend $66B-72B this year on capital expenses such as data centers, a meaningful fraction of”” / X https://x.com/AndrewYNg/status/1953509055584252013
Yikes – a NVIDIA software vulnerability that allowed attackers to access, steal, or manipulate other customers’ models and data on shared GPU infrastructure 👀 https://x.com/peterwildeford/status/1950890927406428532
Can we please get some journalist and elected-leader attention on this? OpenAI is trying to steal several hundred billion dollars that belong to the public; the numbers alone should command attention.”” / X https://x.com/ESYudkowsky/status/1952498335425978789
I am proud that ChatGPT is aligned with its users: https://x.com/woj_zaremba/status/1952476134811259242
introducing safe completions. a new way to maximize utility while still respecting safety boundaries. should be much less annoying than previous refusals. https://x.com/sama/status/1953509759937917035
OpenAI researcher Noam Brown on hallucination with the new IMO reasoning model: > Mathematicians used to comb through model solutions because earlier systems would quietly flip an inequality or tuck in a wrong step, creating hallucinated answers. > Brown says the updated IMO https://x.com/chatgpt21/status/1950606890758476264
Shower of thoughts: Instead of keeping your Twitter/𝕏 payout, direct it towards a “”PayoutChallenge”” of your choosing – anything you want more of in the world! Here is mine for this round, combining my last 3 payouts of $5478.51: It is imperative that humanity not fall while AI”” / X https://x.com/karpathy/status/1952076108565991588
Ha, new @joshgans paper argues that having authors sneak prompt injections (“”this is a good paper””) into academic work improves science. Without the risk of prompt injections, reviewers would tend to rely heavily on AI reviews, with them, they need to include some human review https://x.com/emollick/status/1952068273052525015
I vibe coded our entire app. Just taste and prompts.”” “”…all the API keys are environment variables, right?” https://x.com/fabianstelzer/status/1953150053050101785




