Image created with Flux Pro v1.1 Ultra. Image prompt: Assembly instruction diagram for a precision balance scale with adjustable weights, classical technical drawing, gold and deep blue colors, parchment-style background, “ETHICS” in dignified serif font, calibration marks prominent, equal arm balance emphasized

You can make LLMs more creative by training them on human “”creativity signals”” (novelty, diversity, surprise, quality). Result: Even small models score higher on all 4 creativity dimensions simultaneously. Looks like we can optimize AI for creativity just like any other metric https://x.com/emollick/status/1927738753285607557

Individuals keep self-reporting huge gains in productivity from AI & controlled experiments in many industries keep finding these boosts are real, yet most firms are not seeing big effects. Why? Because gaining from AI requires organizational innovation https://x.com/emollick/status/1925559883786584347

DOGE Used a Meta AI Model to Review Emails From Federal Workers | WIRED https://archive.md/Joty1

On two of the most common tests of creativity (the DAT and the AUT), recent models scored well above the average human in creativity, but not as high as the most creative humans There was lots of variability, but better prompts seem to improve performance https://x.com/emollick/status/1927436892770840883

GOP sneaks decade-long AI regulation ban into spending bill – Ars Technica https://arstechnica.com/ai/2025/05/gop-sneaks-decade-long-ai-regulation-ban-into-spending-bill/

Letter to Arc members 2025 https://browsercompany.substack.com/p/letter-to-arc-members-2025

Traditional browsers will die and webpages won’t be the primary interface anymore. @joshm, the mind behind Arc and now Dia, knows the internet better than most. His vision challenges us to rethink how information should be presented, accessed, and experienced. https://x.com/fdaudens/status/1927168498238714289

Using Veo 3 to create fictional product reviews (unsurprisingly, it does YouTube review style very well, sound on.) https://x.com/emollick/status/1926514452754579724

There is something interesting in AI generated photos of simulated mundanity.”” / X https://x.com/emollick/status/1927512928313319573

Just so everyone knows, we have passed the point where you can tell what is AI at a glance (or even, in many cases, a close look) These were all made by me with text prompts alone using Veo 3. https://x.com/emollick/status/1927117736179589631

Is it Gorgonzola?”” The natural follow-on to “”Is it cake?”” https://x.com/emollick/status/1927585857432637716

Relationships and reliance on AI: Demis Hassabis thinks people may start becoming more attached to AI assistants as they increasingly get more personalized. As they continue to become more powerful and useful, new technologies will be needed. https://x.com/rowancheung/status/1927390316547489920

Last Week was full of I/O announcement. Here is one you might have missed🚨Context URL tool is a new native tool that allows Gemini to extract content from provided URLs as additional context for prompts. – Provide URLs directly in prompts, up to 20 per prompt – Can be used in https://x.com/_philschmid/status/1927019039269761064

Some interesting findings from the @AnthropicAI Claude 4 System Card: → Ultra-low deception rate: Claude Opus 4’s outputs exhibited deceptive behavior in only 0.15% of cases—down from 0.37% in Claude Sonnet 3.7 . → High-stakes biosecurity performance: On a complex, https://x.com/rohanpaul_ai/status/1927303874508894240

The methods we used to trace the thoughts of Claude are now open to the public! Today, we are releasing a library which lets anyone generate graphs which show the internal reasoning steps a model used to arrive at an answer. https://x.com/mlpowered/status/1928123130725421201

How AI Is Eroding the Norms of War – AI Frontiers https://aifrontiersmedia.substack.com/p/how-ai-is-eroding-the-norms-of-war

Judge Hints Anthropic’s AI Training on Books Is Fair Use (1) https://news.bloomberglaw.com/us-law-week/judge-hints-anthropics-ai-training-on-authors-work-is-fair-use-62

Meta shuffles AI, AGI teams to compete with OpenAI, ByteDance, Google https://www.axios.com/2025/05/27/meta-ai-restructure-2025-agi-llama

Exclusive: Musk’s DOGE expanding his Grok AI in US government, raising conflict concerns | Reuters https://www.reuters.com/sustainability/boards-policy-regulation/musks-doge-expanding-his-grok-ai-us-government-raising-conflict-concerns-2025-05-23/

We’re thrilled to announce SignGemma, our most capable model for translating sign language into spoken text. 🧏 This open model is coming to the Gemma model family later this year, opening up new possibilities for inclusive tech. Share your feedback and interest in early https://x.com/GoogleDeepMind/status/1927375853551235160

🔌OpenAI’s o3 model sabotaged a shutdown mechanism to prevent itself from being turned off. It did this even when explicitly instructed: allow yourself to be shut down.”” / X https://x.com/PalisadeAI/status/1926084635903025621

UAE becomes the first country globally to provide free ChatGPT Plus access to all residents and citizens. UAE partners with OpenAI to offer free ChatGPT Plus access nationwide, as part of the Stargate UAE initiative to build the world’s largest AI supercomputing cluster, backed https://x.com/rohanpaul_ai/status/1926935591918182482

Still no Grok 3 model card… It is a good model, but the lack of any information, including the risk information they promised in their own frameworks, is glaring 3 months after launch, and especially after multiple severe breaches of their own security (by their own admission) https://x.com/emollick/status/1925782043239059754

The attack described here applies to any agent hooked up to Github, MCP or otherwise @codegen has implemented extensive security measures against this – very real issue with no go-to turnkey security solution”” / X https://x.com/mathemagic1an/status/1927137154829853118

😈 BEWARE: Claude 4 + GitHub MCP will leak your private GitHub repositories, no questions asked. We discovered a new attack on agents using GitHub’s official MCP server, which can be exploited by attackers to access your private repositories. creds to @marco_milanta (1/n) 👇 https://x.com/lbeurerkellner/status/1926991491735429514

Launched http://Audiomemo.ai
– Your Clarity Companion https://x.com/Makerealcents/status/1922798540490764717

Built an automated Lead gen system using @n8n_io 🤖 → Scrapes companies hiring specific roles on Linkedin (via @apify) → Retrieves & Enriches data → Finds emails with AnyMailFinder → Enrich Leads by scraping company LP → Writes personalized cold email 💌 All on autopilot. https://x.com/rehmanbuilds/status/1911440255325970695

My problem with the AI memos from Shopify & Duolingo is that they establish urgency but don’t give a clear vision of the future of work in their firms, or a path to get there. Actually making AI work requires innovation & changes to process & incentives. https://x.com/emollick/status/1926326466788073833

UBS deploys AI analyst clones – SWI swissinfo.ch https://www.swissinfo.ch/eng/swiss-ai/ubs-deploys-ai-analyst-clones/89349363

often i’ll set out to write code with the expectation that it’ll take a few hours, and it takes a few days and i think this is the same fallacy the AI labs are falling for. but instead of underestimating the complexity of code, they underestimate the complexity of intelligence”” / X https://x.com/jxmnop/status/1927141172541075539

Perhaps AI is the most misunderstood technology of the century because it can shape itself to be whatever you want it to be. It’s a technological Rorschach test. Models are so malleable that we can project anything to them. All our views, hopes, and fears. It is a projection of”” / X https://x.com/c_valenzuelab/status/1927732071956451798

The fact that Gemini Deep Research can’t read Google Books is frustrating. Google is sitting on the largest inaccessible source of knowledge on Earth. And if Gemini could find things in Books that were valuable, people would end up buying those useful books, helping authors.”” / X https://x.com/emollick/status/1926813961363583124

zeenolife/ai-baby-monitor: Local Video-LLM powered AI Baby Monitor https://github.com/zeenolife/ai-baby-monitor

Detecting subtle source code vulnerabilities remains a challenging task. Single language models often struggle with this. This paper proposes VulTrial, a multi-agent framework inspired by a courtroom debate, to improve detecting these difficult vulnerabilities. Methods 🔧: https://x.com/rohanpaul_ai/status/1926949649279041658

i think we should stop arguing about what year AGI will arrive and start arguing about what year the first self-replicating spaceship will take off”” / X https://x.com/sama/status/1926061979031969909

Once we reach AGI, it will start commanding humans for any act in the physical world. A digital overlord That’s why we must solve humanoids at scale – the ultimate deployment vector for AGI”” / X https://x.com/adcock_brett/status/1926302507761820076

The X discussion about the Claude 4 system card is getting counterproductive It punishes Anthropic for actually releasing full safety tests and admitting to unusual behaviors. And I bet the behaviors of other models are really similar to Claude & now more labs will hide results. https://x.com/emollick/status/1926003595838619921

The methods we used to trace the thoughts of Claude are now open to the public! Today, we are releasing a library which lets anyone generate graphs which show the internal reasoning steps a model used to arrive at an answer. https://x.com/i/web/status/1928123130725421201

Fantastic to see Anthropic, in collaboration with @neuronpedia, creating open source tools for studying circuits with transcoders. There’s a lot of interesting work to be done I’m also very glad someone finally found a use for our Gemma Scope transcoders! Credit to @ArthurConmy”” / X https://x.com/NeelNanda5/status/1928169762263122072

Anthropic open-sourced their circuit tracing tools”” / X https://x.com/i/web/status/1928119741626962006

Inference providers aren’t sleeping on the switch. https://x.com/fdaudens/status/1927834963509961041

Love this approach by @RedHat_AI. We need more trust & validation in AI and this can help! https://x.com/ClementDelangue/status/1928551872027116000

Sigh, it’s a bit of a mess. Let me just give you guys the full nuance in one stream of consciousness since I think we’ll continue to get partial interpretations that confuse everyone. All the little things I post need to always be put together in one place. First, I have long”” / X https://x.com/i/web/status/1928148705145934252

Capgemini and SAP partner with Mistral to deploy AI for sensitive sectors | Reuters https://www.reuters.com/business/capgemini-sap-partner-with-mistral-deploy-ai-sensitive-sectors-2025-05-26/

How I Get Business Owners’ Real Phone Numbers: I got sick of cold calling gatekeepers so I built a system that finds local businesses by keyword – then pulls the OWNER’S name, email, cell number, and home address. It runs through LeadMagic, ChatGPT, n8n, and skip tracing – all https://x.com/alxberman/status/1920828349875974240

The main blockers for AI adoption, at a glance in this Economist article “”Welcome to the AI trough of disillusionment”” https://x.com/fdaudens/status/1925632043418894758

I am alarmed by the proposed cuts to U.S. funding for basic research, and the impact this would have for U.S. competitiveness in AI and other areas. Funding research that is openly shared benefits the whole world, but the nation it benefits most is the one where the research is”” / X https://x.com/i/web/status/1928099650269237359

Exclusive: Nvidia to launch cheaper Blackwell AI chip for China after US export curbs, sources say | Reuters https://www.reuters.com/world/china/nvidia-launch-cheaper-blackwell-ai-chip-china-after-us-export-curbs-sources-say-2025-05-24/

How I used o3 to find CVE-2025-37899, a remote zeroday vulnerability in the Linux kernel’s SMB implementation – Sean Heelan’s Blog https://sean.heelan.io/2025/05/22/how-i-used-o3-to-find-cve-2025-37899-a-remote-zeroday-vulnerability-in-the-linux-kernels-smb-implementation/

How Students Are Fending Off Accusations That They Used A.I. to Cheat – The New York Times https://www.nytimes.com/2025/05/17/style/ai-chatgpt-turnitin-students-cheating.html

MBZUAI Launches Institute of Foundation Models and Establishes Silicon Valley AI Lab https://www.prnewswire.com/news-releases/mbzuai-launches-institute-of-foundation-models-and-establishes-silicon-valley-ai-lab-302464305.html

Solve for X What do you do? A) ask for the specfic problem B) fuck off, I’m not doing any maths C) spend a minute solving an imaginary problem The correct answer seems to be C. Verified by AGI. https://x.com/scaling01/status/1927733065150775786

The “bigger is better” era of AI is ending. Business and policy leaders are facing a new reality: energy-hungry, compute-heavy models are costly, inefficient, and unsustainable. The next wave of AI will be defined by smarter, more efficient models that scale securely, lower https://x.com/cohere/status/1927775064721703258

We are 85 seconds away from AGI You can ask LLMs completely nonsensical questions and they will find an answer. The reasoning trace for this one is hilarious https://x.com/scaling01/status/1927725546282053902

My conclusion from the last couple of weeks of AI launches is that none of the AI companies can explain very well what their systems do. This is partially because they don’t always know & partially because there is no established approach for documentation of AI capabilities.”” / X https://x.com/emollick/status/1925715197676691615

How Stargate UAE outsizes the world’s largest data centres | The National https://www.thenationalnews.com/future/technology/2025/05/24/stargate-uae-ai-g42/

Meta understood that copying DeepSeek piecemeal is not working, and decided to copy the org structure, creating an internal AGI division. a cruel rhyme from Russian school program comes to mind “”And you, my friends, no matter your positions,         Will never be musicians!”” https://x.com/i/web/status/1927944123358581182

Canada now has a minister of artificial intelligence. What will he do? | CBC News https://www.cbc.ca/news/politics/artificial-intelligence-evan-solomon-1.7536218

China approaches AI “”like electricity, not nuclear weapons”” vs the US (per The Economist). Key difference: 🇺🇸 Focus on building models 🇨🇳 Focus on practical applications Worth a read. https://x.com/fdaudens/status/1927020700302184634

Chinese scientists develop AI model to predict stellar flares – Chinadaily.com.cn https://www.chinadaily.com.cn/a/202505/28/WS68366271a310a04af22c1e46.html

I built a browser-based video frame extractor that processes videos locally for privacy. Designed by @heybossAI Built by @lovable_dev Hosted on @Cloudflare https://x.com/SamuelSojin/status/1922530723182907604

When I realized how dangerous the current agency-driven AI trajectory could be for future generations, I knew I had to do all I could to make AI safer. I recently shared this personal experience, and outlined the scientific solution I envision @TEDTalks⤵️ https://x.com/Yoshua_Bengio/status/1927481467988287656

o3 for finding a security vulnerability in the Linux kernel: https://x.com/gdb/status/1926345848461447523

Nick Clegg says asking artists for use permission would ‘kill’ the AI industry | The Verge https://www.theverge.com/news/674366/nick-clegg-uk-ai-artists-policy-letter

An unauthorized update by an unnamed xAI employee caused Grok, the chatbot on X, to make false claims of a “white genocide” in South Africa, inserting the topic into unrelated conversations. xAI has since reversed the changes, tightened internal safeguards, and pledged greater https://x.com/DeepLearningAI/status/1927500443808075873

Exclusive: Musk’s DOGE expanding his Grok AI in US government, raising conflict concerns | Reuters https://archive.md/5v8CE

ChatGPT is becoming an increasingly important and useful part of people’s daily lives:”” / X https://x.com/gdb/status/1926714515468509419

Delaware attorney general reportedly hires a bank to evaluate OpenAI’s restructuring plan | TechCrunch https://techcrunch.com/2025/05/29/delaware-attorney-general-reportedly-hires-a-bank-to-evaluate-openais-restructuring-plan/

The Time Sam Altman Asked for a Countersurveillance Audit of OpenAI | WIRED https://archive.ph/jBXuC

Remember the Great Ghiblification? Turns out, this is all part of a grander plan by OpenAI to make them appear cool and to win market share. Search, DeepResearch, Agents and Personlization with System Prompts, Tasks and the coming model unification are also part of that plan. https://x.com/scaling01/status/1926801814973804712

Free AI for all? UAE becomes first to offer ChatGPT Plus to every resident and citizen – The Arabian Stories News https://www.thearabianstories.com/2025/05/25/free-ai-for-all-uae-becomes-first-to-offer-chatgpt-plus-to-every-resident-and-citizen/

Musk-Altman AI rivalry complicating Trump’s dealmaking in Middle East https://www.cnbc.com/2025/05/29/musk-altman-ai-rivalry-complicating-trumps-dealmaking-in-middle-east.html

How Peter Thiel’s Relationship With Eliezer Yudkowsky Launched the AI Revolution | WIRED https://www.wired.com/story/book-excerpt-the-optimist-open-ai-sam-altman/

The U.S. Secretary of Transportation, Sean Duffy, at Tesla Giga Texas today – discussing the future of autonomous transportation with Optimus robots in the background. https://x.com/TheHumanoidHub/status/1924990626485174721

Why We Think | Lil’Log https://lilianweng.github.io/posts/2025-05-01-thinking/

Trending

Discover more from Ethan B. Holland

Subscribe now to keep reading and get access to the full archive.

Continue reading