Image created with OpenAI GPT-Image-1. Image prompt: vintage Sly & the Family Stone album-cover style, close-up portrait on fiery orange backdrop, soft grain featuring primary-color Google “G” hologram; grainy retro print texture, vibrant 60s funk color palette, high-resolution
These wild street interviews have completely taken over my feed. AI video had hit a quality bar already, but Veo 3 with native audio output has unlocked whole new category of creators (and thus content). Ppl who would’ve never bothered duct taping multiple tools together can https://x.com/bilawalsidhu/status/1929568408820949350
Very impressed with Veo 3 and all the things people are finding on r/aivideo etc. Makes a big difference qualitatively when you add audio. There are a few macro aspects to video generation that may not be fully appreciated: 1. Video is the highest bandwidth input to brain. Not”” / X https://x.com/karpathy/status/1929634696474120576
Veo 3 is now the first model to top both the Image to Video and Text to Video leaderboards, outperforming Kling 2.0 and Runway Gen 4 to secure the #1 spot across both modalities! Veo 3 represents a significant leap in Image to Video generation, where Google’s previous Veo 2 had https://x.com/ArtificialAnlys/status/1928318831761707224
I got access to Gemini Diffusion for a few weeks. I believe this is one of the most important upcoming technologies in the AI space, it’s like a GPT-4 moment all over again. Once the model gets better, it will completely reshape AI experiences for us. Video is in real time. https://x.com/skirano/status/1930332481078616296
Google quietly released an app that lets you download and run AI models locally | TechCrunch https://techcrunch.com/2025/05/31/google-quietly-released-an-app-that-lets-you-download-and-run-ai-models-locally/
Lmao. What niche even is this — grassroots dirt track racing meets google maps nerds? Veo 3 videos are seriously ridiculous and fun. Turn audio on for max enjoyment. https://x.com/bilawalsidhu/status/1930405285253767296
Two brutalist office buildings, floating in stormy seas and brimming with antique brass cannons, fire paint at each other, staining their surfaces. https://x.com/emollick/status/1929727288901550523
An inexplicable failure of Microsoft & Google’s AI tools is that they have access to my email but won’t actually use their smarts to help me When I ask for “”urgent messages,”” Google just gives me unread ones and Microsoft literally searches for “”urgent”” Yet Claude does better. https://x.com/emollick/status/1929770530472616075
@karpathy Right now I flip between Claude 4 Opus (usually), sometimes Gemini 2.5 Pro for the coding drugery, both helpful once you know limits. I throw some more challenging design, algorithmic things to o3. It often impresses, but it’s a bit of an ass and too exhausting to argue with for”” / X https://x.com/i/web/status/1929602224218644985
@karpathy Daily driver these days is Gemini 2.5 Pro and sometimes Claude Sonnet 4 For simple brainstorming/ creative writing DeepSeek v3″” / X https://x.com/i/web/status/1929613466475659662
New native audio capabilities in Gemini 2.5 enable text-to-speech in over 24 languages. 🔊Voices are more natural and expressive, and you can seamlessly switch between languages. https://x.com/i/web/status/1929960513779204198
Amidst the massive demand for Gemini 2.5 and Veo 3 models, wanted to also give a big shout out to our world-class infrastructure, chip and SRE teams, who work tirelessly to keep our wonderful TPUs from melting, and without whose incredible work none of this would be possible. https://x.com/demishassabis/status/1928604371157233918
Here is my 2 hour long workshop i just finished at the @aiDotEngineer World’s fair. This is all you need to know to learn on how to use Gemini 2.5! It is beginner friendly from getting your first API key to multimodality, function calling and MCP. 🆓 Completely free – runs https://x.com/i/web/status/1930055992051675312
At this point, having an obvious wrong citations in your AI generated report is a sign that you did not know to hit the Deep Research button on ChatGPT, Gemini, or Claude. The end of the fake-citation problem is good in some ways, but removes a major tell of bad AI use.”” / X https://x.com/emollick/status/1928271628447662367
ChatGPT can now read your Google Drive and Dropbox | The Verge https://www.theverge.com/news/679580/chatgpt-google-drive-dropbox-meeting-notes
ChatGPT for utilizing the context in your Google Drive:”” / X https://x.com/gdb/status/1929965894425555270
Google Labs Portraits experiment features author Kim Scott https://blog.google/technology/google-labs/portraits/
AI Agents can now talk to any website directly. Microsoft’s NLWeb converts website data into APIs that AI agents can query as an MCP server. Works with OpenAI, DeepSeek, Gemini, Claude and other LLMs. 100% Opensource. https://x.com/Saboo_Shubham_/status/1927379307371864428
🔥 Gemini 2.5 with real-time, multilingual, emotion-aware audio dialogue and fully controllable TTS is unbelievably good. It can whisper, argue, and tell stories in your accent. → Native real-time audio dialog with high expressivity, fast response, and style control — users https://x.com/rohanpaul_ai/status/1930257704427143650
Notebooks can now be public in @NotebookLM. Just hit “share” and set the access to “anyone with the link.” Easy-peasy. Share ideas, study guides and team docs, and your peers can explore, ask questions, get instant summaries and audio overviews. https://x.com/i/web/status/1930005768587112755
At Google I/O 2025, Google announced updated versions of Gemini 2.5 Pro and Flash with audio capabilities, a preview of the Gemma 3n open models optimized for mobile, and Veo 3, a model that can generate 4K video with dialogue and other audio. Google Search will be more deeply https://x.com/i/web/status/1929734794033660139
On GPQA Diamond, a set of PhD-level multiple-choice science questions, DeepSeek-R1-0528 scores 76% (±2%), outperforming the previous R1’s 72% (±3%). This is generally competitive with other frontier models, but below Gemini 2.5 Pro’s 84% (±3%). https://x.com/EpochAIResearch/status/1928489527204589680
DeepSeek R1 05-28 LiveBench results: – 8th in the Overall ahead of o4-mini, Gemini 2.5 Flash Preview and Qwen3-235B-A22B (biggest competitors) – 1st on Data Analysis !!! – 3rd on Reasoning !! – 4th on Mathematics ! – 11th on Language – 20th on Instruction Following – 23rd on https://x.com/scaling01/status/1928173385399308639
sharing negative results >>> 20k google scholar citations”” / X https://x.com/i/web/status/1929449598524821509
Cloud Run GPUs are now generally available | Google Cloud Blog https://cloud.google.com/blog/products/serverless/cloud-run-gpus-are-now-generally-available
Differential privacy on trust graphs https://research.google/blog/differential-privacy-on-trust-graphs/
Gemini 2.5 Pro: Access Google’s latest preview AI model https://blog.google/products/gemini/gemini-2-5-pro-latest-preview/
1. Download Google AI Edge Gallery Access the official Google AI Edge Gallery GitHub repo. Go to the “”Releases”” section, then download and install the .apk file (Android). The iOS version is coming soon. Link: https://x.com/itsPaulAi/status/1927453538629624112
Fuck Yes! Serverless GPU for everyone! The Cloud Run just shipped Serverless GPU with no quota request required. 🤯 Deploy @GoogleDeepMind Gemma with a single command! – Pay-per-second GPU billing. – Scale to zero instances. – TTFT of 19 seconds for a Gemma3 as cold start. – No https://x.com/i/web/status/1929638760758874428
Google released an app that allows you to run LLMs from Hugging Face, fully privately and 100% local 🔥 > Generate code on-the-fly > Chat with images > Supports multi-turn conversations > Choose any model from Hugging Face > Based on LiteRT 🔥 > Sign in with HF Support for iOS https://x.com/reach_vb/status/1929450131843137691
NotebookLM introduces public notebooks for sharing https://blog.google/technology/google-labs/notebooklm-public-notebooks/
Last week, @Google dropped a paper on ATLAS, a new architecture that reimagines how models learn and use memory. Unfortunately, it flew under everyone’s radar – but it shouldn’t have! So what’s Atlas bringing to the table? ▪️ Active memory via Google’s so-called Omega rule. It https://x.com/i/web/status/1929992259019432115
Create videos with Veo 3, now available in 73 countries right in the Gemini app 🎬 https://x.com/Google/status/1928573869893230705
Veo 3 is really fun to use for historical what-ifs. I put together a 1940s video newsreel as if Project Habakkuk, the World War Two British plan to build a giant aircraft carrier out of pykrete, a mix of ice and woodpulp, had actually happened. https://x.com/emollick/status/1929386464917307514




