Image created with gemini-2.5-flash-image with claude-sonnet-4-5. Image prompt: Cinematic wide shot of Elphaba soaring above Emerald City during Defying Gravity, her body artistically segmented into glowing modal zones showing text overlays on her black cloak, golden crystallized sound waves emanating from her mouth, violet depth-map skeletal structure visible through her green skin, and luminous wind currents around her outstretched hands, dramatic theatrical lighting from below, moody Wicked color palette with segmentation lines glowing like magical seams, large movie title overlay reading MULTIMODALITY
For AI to be able to help humans in the physical world, we need systems that can understand and simulate the universe. To exponentially accelerate Luma’s path to Multimodal AGI we are building a 2GW compute cluster with Humain and we have raised a $900M Series C. I am incredibly”” / X https://x.com/gravicle/status/1991202746871988680
Meta just dropped SAM 3D, but more interestingly, they basically cracked the 3D data bottleneck that’s been holding the field back for years. Manually creating or scanning 3D ground truth for the messy real world is basically impossible at scale. But what if you just have https://x.com/bilawalsidhu/status/1991237143898017854
Introducing SAM 3D: Powerful 3D Reconstruction for Physical World Images https://ai.meta.com/blog/sam-3d/
SAM 3D enables accurate 3D reconstruction from a single image, supporting real-world applications in editing, robotics, and interactive scene generation. Matt, a SAM 3D researcher, explains how the two-model design makes this possible for both people and complex environments. https://x.com/AIatMeta/status/1991605451809513685
Introducing SAM 3D, the newest addition to the SAM collection, bringing common sense 3D understanding of everyday images. SAM 3D includes two models: 🛋️ SAM 3D Objects for object and scene reconstruction 🧑🤝🧑 SAM 3D Body for human pose and shape estimation Both models achieve https://x.com/AIatMeta/status/1991184188402237877
We’re sharing model checkpoints, an evaluation benchmark, human body training data, and inference code with the community to support creative applications in fields like robotics, interactive media, science, sports medicine, and beyond. 🔗 SAM 3D Body: https://x.com/AIatMeta/status/1991184190323212661
Meta AI Demos https://aidemos.meta.com/segment-anything
Introducing Meta Segment Anything Model 3 and Segment Anything Playground https://ai.meta.com/blog/segment-anything-model-3/
SAM-3 is out on @huggingface! A big upgrade from SAM-2, and Meta finally added support for text prompts. Here I tried it out on @hazardeden10’s magical goal against @Arsenal using the text prompt “”Chelsea player”” Works pretty well! https://x.com/NielsRogge/status/1991213874687758799
Collecting a high quality dataset with 4M unique phrases and 52M corresponding object masks helped SAM 3 achieve 2x the performance of baseline models. Kate, a researcher on SAM 3, explains how the data engine made this leap possible. 🔗 Read the SAM 3 research paper: https://x.com/AIatMeta/status/1991640180185317644
SAM3 video tracking is so good yesterday: collect data, train custom object detector, use tracker to estimate object motion – days today: track anything with text prompt – seconds https://x.com/skalskip92/status/1991232397686219032
We’ve partnered with @Roboflow to enable people to annotate data, fine-tune, and deploy SAM 3 for their particular needs. Try it here: https://x.com/AIatMeta/status/1991191530367799379
SAM 3 tackles a challenging problem in vision: unifying a model architecture for detection and tracking. Christoph, a researcher on SAM 3, shares how the team made it possible. 🔗 Read the SAM 3 research paper: https://x.com/AIatMeta/status/1991538570402934980
SAM3 is open-source model. You can use the models in commercial. You can modify or fine tune. You keep ownership of your modifications. You do not need to release your source code.”” / X https://x.com/skalskip92/status/1991626755782877234
Today we are releasing & open-sourcing Segment Anything 3 (SAM 3). It is a state-of-the-art model for image & video segmentation, and builds upon the work of SAM & SAM 2. SAM3 will also power features in Edits, Meta AI, & Facebook Marketplace soon. https://x.com/alexandr_wang/status/1991198465628459494
Today we’re excited to unveil a new generation of Segment Anything Models: 1️⃣ SAM 3 enables detecting, segmenting and tracking of objects across images and videos, now with short text phrases and exemplar prompts. 🔗 Learn more about SAM 3: https://x.com/AIatMeta/status/1991178519557046380
We release following ⤵️ > video segmentation demo with visual/concept prompting ⏯️ https://x.com/mervenoyann/status/1991182168161136684
AI can now create AND explore 3D worlds. World models and agentic AI are on a collision course. World Labs is making world-building effortless. Google DeepMind’s SIMA-2 is making agency inside those worlds possible. Together, they hint at a new paradigm–AI that both creates https://x.com/bilawalsidhu/status/1990994808626950579
Google DeepMind has introduced SIMA 2, a reasoning, conversational AI agent for 3D worlds including games and generative world-model scenes. – Handles complex goals, explains steps, supports multilingual/emojis for collaborative play. – Adapts to real-time generated 3D worlds https://x.com/TheHumanoidHub/status/1989424462085960082
Google DeepMind’s SIMA 1 vs SIMA 2 The bitter lesson continues to be bitter sweet https://x.com/bilawalsidhu/status/1989001120849735898
I had access to Gemini 3. It is a very good, very fast model. It also demonstrates the change from chatbot to agent. https://x.com/emollick/status/1990827310082330971
Gemini 3 has impressive benchmarks for building agents (Vending-Bench 2, Terminal Bench, and Sierra). So we tested its performance as a research agent using Deep Agents. We found Gemini 3 is very effective at using research tools like file manipulation, planning, and subagent”” / X https://x.com/LangChainAI/status/1991220334578848209
Google Antigravity is an ‘agent-first’ coding tool built for Gemini 3 | The Verge https://www.theverge.com/news/822833/google-antigravity-ide-coding-agent-gemini-3-pro
Agentic coding today: Gemini 3 spent a few minutes correctly diagnosing the issue. Then, across several rounds of a few minutes of work, failed to actually fix it Then, GPT 5.1 codex max was able to work for about 15 minutes and solve the problem but introduced a small bug.”” / X https://x.com/kylebrussell/status/1991247685672923302
Google @Antigravity is a new agentic platform designed to autonomously plan and execute complex software development tasks. – Access Gemini 3 Pro Preview and other models directly. – Distinct Editor and Agent Manager for synchronous and asynchronous workflows. – Browser Subagent https://x.com/_philschmid/status/1990816850792337454
State of the art reasoning right within an autonomous agent. Complex instruction following and advanced coding capabilities from Gemini 3 Pro helps Jules complete more complex tasks in parallel. Available now for Ultra, Pro coming real soon. 2.5 Pro available to everyone. https://x.com/julesagent/status/1991207201487352222
Google Antigravity is our new agentic development platform. It helps developers build faster by collaborating with AI agents that can autonomously operate across the editor, terminal, and browser. It uses Gemini 3 Pro 🧠 to reason about problems, Gemini 2.5 Computer Use 💻 for https://x.com/GoogleDeepMind/status/1990827890435346787
Introducing Gemini 3 ✨ It’s the best model in the world for multimodal understanding, and our most powerful agentic + vibe coding model yet. Gemini 3 can bring any idea to life, quickly grasping context and intent so you can get what you need with less prompting. Find Gemini https://x.com/sundarpichai/status/1990812770762215649
Gemini 3 Pro takes first place on Stagehands agentic browsing benchmark https://x.com/scaling01/status/1990872758872387939
🚀 Deep Agents: The Weekly Roundup 🚀 We’ve shipped new resources to help you build Deep Agents capable of handling complex, long-running tasks. 1/🥉 Build a Research Agent with Gemini 3 – We tested Gemini 3’s impressive benchmarks in practice using Deep Agents. We found Gemini https://x.com/LangChainAI/status/1991928474404311493
New Gemini 3 reasoning and tool use capabilities are a big step forward for agents 👀 – thinking_level to set reasoning for task requirements – thought signatures for stateful tool use – larger context window for managing drift on complex tasks LangGraph, LangChain, and Deep”” / X https://x.com/LangChainAI/status/1991222443298660722
ollama run gemini-3-pro-preview 🧠 State-of-the-art reasoning 🖼️ Deep multimodal understanding 💻 Powerful vibe coding so you can go from prompt to app in one shot ⭐ Improved agentic capabilities, so it can get things done on your behalf, at your direction Gemini 3 Pro is https://x.com/ollama/status/1990839646876553543
Meet Google Antigravity, your new agentic development platform. An evolution of the IDE, it’s built to help you: – Orchestrate agents operating at a higher, task-oriented level – Run parallel tasks with agents across workspaces – Build anything with Gemini 3 Pro. https://x.com/antigravity/status/1990813606217236828
The @GoogleDeepMind team just dropped Gemini 3, and we at LlamaIndex have day-zero support! We also made a little demo to show how you can leverage the advanced agentic capabilities and structured output accuracy of Gemini 3 to automate your GitHub workflow around PRs, you just https://x.com/llama_index/status/1990902918388855185
Hot off the presses is Gemini 3 Pro, Google’s new SOTA model – tops LMArena with a score of 1501 points. Launching simultaneously as an API, inside the consumer Gemini app, Google Search, oh and a new agentic IDE. Here’s the TL;DR: 1. Gemini 3 Pro: new SOTA for multimodality https://x.com/bilawalsidhu/status/1990812584019439988
Gemini 3 Pro sets new record on SWE-bench verified: 74%! (evaluated with minimal agent) Costs are 1.6x of GPT-5, but still cheaper than Sonnet 4.5. Gemini iterates longer than everyone; run your agent with a step limit of >100 for max performance. Details & full agent logs in 🧵 https://x.com/KLieret/status/1991164693839270372
Introducing Gemini 3, the best model in the world for multimodal understanding and our most powerful agentic and vibe-coding model yet. Gemini 3 Pro tops the LMArena Leaderboard at 1501 Elo! Available now in Gemini Enterprise & Vertex AI → https://x.com/GoogleCloudTech/status/1990813342189887831
This is Gemini 3: our most intelligent model that helps you learn, build and plan anything. It comes with state-of-the-art reasoning capabilities, world-leading multimodal understanding, and enables new agentic coding experiences. 🧵 https://x.com/GoogleDeepMind/status/1990812966074376261
On an apples:apples harness comparison (mini-swe-agent: bash-only, same prompts), Gemini 3 Pro sets a new sota on SWE-bench verified: 74.20! https://x.com/ankesh_anand/status/1991199945798365384
Gemini 3 models from @Google @GoogleDeepMind have made a significant 2X SOTA jump on ARC-AGI-2 (Semi-Private Eval) Gemini 3 Pro: 31.11%, $0.81/task Gemini 3 Deep Think (Preview): 45.14%, $77.16/task https://x.com/arcprize/status/1990820655411909018
#Gemini3 is finally out! Congrats to everyone on this amazing launch! Also very excited to see how #DeepThink can power Gemini3 to further the state-of-the-art performances across reasoning, deep knowledge, and multimodality: 41% HLE, 93.8% GPQA, & 45.1% on ARC-AGI-2 (big jump)! https://x.com/lmthang/status/1990816762300960954
Introducing Gemini 3 — our most intelligent model that helps you bring any idea to life. Gemini 3 is our next step on the path toward AGI and has: 🧠 State-of-the-art reasoning 🖼️ Deep multimodal understanding 💻 Powerful vibe coding so you can go from prompt to app in one shot https://x.com/Google/status/1990813116045602942
Gemini 3 scores 31.1% on ARC-AGI-2. Impressive progress.”” / X https://x.com/fchollet/status/1990813908483928178
Google to enable research automation on Gemini Enterprise https://www.testingcatalog.com/google-to-enable-research-automation-on-gemini-enterprise/
Gemini 3: Introducing the latest Gemini AI model from Google https://blog.google/products/gemini/gemini-3/#responsible-development
Gemini 3 Pro is the new leader in AI. Google has the leading language model for the first time, with Gemini 3 Pro debuting +3 points above GPT-5.1 in our Artificial Analysis Intelligence Index @GoogleDeepMind gave us pre-release access to Gemini 3 Pro Preview. The model https://x.com/ArtificialAnlys/status/1990813106478715098
My Gemini 3 Review — matt shumer https://shumer.dev/gemini3review
Gemini 3 Pro is rolling out to @code developers! https://x.com/pierceboggan/status/1990817374799528259
This is Gemini 3 ⚡ https://x.com/Google/status/1991196250499133809
Gemini 3 Pro has around ~7.5T params (vibe-mathing with explanation) > the naive fit with with an R^2 of 0.8816 yields a mean estimation of 2.325 Quadrillion parameters > ummm, that’s not it > let’s only take sparse MoE reasoning models > this includes gpt-oss-20B and 120B, https://x.com/scaling01/status/1990967279282987068
Google Has Your Data. Gemini Barely Uses It. | Shlok Khemani https://www.shloked.com/writing/gemini-memory
After lots of testing, Gemini 3 is a mixed bag – but a useful addition. Compared to 2.5 Pro, its: 1. Worse at transcription and diarization: Adds words that weren’t said, projects emotion, almost like it’s too smart. 2. Not as good at translation or writing: Baseline”” / X https://x.com/hrishioa/status/1991691037035884754
Google AI Studio https://aistudio.google.com/apps/bundled/info_genius?show=&showPreview=true&showAssistant=true
One of the most striking thing about gemini-3-pro is how much better it is with several iterations. It makes better use of the information from the previous iterations than other models. After one iteration is is barely better than gpt-5.1, while after 5 it is almost 10pp ahead. https://x.com/htihle/status/1991137526480810470
BREAKING: Gemini 3 Pro is out!! It’s state of the art (or close) on coding, reasoning, computer use and more. It’s also extremely fast. We’ve been testing it internally @every for a few hours. Here’s what we’ve noticed so far: – Coding. It rips in @FactoryAI’s Droid. So fast https://x.com/danshipper/status/1990812588511567898
📍GeoGuessr isn’t just a game; it’s a massive test of complicated visual reasoning and world knowledge. Very satisfied to see my efforts helped Gemini pass this test and beat human pros for the first time! Still a long way to go, but a dream milestone just unlocked 🔓”” / X https://x.com/songyoupeng/status/1991214812316201131
Gemini 3 Pro is the first LLM to beat professional human players at GeoGuessr https://x.com/scaling01/status/1990904842488066518
We wrote a Gemini 3 Developer Guide including all new API features, Migration strategies, and technical details for building with Gemini 3 Pro preview: – Control reasoning via `thinking_level` low and high modes. – per part `media_resolution` for better multimodal reasoning – https://x.com/_philschmid/status/1990836465647984969
Gemini 3 Pro Preview now on aistudio Pricing: <=200K tokens • Input: $2.00 / Output: $12.00 > 200K tokens • Input: $4.00 / Output: $18.00″” / X https://x.com/scaling01/status/1990797742629925073
The model also shows increased resistance to prompt injections and improved protection against cyberattacks. As we continue to advance AI, we are relentlessly focused on ensuring this transformative technology benefits humanity while minimizing potential harms. See our Gemini 3″” / X https://x.com/GoogleDeepMind/status/1991118579119304990
Gemini 3 is here, and it’s built to be our most secure model yet. 🔒 ✅The most comprehensive safety evaluations of any Google AI model to date ✅Rigorous testing against our Frontier Safety Framework ✅Independent assessment by external industry experts https://x.com/GoogleDeepMind/status/1991118575554408556
From Google’s Frontier Safety Report on Gemini 3 Pro: – clear improvements on all CBRN benchmarks especially in LabBench, a benchmark designed that measures performance on practical tasks required for scientific research in biology – on the hardest subset of their https://x.com/scaling01/status/1991177438789857661
The secret behind Gemini 3? Simple: Improving pre-training & post-training 🤯 Pre-training: Contra the popular belief that scaling is over–which we discussed in our NeurIPS ’25 talk with @ilyasut and @quocleix–the team delivered a drastic jump. The delta between 2.5 and 3.0 is https://x.com/OriolVinyalsML/status/1990854455802343680
One great thing about AntiGravity IDE is its agentic Chrome integration It doesn’t just build the frontend, it drives the UI, pokes the controls and then auto tests fixes in the same loop Cursor has this as well but not nearly as smooth. Playwright MCP is too slow in”” / X https://x.com/cto_junior/status/1990965505243689094
🚨BREAKING: @GoogleDeepMind’s Gemini-3-Pro is now #1 across all major Arena leaderboards 🥇#1 in Text, Vision, and WebDev – surpassing Grok-4.1, Claude-4.5, and GPT-5 🥇#1 in Coding, Math, Creative Writing, Long Queries, and nearly all occupational leaderboards. Massive gains https://x.com/arena/status/1990813759938703570
This is the biggest performance delta we’ve seen since launching Design Arena Gemini 3.0 Pro has taken #1 overall and #1 in 4 of our 5 code arenas – Website, Game Dev, 3D Design, and UI Components Well-earned congratulations to the @GoogleDeepMind team on a remarkable https://x.com/grx_xce/status/1990815340893245481
Students in the US (and many other countries) can get their hands on all the Gemini 3 Pro goodness for free!”” / X https://x.com/demishassabis/status/1990993251247997381
Gemini 3 Pro just took the #1 spot in our new AA-Omniscience Index — but it is a nuanced story AA-Omniscience is our new knowledge and hallucination eval. Gemini 3 Pro’s leadership is driven by its high Accuracy (percentage correct); the model scored a massive 14 points higher https://x.com/ArtificialAnlys/status/1990926803087892506
The Scaling Wall Was A Mirage | Tomasz Tunguz https://tomtunguz.com/gemini-3-proves-pretraining-scaling-laws-intact/
Gemini 3.0 is the next-generation frontier model on our LiveCodeBench Pro benchmark, better than GPT-5/5.1. We’re very excited that Google has adopted our benchmark: a continuously updated collection of problems from Codeforces, ICPC, and IOI designed specifically to minimize https://x.com/wenhaocha1/status/1990818535640088585
Just how significant is the jump with Gemini 3? We just released a new leaderboard to track AI developments. Gemini 3 is the largest leap in a long time. https://x.com/hendrycks/status/1991188096302338491
I played with Gemini 3 yesterday via early access. Few thoughts – First I usually urge caution with public benchmarks because imo they can be quite possible to game. It comes down to discipline and self-restraint of the team (who is meanwhile strongly incentivized otherwise) to”” / X https://x.com/karpathy/status/1990854771058913347
Gemini 3 Pro is live in Cline! 1M token context window & a new SOTA on benchmarks. https://x.com/cline/status/1990820473555595389
Gemini 3 Pro is now available in Windsurf”” / X https://x.com/cognition/status/1990856307616985163
Gemini 3 Pro set a new record on FrontierMath: 38% on Tiers 1-3 and 19% on Tier 4. On the Epoch Capabilities Index (ECI), which combines multiple benchmarks, Gemini 3 Pro scored 154, up from GPT-5.1’s previous high score of 151. https://x.com/EpochAIResearch/status/1991945942174761050
This is cope. Gemini 3’s behavioral problems are structurally the same as in Gemini 2.5 and earlier. To the extent that it does better, it’s just overcoming its biases with raw horsepower. Gemini post-training is malign since the very first experiments. Cursed bloodline. https://x.com/teortaxesTex/status/1991086733962715540
Introducing Gemini 3 Pro, the world’s most intelligent model that can help you being anything to life. It is state of the art across most benchmarks, but really comes to life across our products (AI Studio, the Gemini API, Gemini App, etc) 🤯 https://x.com/OfficialLoganK/status/1990813077172822143
Gemini 3 is now in AI Mode — making it even easier to ask anything in Search. Here’s more on this update from @rmstein, VP of Product for Search.”” / X https://x.com/Google/status/1991212868620951747
Gemini 3 Pro takes the crown on Scale AI’s VisualToolBench https://x.com/scaling01/status/1991932333147213834
Gemini 3 Pro is the best multimodal model ever. You can now turn a single picture into an almost pixel-perfect website. It’s honestly incredible. And it’s now the default model in @MagicPathAI https://x.com/skirano/status/1991175569388494972
Crushing superiority of Gemini-3-pro-“”””””preview”””””” on WeirdML. The gap between 5.1(high) and Gemini is equal to one between o1(high) and o3(high). A generation’s worth of advantage. https://x.com/teortaxesTex/status/1991156784719888588
Google just released Gemini 3, its most intelligent AI model yet. I caught up with Demis Hassabis, CEO of Google DeepMind, to ask about: -Gemini 3 and Google’s AI strategy -Google’s new Antigravity tool -Medical-grade AI Here’s what he said: https://x.com/rowancheung/status/1990814463428059597
Gemini 3 Pro is now available in Cursor!”” / X https://x.com/cursor_ai/status/1990814174264381910
Gemini 3 Pro with the largest delta recorded thus far on @Designarena 🤯 https://x.com/OfficialLoganK/status/1990826955730489733
One of the early Gemini 3 tests I did was take the bouncing ball example and try to make it 10x harder, Gemini 3 Pro crush it in 1 shot… (not best of N, literally first prompt made this) https://x.com/OfficialLoganK/status/1990819310072443340
Google just released Gemini 3, its most intelligent AI model yet. I caught up with Demis Hassabis, CEO of Google DeepMind, to ask about: -Gemini 3 and Google’s AI strategy -Google’s new Antigravity tool -Medical-grade AI Here’s what he said: https://x.com/rowancheung/status/1990814463428059597?s=20
Gemini 3 Pro is now available on OpenRouter https://x.com/scaling01/status/1990817957497155848
Amp’s new default model: Gemini 3 Pro https://x.com/thorstenball/status/1990821112750481744
Hey, Gemini 3, So I need DOOM, but more root vegetables, also no guns or demons or mars. And more of a focus on different flooring styles. but otherwise EXACTLY the same as DOOM.”” Gemini: “”Here is F.L.O.O.R. (First-person Lino Observation & Ornamental Review).”” Pretty good! https://x.com/emollick/status/1991249261816594896
Gemini is such a weird model – I find it too jagged and unreliable at instruction following to switch to it but where it’s i guess been RL’ed it seems to be big leaps..”” / X https://x.com/Teknium/status/1991815251084628196
We’ve been intensely cooking Gemini 3 for a while now, and we’re so excited and proud to share the results with you all. Of course it tops the leaderboards, including @arena, HLE, GPQA etc, but beyond the benchmarks it’s been by far my favourite model to use for its style and https://x.com/demishassabis/status/1990818891392496005
Gemini 3 Pro has taken the #1 spot on Dubesor Bench https://x.com/scaling01/status/1991931844347207887
Congrats to Google on Gemini 3! Looks like a great model.”” / X https://x.com/sama/status/1990828659981144462
Look what we have been cooking for you #Gemini3 ! ✨ Beyond other capabilities, Gemini 3’s spatial understanding and world knowledge are also truly next-level! Incredible to see the progress, and proud to have helped chart some of those new territories!!🚀”” / X https://x.com/songyoupeng/status/1990835604767322523
Ohh no, this is so much worse Maybe the antigravity IDE doesn’t have really good style guidelines for Gemini 3 Should give it a try in cursor https://x.com/cto_junior/status/1990966750746484920
Gemini 3 Pro is still undefeated on the Snake Arena https://x.com/scaling01/status/1991932651968852333
gemini 3 pro • our most intelligent model yet • SOTA reasoning • 1501 Elo on LMArena • next-level vibe coding capabilities • complex multimodal understanding available now in Google AI Studio and the Gemini API https://x.com/GoogleAIStudio/status/1990813281414455385
Gemini 3 Pro on the new Vending-Bench Arena 🤯 tool calling is impressive with this model. https://x.com/OfficialLoganK/status/1990833534672797703
At Box, we’ve been testing Gemini 3 Pro in early access with Box AI on our most complex advanced reasoning eval, and Gemini Pro was a massive 22 percentage point improvement over Gemini 2.5 Pro. For this test, we ask the model a series of complex, real-world questions with a set https://x.com/levie/status/1990820579981840746
And say hello to Gemini 3 Deep Think, even more SOTA compared to Gemini 3 Pro 🤯 https://x.com/OfficialLoganK/status/1990814722250146277
Gemini 3 Pro #1 on PMPP-Eval PMPP = Programming Massively Parallel Processors aka coding with CUDA”” / X https://x.com/scaling01/status/1990920793887273396
Gemini 3 is now available in the @GeminiApp. ⚡ Starting today, you’ll be able to: 🧠 Get more helpful, concise responses with easier-to-read formatting. 🧪 Try our new experiments, visual layout and dynamic view, that use Gemini 3 capabilities to make your responses more visual https://x.com/Google/status/1990829896562548855
Gemini 3 Prompting: Best Practices for General Usage https://www.philschmid.de/gemini-3-prompt-practices
Gemini 3: Introducing the latest Gemini AI model from Google https://blog.google/products/gemini/gemini-3/
Gemini 3 Pro (preview) scores 91% on VPCT (spatial reasoning) Uhhhh jesus christ https://x.com/ChaseBrowe32432/status/1990810992931135909
Google Antigravity Blog: introducing-google-antigravity https://antigravity.google/blog/introducing-google-antigravity
nano banana pro (gemini 3 pro image) • SOTA text rendering & localization • granular physics & lighting control • up to 4k studio-quality output • precise character consistency now available in preview on the Gemini API and in Google AI Studio with paid API key https://x.com/GoogleAIStudio/status/1991537543989588445
Nano Banana Pro is taking off. Here are some standout examples from the community so far 🧵”” / X https://x.com/GeminiApp/status/1991570302720163988
Introducing Nano Banana Pro (Gemini 3 Pro Image), Google DeepMind’s most advanced image generation and editing model. Now available on Together AI for production-scale visual content creation with reliable inference. https://x.com/togethercompute/status/1991614379394203973
If you see an image and want to confirm it has been made with Google AI, upload it to the Gemini app and ask a question like “”Was this generated with Google AI?”” Gemini will check for the SynthID watermark and use its own reasoning to return a response that helps you quickly make”” / X https://x.com/Google/status/1991552945754612118
The Gemini app gets new image verification features https://blog.google/technology/ai/ai-image-verification-gemini-app/
We’re launching generative UI features in the @GeminiApp and Google Search, starting with AI Mode, to make information more accessible in new ways. Here’s what to know: – Our generative UI dynamically creates visual layouts and interactive interfaces — such as webpages, games,”” / X https://x.com/Google/status/1991270067934216372
Gemini 3 Pro Preview has comparable speeds to Gemini 2.5 Pro, with 128 output tokens per second. This places it ahead of other frontier models including GPT-5.1 (high), Kimi K2 Thinking and Grok 4 https://x.com/ArtificialAnlys/status/1990813128226189811
OpenAI can’t beat Google in consumer AI – by John Hwang https://nextword.substack.com/p/openai-cant-beat-google-in-consumer
If Google really wanted to accelerate science, it should make Deep Research (and Gemini in general) have better retrieval from Google Scholar and Google Books. These are unique repositories that contain a remarkable amount of the world’s academic knowledge in hard-to-access form.”” / X https://x.com/emollick/status/1989755741549597039
Try this: have Nano Banana Pro search for you online, then ask it to create what your Instagram profile would look like. It’s a surprisingly good way to visualize your online persona. https://x.com/skirano/status/1991921872330735982
Senior Director of Product Management for Gemini @tulseedoshi breaks down the latest on Gemini 3 and Nano Banana Pro ⬇️ https://x.com/Google/status/1991652494032732443
Fun little Gemini 3 experiment where I asked it “”build me a time machine simulator, make it very very good”” and then “”make it better”” a few times. I like that it added calls to Gemini within the application, including adding speech & nano banana images. https://x.com/emollick/status/1990904243239473351
btw if you want extra precision in your editing with Nano Banana Pro and you are an Ultra subscriber, you can use it in the Flow app https://x.com/demishassabis/status/1991662935983419424
Nano Banana PRO is live in LTX. We took it for a test drive, and the results are wild. You’re going to want to save this one… Here’s what’s new 🧵 https://x.com/LTXStudio/status/1991943188379250933
🍌⚡ We put Gemini 2.5 Flash Image “Nano Banana” vs. Gemini 3 Pro Image “Nano Banana Pro” head-to-head… Same prompt. Two different outcomes. Here’s what @GoogleDeepMind shared is new: 🔶 Crisp, clearer text 🔶 4K-ready visuals 🔶 Stronger Gemini 3 reasoning 🔶 Adjustable https://x.com/arena/status/1991652781879620088
Over the past ~8 hours, @yupp_ai users from around the world have been going 🍌🍌for the new Google Nano Banana Pro model – it sits atop our Image leaderboard by a wide margin! Congrats @sundarpichai and @Google for building on Gemini 3.0 to produce the world’s best image model! https://x.com/lintool/status/1991693200822768033
From Gemini 3 to Nano Banana Pro & more, the team has been shipping. Here’s a look at the latest Drops 🧵1/10″” / X https://x.com/GeminiApp/status/1991953958257205641
Google to release Nano Banana Pro next week https://www.testingcatalog.com/google-to-release-nano-banana-pro-powered-by-gemini-3-pro-next-week/
Nano Banana Pro image generation in Gemini: Prompt tips https://blog.google/products/gemini/prompting-tips-nano-banana-pro/
Gemini 3 Pro Image (Nano Banana Pro) – Google DeepMind https://deepmind.google/models/gemini-image/pro/
Nano Banana Pro is wild. I just built a little app in Google AI Studio to help build intuition around AI papers. Paper reading is more fun than ever. 🙂 Images generated by Nano Banana Pro. Gemini 3 + Nano Banana Pro is an insane combo. https://x.com/omarsar0/status/1991657126188773878
Developers can build with Nano Banana Pro (Gemini 3 Pro Image) https://blog.google/technology/developers/gemini-3-pro-image-developers/
These major improvements in accuracy of rendered text are part of why the Nano Banana Pro model is such an upgrade over our earlier Nano Banana model (e.g. error rate goes from 56% for Nano Banana, aka Gemini 2.5 Flash Image, to 8% for Nano Banana Pro, aka Gemini 3 Pro Image).”” / X https://x.com/JeffDean/status/1991573065994744091
Nano Banana Pro aka gemini-3-pro-image-preview is the best available image generation model https://simonwillison.net/2025/Nov/20/nano-banana-pro/
🚨🍌BREAKING: @GoogleDeepMind’s Gemini 3 Pro Image aka Nano Banana Pro is in the Arena! Built on Gemini 3, which only two days ago landed as #1 across all major Arena leaderboards. Put it head-to-head in Battle mode with the latest models and judge for yourself if it’s SOTA for https://x.com/arena/status/1991540746114199960
Gemini 3 Pro Image vs GPT-Image 1 https://x.com/scaling01/status/1991546597013160290
Nano Banana Pro: Gemini 3 Pro Image model from Google DeepMind https://blog.google/technology/ai/nano-banana-pro/
Real world users on @yupp_ai prefer Google Nano Banana Pro 🍌🍌an incredible 80+% of the time when compared to competitor models for everyday use cases! https://x.com/lintool/status/1991693562820587926
Starting today for Google AI Ultra subscribers, creating with Nano Banana Pro in Flow means mastering the elements and the lens with precision and control. Watch @sanchitsawaria break down how to transform a single static frame into a cinematic shot: ✅ Change focus to guide the https://x.com/FlowbyGoogle/status/1991620311637283138
Nano Banana Pro (Gemini 3 Pro Image) now available in @GoogleAIStudio and Gemini API 🍌🍌🍌 The model “thinks”” through a prompt and can retrieve real-time data, such as weather forecasts or stock charts, using Google Search grounding before generating high-fidelity images https://x.com/_philschmid/status/1991537712420020225
Nano Banana Pro is great at making paper illustrations Here is Attention is All You Need https://x.com/osanseviero/status/1991804629554995247
Nano Banana Pro marks a significant jump in accuracy of rendered text within images across many languages. https://x.com/19kaushiks/status/1991535638676664399
A powerful way to use Nano Banana Pro in @FlowbyGoogle Step 1: Upload an image or generate an image using Imagen or Nano Banana https://x.com/nmatares/status/1991696375403409765
Nano Banana Pro🍌is a bigger milestone than it seems. Watch how it can generate high-fidelity annotated figures and equations from papers. And you can iterate on images using chat! 🤯 Watch until the end. If enough interest, I will try to release the app over the weekend. https://x.com/omarsar0/status/1991911424868970662
Nano Banana Pro, released this morning, is clearly the best image generation model. Superb instruction following, plus it can generate full infographics (with correct spelling and properly rendered text!) from a short prompt based on running extra searches https://x.com/simonw/status/1991545654901133797
Gemini 3 is coming to Google Search, starting with AI Mode. ⚡ Here’s what to know: 🏆 This marks the first time we’ve brought a Gemini model to Search on day one. 🔎 Our newest model brings incredible reasoning power to Search because it’s built to grasp unprecedented depth https://x.com/Google/status/1990845314551447838
The most crushing defeat for OpenAI I did not expect Gemini 3 Pro to be SOTA on WeirdML WeirdML has been an OpenAI stronghold for quite some time. https://x.com/scaling01/status/1991154001283358992
Gemini 3 Pro takes the crown on LisanBench – it scores 2.2x higher than GPT-5 while using 2.4x fewer reasoning tokens – it has the highest score on 23 out of 50 words – Grok-4 is the only model that can keep up https://x.com/scaling01/status/1990845163652993166
The Artificial Analysis leaderboard shows Gemini 3 at 73%, GPT-5.1 at 70%, and Kimi at 67% – minor differences. On our leaderboard, Gemini is 47%, GPT-5.1 is 38%, and Kimi is 27% – Gemini 3 is substantially more capable on hard benchmarks. https://x.com/hendrycks/status/1991188104804208736
Gemini Robotics 1.5 features a separate reasoning engine (ER), but its VLA model is also capable of thinking due to interleaved reasoning tokens. The VLA is able to independently operate long autonomous sequences (15+ minutes) without aid from the ER/VLM. https://x.com/TheHumanoidHub/status/1989393094631199088
Chase Brower on X: “Gemini 3 Pro (preview) scores 91% on VPCT (spatial reasoning) Uhhhh jesus christ https://t.co/fbyTHE47E1″ / X
https://x.com/ChaseBrowe32432/status/1990810992931135909
NVIDIA just released Nemotron Parse on Hugging Face A new vision model that goes beyond traditional OCR to understand complex document layouts. It extracts text, tables, and other elements with spatial grounding, turning unstructured documents into actionable data.”” / X https://x.com/HuggingPapers/status/1991108589235372286
If you work with robotics, AV, or 3D vision, this update will save you months of engineering. Most models need complex engineering to get reliable 3D geometry. This one does it with a plain transformer. Depth Anything 3 is the new model from @BytedanceTalk that predicts stable, https://x.com/IlirAliu_/status/1989622721366446190
Depth Anything 3 proves most 3D vision research has been overengineering the problem. Vanilla DINOv2 transformer + depth-ray pairs crushes SOTA by 44% on pose, 25% on geometry. One approach for SOTA monocular depth, multi-view geometry, pose estimation, and novel view synthesis”” / X https://x.com/bilawalsidhu/status/1989444908357488832
ByteDance-Seed/Depth-Anything-3: Depth Anything 3 https://github.com/ByteDance-Seed/Depth-Anything-3
Depth Anything 3 is here! It’s a beefy one! https://x.com/Almorgand/status/1989370456131215514
After a year of team work, we’re thrilled to introduce Depth Anything 3 (DA3)! 🚀 Aiming for human-like spatial perception, DA3 extends monocular depth estimation to any-view scenarios, including single images, multi-view images, and video. In pursuit of minimal modeling, DA3 https://x.com/bingyikang/status/1989358267668336841
Input tokens: 1048576 Output tokens: 65536 “”Our most intelligent model with SOTA reasoning and multimodal understanding, and powerful agentic and vibe coding capabilities”””” / X https://x.com/scaling01/status/1990803527887626446
Hand-controlled hologram boids: Most people have seen hologram tricks. Very few know you can build one that reacts to your hand for about 100 dollars. This demo shows a hand-controlled boids simulation: a small flock of digital particles that moves based on your gestures. No https://x.com/IlirAliu_/status/1989259566740054065
What was a complex hacky pipeline in 2023 to take indoor 3d scans and reskin them to different types of decor is now just a few clicks in 2025. World labs marble has collapsed a lot of the complexity involved in generating and editing 3d worlds: https://x.com/bilawalsidhu/status/1988958359412961743
we built OCR Arena, a free playground for the community to compare leading VLMs and OCR models side-by-side! upload any doc, run 10+ OCR models, and vote for the best ones on a public leaderboard: https://x.com/kushalbyatnal/status/1991898369372082197
Introducing IBench A visual reasoning benchmark designed to test LLMs to spot fine details in images. We test the model on images containing line segments, and ask it to identify and count each intersection of the line segments. The current SOTA on this benchmark is (not https://x.com/adonis_singh/status/1990963148770119889
Gemini 3 and GPT 5.1 are now live in the W&B Weave Playground. You can now test these new models side-by-side, refine your prompts, and evaluate their performance against your actual production traces. Weave helps you experiment and build better AI agents faster. Try it below! https://x.com/weave_wb/status/1991601539728003200
Are you seriously telling me Gemini 3 chat does not have search capabilities?? Is this even a Google product?”” / X https://x.com/Teknium/status/1991059260193792204
Cline 3.38.0 is out now. This release brings Gemini 3 Pro Preview support, @aquavoice_ Avalon as the new model for speech to text and a series of bug fixes in context truncation and native tool calling. Here is what’s new: https://x.com/cline/status/1991215206413017252
Gemini 3 Pro is live across Vercel AI Cloud, and it’s available: • on AI Gateway using 𝚐𝚘𝚘𝚐𝚕𝚎/𝚐𝚎𝚖𝚒𝚗𝚒-𝟹-𝚙𝚛𝚘-𝚙𝚛𝚎𝚟𝚒𝚎𝚠 • as a model on https://x.com/vercel/status/1990816243138633917
Noooooo Gemini, what is this unholy fix https://x.com/cto_junior/status/1990988738298839278
Announcing Design, a new Replit experience focused on beautiful UIs. The first non-slop AI design experience, powered by Gemini 3.0 https://x.com/amasad/status/1990859423942893816?s=20
Try Nano Banana Pro NOW on Together AI: https://x.com/togethercompute/status/1991954662606635391
PRO for PROs Nano Banana PRO is available at no cost for @huggingface PRO subscribers on Spaces, go bananas 🍌 https://x.com/multimodalart/status/1991549140627775511
Don’t underestimate the importance of a good harness that fits the model. In terminal-bench2, GPT-5.1-Codex goes from 16th place (36%) using Terminus 2 to 1st place (57%) using Codex CLI. Gemini 3 Pro enters at #2 with Terminus 2. https://x.com/tristanzajonc/status/1990879703935103256
Computer vision is fucking cool. Matic robo building a real time 3d map of your house. https://x.com/bilawalsidhu/status/1989692317041922270
VisPlay: Self-Evolving Vision-Language Models from Images Introducing a self-evolving RL framework for VLMs to autonomously improve reasoning from unlabeled image data. Achieves SOTA on visual reasoning, compositional generalization, and hallucination reduction! https://x.com/HuggingPapers/status/1991539261175394578
Document AI goes beyond traditional OCR to create intelligent systems that read, understand, and act on documents like humans do. Our latest blog post explains how agentic OCR combined with LLM-powered workflows is transforming document automation across industries: 🧠 Agentic https://x.com/llama_index/status/1990465974357950625
Pretty cool hack to blend between different video feeds to give you the feeling of free viewpoint video AKA. god’s eye view. TL;DR 36 cameras deployed at basketball & badminton venues for China’s National Games, letting viewers drag around on their phones for different angles https://x.com/bilawalsidhu/status/1989362893243154501
🚨👀Vision Leaderboard Update There is a new model provider in the Vision Arena! ERNIE-5.0-Preview-1120 by Baidu @ErnieforDevs has landed with a score of 1206. Just two weeks ago, we shared the Text results, where ERNIE-5.0-Preview shined in Creative Writing, Longer Query, &”” / X https://x.com/arena/status/1991913408221061353
Excited to share another milestone! Our newly released ERNIE-5.0-Preview-1120 has entered the @arena Vision Leaderboard for the very first time! It lands straight in the Top 15 with a score of 1206, on par with Claude Sonnet 4 and GPT-5-high! 🚀 ERNIE-5.0 is natively https://x.com/ErnieforDevs/status/1991898146981789718
Physical Intelligence unveiled π*0.6 (Pi-Star 0.6): a vision-language-action (VLA) model upgraded via their new Recap method (RL with Experience & Corrections via Advantage-conditioned Policies). Recap combines three human-like learning stages: initial demonstrations, real-time https://x.com/TheHumanoidHub/status/1990585956269965743
Ego-VCP: a learned ego-vision world model trained offline on demonstration-free random data that predicts dynamics in latent space for humanoid robots. Achieved robust real-time contact-rich planning on a real Unitree G1: bracing against walls, blocking flying objects, https://x.com/TheHumanoidHub/status/1990558687703019539
Chinese startup MindOn trained Unitree G1 to do house chores. “No speed up, no teleoperation” https://x.com/TheHumanoidHub/status/1989364406850044284





Leave a Reply