Image created with gemini-3.1-flash-image-preview with claude-opus-4.7. Image prompt: A symmetrical Byzantine gold-ground mosaic apse featuring a single centered flattened iconic saint with a circuit-ring halo, holding a hammered-gold film reel and a golden clapperboard, flanked by a vertical strip of mosaic film-frame tesserae showing the figure in successive poses like a sacred zoetrope, faint spiral gyre of frames behind the halo, burnished antique gold with imperial purple and Tyrian crimson robes, warm candlelit directional glow catching hammered metal, tactile tesserae grout texture, the bold Trajan-capital title VIDEO in ivory-and-gold across the lower third, 16:9 full-bleed, painterly reverent illuminated-manuscript render.
I think people don’t realize why Gemini Omni is different than other video AIs. It is fully multimodal, so it can edit video natively, too I took the famous “”train “” movie from 1896 & made it a bullet train, LEGO, added a time traveler, a centipede, muppets… (see reflections?)
https://x.com/emollick/status/2057874739817808223
Omni is pretty nuts. It is NOT seedance. Any input in/out. It’s more than nano banana for video – it’s quite literally industrial light & magic. Now effectively reduced to an insanely realistic AR filter that you can apply on demand to anyone’s footage.
https://x.com/bilawalsidhu/status/2057300479340695960
Gave google omni a sketched camera path and asked it to generate drone POV footage.
https://x.com/bilawalsidhu/status/2059419767417487718
Google just revealed Omni, personalized cross-device intelligence, and Spark agents at I/O 2025. I sat down with CEO Sundar Pichai to figure out what comes next: 1:46 Omni: “”Nano Banana for video”” 4:59 The future of YouTube 7:04 Advice for AI skeptics 9:33 Why your mom should
https://x.com/rowancheung/status/2057491344697012384
Nano Banana for video is here 🍌🎥 Gemini Omni is our new AI model that makes creating and editing videos as easy as having a conversation. Here’s how it works ↓
https://x.com/Google/status/2057881884219035752
We came across a really interesting tool that fixes a pretty common problem inside AI video workflows: extending AI cinematic scenes with seamless continuity. Using Omni models, you can take the last frame from an existing video clip and prompt something like: “show me what
https://x.com/CuriousRefuge/status/2057920807389806699
Gemini Embedding 2: A Native Multimodal Embedding Model from Gemini 🚀 Today, we’re sharing the @GoogleDeepMind white paper for GE 2, our first native multimodal embedding model. Whether it’s text, audio, video, or image, GE 2 provides a unified representation of the input.
https://x.com/mseyed/status/2059504005387284629
Google has the only true Omni model, but the elements aren’t hooked up. It appears it can take in & output audio, images. video, songs, text, code, etc. But right now each type of output is separate. When you can access the model directly, blending modes, a lot becomes possible.
https://x.com/emollick/status/2059774997535584325
Yesterday we released Aleph 2.0, our upgraded video editing model that lets you change exactly what you want while keeping everything else the same. Available inside our new Edit Studio, you can work with multishot sequences up to 30 seconds long at 1080p. Learn how to get
https://x.com/runwayml/status/2057826728769134599
360 drone has arrived. Looking fwd to capturing some big ass gaussian splats.
https://x.com/bilawalsidhu/status/2057635347576524823
Bridging 3D prior from generative models and real 3D scan (only RGB!) is getting sharper everyday!
https://x.com/Almorgand/status/2059011590167392765
Fast 4D Mesh Generation by Spatio-Temporal Attention Chains”” TL;DR: training-free spatio-temporal attention chains generate topology-consistent 4D meshes from video 13× faster while improving temporal correspondence quality
https://x.com/Almorgand/status/2059206856589844552
Large-Scale High-Quality 3D Gaussian Head Reconstruction from Multi-View Captures”” TL;DR: feed-forward UV-parameterized Gaussian head reconstruction scales to 10k+ identities while enabling high-quality avatars and real-time facial animation
https://x.com/Almorgand/status/2059388320078037237
Made a free Pixal3D demo (Tencent’s new image-to-3D model) because I like it a lot 🔥 What’s interesting: pixel-aligned generation: every point in the mesh ties back to a specific input pixel, so silhouettes, textures, and tiny details actually survive into the 3D asset instead
https://x.com/victormustar/status/2057752615396557225
RecGen: 3D Multi-Object Scene Reconstruction from Sparse Observations”” TL;DR: generative 3D scene reconstruction framework that recovers geometry, texture, and pose from sparse RGB-D observations while remaining robust to heavy occlusions and object symmetries
https://x.com/Almorgand/status/2059568924841091282
This is pretty cool — basically “creative upscaling” for 3d scans
https://x.com/bilawalsidhu/status/2059029365002744034
This is text-to-CAD right now. A planetary gear assembly in CAD Explorer where users adjust drive parameters and animation speed, showing gears rotating and meshing realistically. Simulating forces on complex mechanisms strains AI spatial reasoning, yet produces impressive
https://x.com/IlirAliu_/status/2057369823676453305
Velox: Learning Representations of 4D Geometry and Appearance”” TL;DR: learns compact dynamic 4D latent tokens that jointly model geometry and appearance for tasks like video-to-4D generation and 3D tracking
https://x.com/Almorgand/status/2057840305588613505
VGGT-Ω”” TL;DR: scales feed-forward 3D reconstruction to dynamic scenes with memory-efficient register attention, enabling large-scale spatial foundation modeling
https://x.com/Almorgand/status/2057505812868641149
Synthesize Realistic 3D Medical Images at Scale to Ship Pre‑Trained Models | NVIDIA Technical Blog
https://developer.nvidia.com/blog/synthesize-realistic-3d-medical-images-at-scale-to-ship-pre-trained-models/?ncid=so-twit-899791
Camera path scribbles are turning into an AI video trend. Cool seedance example by Keety.
https://x.com/bilawalsidhu/status/2059751008746471791





Leave a Reply