Image created with gemini-2.5-flash-image with claude-sonnet-4-5. Image prompt: Cinematic night scene with vast starry sky above dark pastoral field, subtle film strip perforation texture visible in grass at bottom edge, bold white sans-serif text VIDEO centered in upper sky, widescreen composition, deep navy and black tones with silver stars and moonlit grass, film poster aesthetic.

Detail Enhanced Gaussian Splatting for Large-Scale Volumetric Capture”” TL;DR: Full Studio pipeline for 4D volumetric capture (GS); HD Scene capture + Face capture https://x.com/Almorgand/status/1993730815818154258

Turned my real world 360 images into 3d scenes with World Labs. You can use the built-in editing tools to stitch them together into large-scale 3d worlds – then use them as virtual set for your AI videos, games and VR experiences. The holodeck is much closer than you think. https://x.com/bilawalsidhu/status/1992703722473038259

YouTube test features and experiments – YouTube Community https://support.google.com/youtube/thread/18138167/youtube-test-features-and-experiments?sjid=82014890340299316-NA

RynnVLA-002 A Unified Vision-Language-Action and World Model https://x.com/_akhaliq/status/1992973969784213854

𝚁𝚢𝚗𝚗𝚅𝙻𝙰-𝟎𝟎𝟐 RynnVLA-002: a unified action world model that evolves RynnVLA-001 from “VLA and generative priors” into a fully joint Chameleon-based action-world framework, merging VLA policy and world model in one autoregressive transformer with shared token space for https://x.com/gm8xx8/status/1992910628172800412

We’re launching a new frontier physics eval on Artificial Analysis where no model achieves greater than 9%: CritPt (Complex Research using Integrated Thinking – Physics Test) Developed by 60+ researchers from 30+ institutions across the world including the Argonne National https://x.com/ArtificialAnlys/status/1991913465968222555

3d visual positioning experiment — look at the alignment between the 3d mesh and the live camera view. Truly feels magical — like x-ray vision. Workflow: 1. Scanned a street in 15 mins w/ xgrids 2. Localized against that scan at night, in real-time, while sitting in a car The https://x.com/bilawalsidhu/status/1993444825463701958

This AI paper just solved Google Earth’s biggest problem. Satellites look down. Humans look across. That perspective gap is why 3D maps are limited to cities you can blanket with aerial flyovers. Skyfall-GS bridges the gap by synthesizing the views we never captured – https://x.com/bilawalsidhu/status/1992051324096238068

Vibe coded this JARVIS inspired HUD with laser eyes and repulsor hands in 20 mins w/ Gemini 3 Pro using MediaPipe and ThreeJS. Augmented reality filters were always cool – but vibe coding makes them so much more accessible than complex AR authoring tools of the past. https://x.com/bilawalsidhu/status/1993025326520451475

Google DeepMind on X: “To celebrate five years of #AlphaFold, we’re making The Thinking Game available on YouTube. 🧬 Get a candid look at the triumphs, the challenges and the pivotal moments that led to a breakthrough on a 50-year-old grand challenge in biology. Stream for free on @YouTube → https://t.co/CcTf2vivR9” / X
https://x.com/GoogleDeepMind/status/1993714943116386619

STARFlow-V: End-to-End Video Generative Modeling with Normalizing Flows “”In this work, we revisit this design space by presenting STARFlow-V, a normalizing flow-based video generator with substantial benefits such as end-to-end learning, robust causal prediction, and native https://x.com/iScienceLuvr/status/1993629956375822508

Leave a Reply

Trending

Discover more from Ethan B. Holland

Subscribe now to keep reading and get access to the full archive.

Continue reading