Image created with gemini-3.1-flash-image-preview with claude-opus-4.7. Image prompt: A stylized flattened Byzantine saint icon centered in a gold-ground mosaic apse, wearing a hammered-gold VR headset like a sacred blindfold with fine spiral gyres of light radiating from its edges into the tesserae, hands holding two tiny mosaic worlds; burnished gold, imperial purple and Tyrian crimson robes, candlelit sacral glow, tactile grout texture, the words ‘AR / VR’ set as bold ivory Trajan Roman capitals across the lower third, 16:9 full-bleed.

360 drone has arrived. Looking fwd to capturing some big ass gaussian splats.
https://x.com/bilawalsidhu/status/2057635347576524823

Bridging 3D prior from generative models and real 3D scan (only RGB!) is getting sharper everyday!
https://x.com/Almorgand/status/2059011590167392765

Fast 4D Mesh Generation by Spatio-Temporal Attention Chains”” TL;DR: training-free spatio-temporal attention chains generate topology-consistent 4D meshes from video 13× faster while improving temporal correspondence quality
https://x.com/Almorgand/status/2059206856589844552

Large-Scale High-Quality 3D Gaussian Head Reconstruction from Multi-View Captures”” TL;DR: feed-forward UV-parameterized Gaussian head reconstruction scales to 10k+ identities while enabling high-quality avatars and real-time facial animation
https://x.com/Almorgand/status/2059388320078037237

Made a free Pixal3D demo (Tencent’s new image-to-3D model) because I like it a lot 🔥 What’s interesting: pixel-aligned generation: every point in the mesh ties back to a specific input pixel, so silhouettes, textures, and tiny details actually survive into the 3D asset instead
https://x.com/victormustar/status/2057752615396557225

RecGen: 3D Multi-Object Scene Reconstruction from Sparse Observations”” TL;DR: generative 3D scene reconstruction framework that recovers geometry, texture, and pose from sparse RGB-D observations while remaining robust to heavy occlusions and object symmetries
https://x.com/Almorgand/status/2059568924841091282

This is pretty cool — basically “creative upscaling” for 3d scans
https://x.com/bilawalsidhu/status/2059029365002744034

This is text-to-CAD right now. A planetary gear assembly in CAD Explorer where users adjust drive parameters and animation speed, showing gears rotating and meshing realistically. Simulating forces on complex mechanisms strains AI spatial reasoning, yet produces impressive
https://x.com/IlirAliu_/status/2057369823676453305

Velox: Learning Representations of 4D Geometry and Appearance”” TL;DR: learns compact dynamic 4D latent tokens that jointly model geometry and appearance for tasks like video-to-4D generation and 3D tracking
https://x.com/Almorgand/status/2057840305588613505

VGGT-Ω”” TL;DR: scales feed-forward 3D reconstruction to dynamic scenes with memory-efficient register attention, enabling large-scale spatial foundation modeling
https://x.com/Almorgand/status/2057505812868641149

Synthesize Realistic 3D Medical Images at Scale to Ship Pre‑Trained Models | NVIDIA Technical Blog
https://developer.nvidia.com/blog/synthesize-realistic-3d-medical-images-at-scale-to-ship-pre-trained-models/?ncid=so-twit-899791

Leave a Reply

Trending

Discover more from Ethan B. Holland

Subscribe now to keep reading and get access to the full archive.

Continue reading