Image created with Gemini. Image prompt: A flat acrylic collage of slightly offset overlapping rectangular photo panels forming one scene of a swimmer mid-dive splashing into a turquoise pool with hand-drawn white squiggle ripples, each panel capturing a slightly different moment and angle of the same dive, flat saturated colors of pool turquoise terracotta pink lawn green and hot yellow, hard-edged shapes with no shading, high noon light with no shadows, tablet-drawn digital line, generous white space.

Apple Beat Google to It! Hyper-Realistic Maps at WWDC 2026 — Gaussian Splats & the ongoing 3D maps land grab to own the substrate that connects the physical and digital world. Hopping on with @RadianceFields to discuss.
https://x.com/bilawalsidhu/status/2064146894930989138

holy crap! apple just beat google to the punch — 3d gaussian splatting is coming to apple maps. these 3d scenes are made from oblique aerial imagery. but unlike blobby photogrammetry — no more broccoli trees, no more melted powerlines — ground level detail that actually holds
https://x.com/bilawalsidhu/status/2064057313057439795

ICYMI google’s been shipping gaussian splats in maps for a while — but tucked inside immersive view and indoor only. I used it heavily in NYC. Meanwhile apple’s bringing radiance fields to 300+ cities this fall, built from aerial imagery. Surprised they didn’t drop ground-level
https://x.com/bilawalsidhu/status/2064187152930365502

Art Directors Guild Slams Martin Scorsese for AI Partnership
https://variety.com/2026/film/news/art-directors-guild-statement-martin-scorsese-ai-1236770996/

Runway News | Runway and Lionsgate Expand Partnership
https://runwayml.com/news/runway-and-lionsgate-expand-partnership

First Aronofsky, then Scorsese, and now Gareth Edwards. Generative media has been polarizing – but is the Overton window shifting before our eyes?
https://x.com/bilawalsidhu/status/2062358867510440354

Apple’s gaussian splat maps are rolling out on the developer beta! I hope they make a legit vision pro app – would be dope to drop into a city with the homies embodied as persona avatars.
https://x.com/bilawalsidhu/status/2064494805023911966

🎉 Meet vLLM-Omni v0.22.0, a major upgrade for omnimodal world models and production-grade multimodal serving. 🌍 Day-0 @NVIDIAAI Cosmos 3 world models: text, image, audio, video, and action, in and out. 🤖 Robot serving: DreamZero + OpenPI realtime API. 🎙️ Production TTS:
https://x.com/vllm_project/status/2064013506882703421

llama.cpp just added video input support 👀 You can now enjoy Gemma 4 video understanding capabilities in your chat completions endpoint and via mtmd-cli
https://x.com/osanseviero/status/2063985470489448887

You may have recently heard claims that video generation models are “”dumb”” about physics, and only “”world models”” (V-JEPA, specifically) have a valid internal model of physics. This turns out to be false. In a recent paper, researchers show that a LINEAR probe of diffusion
https://x.com/giffmana/status/2064718736783823145

1X has hired Samarth Sinha to lead its new World Models research group. Samarth was previously a founding researcher at Luma AI, scaling large multimodal models. The mandate: build the next generation of foundation world models for NEO humanoid, trained on large-scale,
https://x.com/TheHumanoidHub/status/2062595716946784758

A new v0 robotics benchmark by independent researcher that tests robot foundation models on four tasks: Requiring spatial reasoning, geometric understanding, occlusion handling, and visuospatial planning. Using a low-cost open-source SO-101 robotic arm. The tasks progress in
https://x.com/IlirAliu_/status/2063894548821016989

A zero-shot video-language reward model. Trained on over 1M trajectories from 21 robot embodiments. It predicts: Frame-level task progress and generalizes zero-shot to unseen tasks, scenes, and robots, yielding 2.4-4.5x better success rates in: Online/offline RL, data
https://x.com/IlirAliu_/status/2063168037964877993

Boston Dynamics taught Atlas a Rabona kick. Learned from human mocap, retargeted to Atlas, trained in sim via RL, deployed zero-shot to the real robot. Soccer skills demand whole-body coordination, and similar recipe transfers to warehouse work.
https://x.com/TheHumanoidHub/status/2062607675339497680

If you can teach a robot a new skill in under an hour, you’ve just made your whole lab more flexible overnight. Pharma and biotech need automation that’s flexible enough for unstructured work… yet auditable enough to trust in regulated environments. That combination doesn’t
https://x.com/IlirAliu_/status/2062958380239511894

One response to “Video: AI News Week Ending 06/12/2026”

Leave a Reply

Trending

Discover more from Ethan B. Holland

Subscribe now to keep reading and get access to the full archive.

Continue reading