Image created with Gemini. Image prompt: A flat-scanned 1920s Dada Merz collage on aged kraft board showing a cardboard stereoscope viewer with two paper collage panels inside its lenses that nearly align into one landscape built from torn tickets, ledger grids, and anatomy-book eye diagrams, surrounded by faded vermilion arrows and oxidized ochre numerals. The title ‘AR/VR’ is spelled across the top in mismatched cut-out letterpress typefaces glued at slight angles, with visible torn fiber edges, adhesive stains, foxing, and flat even lighting.
Apple Maps Just Got Insanely Realistic — Here’s The Tech Behind It 00:00 Intro 00:41 What did Apple actually release? 01:30 Why the old way was broken 02:35 What is 3D Gaussian Splatting? 04:06 Live demo: Apple’s photorealistic aerial maps 07:43 The real competition: Apple vs”
https://x.com/bilawalsidhu/status/2068326490656182329
Wow. Seedance is really good at turning greyboxed 3d references into final quality pixels. Like really good. Seedance 2.5 coming with 30 second generations and up to 50 (!) references is gonna be mad. Try the workflow Reid is running below.”
https://x.com/bilawalsidhu/status/2069825605994963047
Is the Strait of Hormuz open or closed? Came up over lunch so I voice noted my agent to “boot up god’s eye view and check.” It sent me back this timelapse clip — refreshed the AIS data, checked oil futures, mapped 120 days of vessel transits and rendered it out. TL;DR it’s”
https://x.com/bilawalsidhu/status/2069149876454314137
Guess what 360 camera support is building up towards? 3d gaussian splatting of course! Only a matter of time before epic rolls it out.”
https://x.com/bilawalsidhu/status/2069855780145103267
This feels like setting cron jobs for the real world. Task semi-autonomous drones to periodically survey your own sanctum or the sandbox. I especially like how the 3d camera frustum changes based on the field of view of the sensor you’re looking through.”
https://x.com/bilawalsidhu/status/2068480671056589308
Watching sports from a god’s eye view. I don’t care what people say, this is how I want to experience sports – as full blown 4d gaussian splats or better. Makes 3d videos look ancient after you try this tech. It’s inevitable. Dare I say new media.”
https://x.com/bilawalsidhu/status/2067978121576362228
📣📣 Meet Qwen-AgentWorld — a native language world model that simulates 7 agent environments (MCP, Search, Terminal, SWE, Web, OS, Android) within a single model. Environment modeling is the training objective from day one, not a post-hoc adaptation. 🤔 LLMs are trained to be”
https://x.com/Alibaba_Qwen/status/2069720365442719867
We open-source Qwen-AgentWorld-35B-A3B (MoE, 35B/3B active, 256K context) and AgentWorldBench. Two routes, one roadmap: 🔬 Build the simulator — scalable, controllable, surpassing real environments 🧠 Internalize world modeling — predict before you act Qwen-AgentWorld is our”
https://x.com/Alibaba_Qwen/status/2069720412481888400
[2606.24597] Qwen-AgentWorld: Language World Models for General Agents
https://arxiv.org/abs/2606.24597
🧠 Paradigm II — Agent Foundation Model: world modeling as agent capability. Single-turn, non-agentic environment prediction → tested directly on multi-turn, tool-calling agent tasks. No agentic RL, no task-specific tuning. Gains across 7 benchmarks, including 3 entirely”
https://x.com/Alibaba_Qwen/status/2069720397747220493
Mondo is the most adorable companion robot. It’s safe around kids, can track and follow a person, and capture high-resolution action videos. Mondo is also doing frontier research in humanoid Video-Action models built on a world model backbone:
https://t.co/vuOIngXzVa. I have no”
https://x.com/TheHumanoidHub/status/2067409544574087632
We taught a brand-new mini-series this year at @SCSatCMU on Modern GPU Programming for ML Systems, as part of the ML Systems course, touching on fun questions like what data layout swizzling is, how to use 3D TMA, and state-of-the-art Blackwell programming. We released a curated”
https://x.com/tqchenml/status/2069382647302734099
Cornell’s Robot Learning Course. Every slide. Every assignment. Free. 📌 CS 4756 covers the full modern stack: → Imitation learning → Reinforcement learning → Model predictive control → 3D perception → Sim-to-real transfer → LLMs for robot control Every lecture slide is”
https://x.com/IlirAliu_/status/2068035229424423366
FiCA: Feed-forward instant Gaussian Codec Avatars from a Single Portrait Image
https://kim-youwang.github.io/FiCA
FLAT | Feedforward Latent Triangle Splatting
https://flat-splat.github.io/
street view grounding now available for google’s offline video models. can’t wait till you can do “spatial RAG” to load in the right panoramas to reference – suddenly large scale real world locations become movie sets!”
https://x.com/bilawalsidhu/status/2069549894335946939





Leave a Reply