Image created with gemini-3.1-flash-image-preview. Image prompt: 1960s Pink Panther cartoon cel of the pink panther leaning casually beside a large easel, holding a dripping paintbrush, having just painted a perfect pink self-portrait on the canvas, flat tangerine orange background with generous negative space, loose black ink outlines and minimalist mid-century graphic design, large hand-lettered playful title text reading IMAGES at the top.
SAM2Matting: Generalized Image and Video Matting” TL;DR: upgrades foundation video trackers like SAM2 into state-of-the-art image and video matting models by decoupling temporal tracking from fine-grained alpha estimation achieving zero-shot video matting without video training”
https://x.com/Almorgand/status/2076705756485746789
i gave 5.6 sol access to my camera roll and had it extract pictures of every piece of clothing i own from my photos then, told it to find new outfits for me and render them on me with gpt-image! its kinda cool to see your entire wardrobe in a collection like this”
https://x.com/cdngdev/status/2076812846793650485
Meta pulls new AI image feature after days of backlash
https://www.bbc.com/news/articles/c2dy6e8klw0o
Newest AI models to explore ↓ Frontier / Commercial GPT-5.6 (Sol, Terra, Luna) GPT-Live Claude Fable 5 Claude Mythos 5 Muse Spark 1.1 Muse Image Muse Video Grok 4.5 Research / Open Gemma 4 InternVLA-A1.5 NVIDIA Audex SenseNova-Vision RynnWorld-4D Vidu S1 AlayaWorld”
https://x.com/TheTuringPost/status/2077203833999229056
WildSplat: Feedforward Gaussian Splatting from Unposed In-the-Wild Images” TL;DR: feed-forward 3D Gaussian Splatting for unposed photos that disentangles geometry from appearance, enabling sota novel view synthesis and appearance editing from sparse in-the-wild images.”
https://x.com/Almorgand/status/2077432202745290838
ChatGPT is available again on WhatsApp in the EEA, part of our work to make AI accessible in the apps people already use every day. Message the verified 1-800-CHATGPT contact to ask questions, upload images, send voice notes, create images, and use ChatGPT in many languages.”
https://x.com/ChatGPT/status/2076654365121855835
Google Images: 25 years of visual search innovation
https://blog.google/products-and-platforms/products/search/google-images-25th-anniversary/
🤗 MOSS-VL-Realtime is now open source on @huggingface . The 11B model family supports text, single and multiple images, single and multiple videos, and interleaved visual-text inputs in Chinese and English.@MosiAI_Official Highlights: 🏗️ Cross-Attention architecture”
https://x.com/Open_MOSS/status/2076993673552879790
muse image explaining schrodinger’s cat!”
https://x.com/alexandr_wang/status/2075804857370599800





Leave a Reply