Image created with Flux Pro v1.1 Ultra. Image prompt: Photo desk with lightbox and contact sheets; the word “Images” printed on a color-checker card in neutral sans; AI upscaling notes taped to the monitor; tactile, sharp, studio glow
Playing with the new mystery “”nano-banana”” image generation model: “”a photo where a woman with a pink mask covering just the left side of her face, and the right side is painted green, she is wearing a duck costume, but the feet are muddy. she stands next to a golden retriever https://x.com/emollick/status/1957588350207938937
bangs successfully removed with 8-step Qwen Image Edit [Fast] too 💨 using Qwen Image Lightning LoRA, now on Spaces👇 https://x.com/linoy_tsaban/status/1957762030393544847
Mechaverse https://mechaverse.dev/
🍌”” / X https://x.com/OfficialLoganK/status/1957908528925909391
Generate now: https://t.co/iRl3CSWuvF Generated with Imagen 4 Fast “` { “”scene””: “”minimalist white studio””, “”subjects””: [ { “”type””: “”smartwatch””, “”description””: “”silver frame with red strap””, “”position””: “”center””, “”pose””: “”lying flat”” } ], “”style””: “”photorealistic””, “”lighting””: https://x.com/_philschmid/status/1956351658381705420
Here’s what we shipped this week 🚢🚢🚢 —We launched a new Imagen 4 Fast model so developers can quickly generate images at only $0.02 per image and updated Imagen 4 and Imagen 4 Ultra to support 2K images. All are now generally available in the Gemini API for developers and”” / X https://x.com/GoogleAI/status/1956400937054163357
Imagen 4 is now GA in AI Studio and the Gemini API! Choose from three models: Ultra, Standard, and Fast starting at 2ct per image. ⚡️ Experience up to 10x faster generation than our previous model. 🖼️ Create images with up to 2k resolution. ✍️ Render longer text strings with https://x.com/_philschmid/status/1956351654753673252
Now @GooglePhotos can help you make custom AI-powered edits in seconds — just by asking. You can ask for simple corrective edits — like “remove the cars in the background” — or more creative ones, like adding a party hat or sunglasses. And if you don’t know where to start, try https://x.com/Google/status/1958946812817019305
Google Pixel 10: 9 new AI features and updates https://blog.google/products/pixel/google-pixel-10-ai-features-updates/
It is interesting to see how much effort is going into making ancillary features of the AI models go viral. Ever since the (organic) Studio Ghibli moment, one focus for Grok & Gemini has been on video as a gateway. A challenge has been whether people have creative video ideas.”” / X https://x.com/emollick/status/1956312948130947339
🖼️ Imagen 4 is now Generally Available! To show how powerful the new ⚡fast⚡ model is, try out this movie-pictionary game forked from @alexanderchen. 📽️ Guess the movie using Imagen 4 Fast. Link is below ⬇️ Here are a couple of my faves. https://x.com/m4rkmc/status/1956238192035663874
Developed a more generalized version. The pattern 東南西北 can be displaced, but remains conserved. https://x.com/RavenKwok/status/1958157337187020973
🖼️🚨 Text-to-Image Leaderboard Update A new contender: Lucid Origin debuts on the Text-to-Image Leaderboard. Ranking at #9, this is a new model provider to enter the Top 10! Congrats to the @LeonardoAi_ team 👏 https://x.com/lmarena_ai/status/1958965415180476654
Google dropped several products and updates, including: —Memory and temporary chat in Gemini —General availability of Imagen 4, with a fast version —A new Gemma 3 variant with 270M parameter https://x.com/adcock_brett/status/1957111015306350998
RotBench Evaluating Multimodal Large Language Models on Identifying Image Rotation https://x.com/_akhaliq/status/1958635243197325625
🎨✨ From simple sketches to stunning 3D interiors — powered by Qwen-Image-Edit! All designs are community contributions, showcasing how AI transforms architectural visions into realistic, stylish, and precise creations. Try it now: https://x.com/Alibaba_Qwen/status/1958744976772198825
📸 Just showed Qwen Chat Vision Understanding how to “”see”” and understand a meal — and it didn’t just identify the food, it analyzed what, where, weight and even how many calories! From a simple photo, we extracted detailed insights: ✅ Object detection ✅ Weight estimation ✅ https://x.com/Alibaba_Qwen/status/1956618027769971070
🖼️ 🚨 Image Edit Leaderboard Update: Qwen-Image-Edit is now the #1 open model for Image Edit in the Arena (Apache 2.0). The model by @alibaba_qwen debuts at #6 overall on the Image Edit leaderboard tied with Gemini 2.0 Flash Preview. https://x.com/lmarena_ai/status/1958206842657743270
🖼️ Image Edit Model Update Qwen-Image-Edit, developed by @Alibaba_Qwen, is now available in the Arena. This model brings image editing capabilities, and we encourage you to test it with your most complex prompts. https://x.com/lmarena_ai/status/1957878222986821711
🚀 Excited to introduce Qwen-Image-Edit! Built on 20B Qwen-Image, it brings precise bilingual text editing (Chinese & English) while preserving style, and supports both semantic and appearance-level editing. ✨ Key Features ✅ Accurate text editing with bilingual support ✅ https://x.com/Alibaba_Qwen/status/1957500569029079083
🚀 Small but mighty update to Vision Understanding in Qwen Chat — now with native 128K context and stronger performance across vision, video, and 3D tasks! 🔥 Key Upgrades: ✅ Significant boost in math & reasoning ✅ More accurate object recognition ✅ OCR support for 30+ https://x.com/Alibaba_Qwen/status/1956289523421470855
NextStep-1: Toward Autoregressive Image Generation with Continuous Tokens at Scale “”Autoregressive models—generating content step-by-step like reading a sentence—excel in language but struggle with images. Traditionally, they either depend on costly diffusion models or https://x.com/iScienceLuvr/status/1956321483183329436
Qwen Image Edit works too well with lightx2v LoRA to run with just 8 and 4 steps, wtf? in my experience, 8 steps keeps the quality of the edits at the same level as the original model, at a 12x speedup 💨 (ofc i built a demo for it) https://x.com/multimodalart/status/1958217824629092568
Qwen-Image Edit in ComfyUI”” / X https://x.com/Alibaba_Qwen/status/1957991583649001555
Qwen-Image-Edit is out in anycoder for image editing in your vibe coded apps Built on 20B Qwen-Image, it brings precise bilingual text editing (Chinese & English) while preserving style, and supports both semantic and appearance-level editing. https://x.com/_akhaliq/status/1957519569016238268
Qwen-Image-Edit is the new open weights leader in Image Editing, with quality comparable to GPT-4o and FLUX.1 Kontext [max] Qwen-Image-Edit is the image editing variant of the recent Qwen-Image release from Alibaba, also released under the Apache 2.0 license with weights https://x.com/ArtificialAnlys/status/1958712568731902241
Qwen-Image-Edit: Image Editing with Higher Quality and Efficiency | Qwen https://qwenlm.github.io/blog/qwen-image-edit/
Relighting images with Qwen Edit impressive directional control and color temperature manipulation w/o additional finetuning crazy how we needed a dedicated model for this not long ago https://x.com/linoy_tsaban/status/1958176756185325931
Thank you! Qwen-Image-Edit is now available in anycoder!”” / X https://x.com/Alibaba_Qwen/status/1957709912202682588
👀🚨 Vision Leaderboard update! Two new models have entered the Vision Top 20 this week: 🔸Qwen-vl-max-2025 by @alibaba_qwen lands at #10 (tied with gemini-1.5-pro & gpt-5-nano-high) 🔸Step 3 by @StepFun_ai ranks at #19 (tied with step-lo-turbo) Congrats to both 🎉 this is https://x.com/lmarena_ai/status/1958957107946168470
Wow — Qwen-Image-Edit just debuted at #2 in the Image Editing Arena 🏆 ELO 1098, with performance on par with GPT-4o — and all at open weights under Apache 2.0. Thanks to @ArtificialAnlys Try it now: https://x.com/Alibaba_Qwen/status/1958725835818770748
New from S-Lab, Nanyang Technological University & SenseTime Research: Next Visual Granularity Generation (NVG)! This novel framework progressively refines images from global layout to fine details, offering fine-grained control over generation. It outperforms the VAR series in https://x.com/HuggingPapers/status/1957836902020612180
🐞 We hit a bug in the inference code for Qwen-Image-Edit on Diffusers, which caused some odd cases. ✅ Fixed now and thanks to Diffusers for the quick merge — give it another try! 🔗 Try it now: https://x.com/Alibaba_Qwen/status/1957840853277290703
AI Toolkit now supports fine tuning Qwen Image Edit and supports caching the text embeddings with the control images. I already trained a 3 bit ARA for it, which will allow you to train a LoRA at 1024 on a 5090 when caching the text embeddings. More in 🧵 https://x.com/ostrisai/status/1958932936620900666
It’s out friends! Really great to see the state of things in image edits, video fidelity being pushed further and further, thanks to the community! This release also features new fine-tuning scripts for Qwen-Image and Flux Kontext (with support for image inputs). So, get busy https://x.com/RisingSayak/status/1957668389935096115
nano-banana, qwen-image-edit, what else? Try @StepFun_ai NextStep-1-Large-Edit – 14B AR model – Apache 2 license – Demo available on @huggingface – Pretrain model also made available Link below https://x.com/Xianbao_QIAN/status/1957749693485838448
qwen image edit is back at #1 trending model at @huggingface 👑 https://x.com/multimodalart/status/1958229738398634171
Qwen-Image pruning experiment. Going from 60 to 30 blocks, 20B params to 10B params. Removed block idx 2, 3, 4, 5, 7, 8, 10, 11, 12, 13, 14, 15, 16, 21, 23, 24, 40, 41, 42, 43, 44, 45, 49, 50, 51, 52, 53, 54, 55, 56 https://x.com/ostrisai/status/1957748358451503166
stepfun-ai/NextStep-1 https://github.com/stepfun-ai/NextStep-1
xAI made its flagship model, Grok 4, free for all, with “”generous”” usage limits for a limited time The company has also made its Imagine video generator free for everyone using the Grok app, for a limited time https://x.com/adcock_brett/status/1957110985824645422
Build an open-source AI video studio with this @nextjs template using Veo 3 and Imagen 4 in the Gemini API. Create text-to-video, image-to-video, and edit videos in the browser for a specific time range. https://x.com/googleaidevs/status/1958599306472206349
Build your own video generation Studio with Veo 3 in Gemini API! Excited to share an open-source @nextjs template for generating videos with Google’s Veo 3 and Imagen 4. Features: – Generate videos from text prompts using the Veo 3 model. – Create videos from images and text https://x.com/_philschmid/status/1957821851331416079




