Clothing Giant H&M Will Use Models’ AI-Made Digital Twins, Consent Included https://www.inc.com/kit-eaton/clothing-giant-hm-will-use-models-ai-made-digital-twins-consent-included/91166352

Ideogram 3.0 https://about.ideogram.ai/3.0

Reve on X: “Halfmoon is Reve Image — and it’s the best image model in the world 🥇 (🔊) https://t.co/Zm1FzNQaFh” / X
https://x.com/reveimage/status/1904211082870456824

“The Halfmoon 🌓 reveal: Congratulations to @reveimage on creating the world’s leading image generation model with Reve Image! Reve Image has been in the Artificial Analysis Image Arena over the past week and is the clear leader, beating strong competition including Recraft V3, https://x.com/ArtificialAnlys/status/1904188980423467472

“Even with multimodal image generation, ChatGPT only gets 1 out of 3 of the hard challenges of image creation: ✅Horse riding an astronaut ❌Clock reading 5:30 ❌Full to the brim glass of wine Presumably, the training image data is mostly not-full glasses and clocks set to 10:10 https://x.com/emollick/status/1904710558919651819

Introducing 4o Image Generation | OpenAI https://openai.com/index/introducing-4o-image-generation/

“New image model from OpenAI is pretty good at UI stuff. https://x.com/skirano/status/1904609866099933272

“// i lead model behavior at openai, and wanted to share some thoughts & nuance that went into setting policy for 4o image generation. features capital letters (!) bc i published it as a blog post: — This week, we launched native image generation in ChatGPT through 4o. It was” / X https://x.com/joannejang/status/1905341734563053979

“This is a pretty cool feature that got drowned by Ghiblification hype: you can ask 4o image gen for transparent backgrounds! Should be super useful for creating all kinds of assets.” / X https://x.com/giffmana/status/1905407013103747422

“Holy crap. Reve is REALLY good. No surprise that it’s currently #1 in the Artificial Analysis text-to-image leaderboard — ahead of Recraft v3, Imagen v3 and FLUX 1.1 Pro. This team cooked, and it shows. https://x.com/bilawalsidhu/status/1904325105267683481

New Reve Image Generator Beats AI Art Heavyweights MidJourney and Flux at a Penny Per Image – Decrypt https://decrypt.co/311375/new-reve-image-generator-beats-ai-art-heavyweights-midjourney-and-flux-at-a-penny-per-image

“So, @Midjourney Weekly Office Hours just ended. Brace yourselves, V7 is coming next week. I use a slide when I present on generative AI with MidJourney version dates. People are always impressed by the fact that they went from V1 in February 2022 to V6 in December 2023. That’s https://x.com/alanxtruc/status/1905009099013521554

“The instruction following capability of gpt-4o native image gen is pretty incredible. Nothing comes close to this, how could every other image gen model miss the mark” / X https://x.com/abacaj/status/1905075484892836308

“GPT-4o image gen is fun. It can do photorealistic & stylized renditions while keeping consistent structure and pose. “Make an image of a UFO parking not allowed sign with a UFO conspicuously parked in that spot, with a grey alien arguing with a police officer in front of it” https://x.com/bilawalsidhu/status/1904627124109271238

“Excited to come out of stealth at @reveimage! Today’s text-to-image/video models, in contrast to LLMs, lack logic. Images seem plausible initially but fall apart under scrutiny: painting techniques don’t match, props don’t carry meaning, and compositions lack intention. (1/4) https://x.com/Taesung/status/1904220824435032528

“we are launching a new thing today—images in chatgpt! two things to say about it: 1. it’s an incredible technology/product. i remember seeing some of the first images come out of this model and having a hard time they were really made by AI. we think people will love it, and we” / X https://x.com/sama/status/1904598788687487422

“Halfmoon is Reve Image — and it’s the best image model in the world 🥇 (🔊) https://x.com/reveimage/status/1904211082870456824

“Kling 1.6 now dominates Image-to-Video Leaderboard in artificialanalysis @Kling_ai https://x.com/rohanpaul_ai/status/1905581276217696720

“💥 Today we’re rolling out a *major* update to image generation in ChatGPT! The model is now quite good at following complex instructions, including detailed visual layouts. It’s very good at generating text. It can do photorealism or any number of other styles. Btw—if you” / X https://x.com/kevinweil/status/1904595752380465645

“Entire ComfyUI workflows just became a text prompt. Open an image in GPT-4o and type “turn us into Roblox / GTA-3 /Minecraft / Studio Ghibli characters” https://x.com/bilawalsidhu/status/1904908540063424898

“My “otter on a plane using wifi” benchmark has now been saturated by ChatGPT 4’s new image generator: “an otter on an airplane using wifi, on their laptop screen is image generation software creating an image of an otter on a plane using wifi,” first try https://x.com/emollick/status/1904943271282934215

“”ChatGPT show me a photorealistic drone shot of a fantasy city, the walls are white alabaster streaked with gold, while massive brass trellises built into the towers allow vines of maroon to climb them. At the center is a tower with a balcony and a hooded figure on it” “in the https://x.com/emollick/status/1904720324530254049

“4o-image is the first time many people will start to think of the post-reality-filter stage rolling out over the next few years. one’s reality is whatever they want it to be (ghibli or pokemon or lotr or…), and as each human finds that which they truly desire, they then” / X https://x.com/nearcyan/status/1905219687547621740

OpenAI GPT image teaser example “the pros and cons https://x.com/ajabri/status/1904599427366739975

“Sure, you could use an annotation tool to create bounding boxes of objects in images for you… or you can ask a multimodal AI to do it freehand. https://x.com/emollick/status/1904028116063822141

“The Screenshot: a fake screenshot generated by ChatGPT 4o of a Wikipedia article about the screenshot itself, with a copy of the screenshot in the article https://x.com/goodside/status/1904743355147235834

“I’ve had access to the new GPT-4o image generator for a bit: “now i need you to perfectly illustrate the moment that Elvis met Napolean at waterloo” “make it photorealistic” “napoleon is wearing a rubber duck on his head elvis’s pants have the ideal gas law printed on them” https://x.com/emollick/status/1904608706970398841

“Multimodal image generation is going to actually impact a lot of economically and culturally meaningful work in ways I don’t think we understand yet. It is very flexible, relevant to many uses & got good all at once. Still flaws, but the gain in capabilities seems rather rapid.” / X https://x.com/emollick/status/1904784680416608259

FLUX.1 [Inpainting] – a Hugging Face Space by SkalskiP https://huggingface.co/spaces/SkalskiP/FLUX.1-inpaint

“Image generation just landed on the xAI API. Pretty cool stuff—developers can now build some wild visuals. Have at it! – API Console: https://x.com/xai/status/1903098565536207256

BRIA.ai | Visual Generative AI Done Right https://platform.bria.ai/register

“ChatGPT, show me the backpack and materials I would need to carry back in time to become Roman Emperor. https://x.com/emollick/status/1904918711124840485

“Thanks to 4o-imagegen, the world finally knows how to draw the rest of the fucking owl: https://x.com/giffmana/status/1904645482024202365

“believe it or not we put a lot of thought into the initial examples we show when we introduce new technology” / X https://x.com/sama/status/1905069374035411209

“Ghiblification will continue until morale improves” / X https://x.com/iScienceLuvr/status/1905363176943759836

“After an incredible 3 years leading model development at Midjourney, I’ve joined Cursor to work on coding agents. I’m incredibly proud of my time at Midjourney and the work we did, of the results of that singular focus on beauty and creativity.” / X https://x.com/gallabytes/status/1902864624510439516

“People are complaining it’s slightly inaccurate. That’s fair but to me that seems mostly an issue of the LLM part of it I am instead highlighting the impressive compositioning, text, overall flow of the diagram, etc. which was generated without needing to be specified at all” / X https://x.com/iScienceLuvr/status/1905067166350884941

Why Bria Stands With Data Owners and Creators For Visual Generative AI https://blog.bria.ai/the-fight-for-a-fair-ai-future

“i have updated the meme https://x.com/swyx/status/1902935741103215103

“Fun image editing prompt you can try in GPT-4o. “Add a huge explosion in the background and add light wrap around me” https://x.com/bilawalsidhu/status/1904763899766767957

Single Image Iterative Subject-driven Generation and Editing https://siso-paper.github.io/

“Today’s visual generative models are mere stochastic parrots of imagery, much like early language models, which could only statistically mimic short sentences with little reasoning. In contrast, modern large language models (LLMs) can comprehend long documents, keep track of” / X https://x.com/m_gharbi/status/1904213903384695280

“do not sleep on RF-DETR 🔥 it’s a sota, real-time and most importantly, open-source object detector released by @roboflow @skalskip92 my vibe tests on very noisy images and videos have passed 👏 currently being integrated to transformers 🤗 https://x.com/mervenoyann/status/1905318925325173000

“We’re excited to announce Gigapixel v8.3.0 with the world’s fastest diffusion model for high-res image restoration, Recover v2. See thread for release details. https://x.com/topazlabs/status/1902742856512446490

Dereflection Any Image with Diffusion Priors and Diversified Data https://abuuu122.github.io/DAI.github.io/

“pro tip: the new image generation is also available on https://x.com/stevenheidel/status/1904601168317399199

“tried ChatGPT image generation to ghiblify myself because I will use whatever @giffmana & friends cook 🤝🏻 https://x.com/mervenoyann/status/1904812225434362204

FRESA: Feedforward Reconstruction of Personalized Skinned Avatars from Few Images https://rongakowang.github.io/fresa/fresa.html

“Native GPT 4o image generation: https://x.com/gdb/status/1904601537487270243

ChatAnyone: Stylized Real-time Portrait Video Generation with Hierarchical Motion Diffusion Model https://humanaigc.github.io/chat-anyone/

“Took this photo of me, @jayspiel_ and @stephendhughes and Ghiblified it with @OpenAI image gen, and it came out perfect minus one detail. See it? https://x.com/raizamrtn/status/1904714762027753633

“yeah confirmed 4o image is autoregressive https://x.com/swyx/status/1904660433203871845

“Native GPT 4o image generation, welcome to llama park https://x.com/_akhaliq/status/1904719228675961014

“tremendous alpha with images in chatgpt rn” / X https://x.com/sama/status/1904917997116153929

“✨ Excited to share QVQ-Max, our visual reasoning model that’s still evolving We’ve been experimenting with this approach for a while – try it out on Qwen Chat! (https://t.co/FmQ0B9tiE7) 🚀 Just upload any image or video, ask away, and hit the “Thinking” button to see how it https://x.com/Alibaba_Qwen/status/1905342260100956210

“the new version of images in chatgpt is still rolling out, so please try again later today if you dont get a great one :)” / X https://x.com/sama/status/1904601193504268680

[2503.20595] Diffusion Counterfactuals for Image Regressors https://arxiv.org/abs/2503.20595

Paper page – LHM: Large Animatable Human Reconstruction Model from a Single Image in Seconds https://huggingface.co/papers/2503.10625

“Today’s text-to-image models are essentially that—random slice-of-the-world generator. There’s no intelligence. This is both a data and representation problem. We need to leverage the equivalent of full documents for images, but we don’t have a good representation for it. (3/4)” / X https://x.com/Taesung/status/1904220827073257483

“Introducing Together Chat! Use DeepSeek R1 (hosted in North America) & other top open source models to do web search, coding, image generation, & image analysis. Available today for free! https://x.com/togethercompute/status/1904204860217500123

“it’s super fun seeing people love images in chatgpt. but our GPUs are melting. we are going to temporarily introduce some rate limits while we work on making it more efficient. hopefully won’t be long! chatgpt free tier will get 3 generations per day soon.” / X https://x.com/sama/status/1905296867145154688

“🛳️Rolling out interactive Mindmaps in NotebookLM! I’m so inspired by the Exploratorium here in SF – What if every notebook generated your own personal set of interactive understanding toys that help you learn through play? What if instead of text or images, Notebook could even https://x.com/tokumin/status/1902251588925915429?s=46

“images in chatgpt are wayyyy more popular than we expected (and we had pretty high expectations). rollout to our free tier is unfortunately going to be delayed for awhile.” / X https://x.com/sama/status/1905000759336620238

“4o image generation has arrived. It’s beginning to roll out today in ChatGPT and Sora to all Plus, Pro, Team, and Free users. https://x.com/OpenAI/status/1904602845221187829

People are using Google’s new AI model to remove watermarks from images | TechCrunch https://techcrunch.com/2025/03/17/people-are-using-googles-new-ai-model-to-remove-watermarks-from-images/

“The funny thing about multimodal image generation as released in the last week by Google and OpenAI is that now LLM image generation works like how most people using LLMs for the past two years always thought LLM image generation works. https://x.com/emollick/status/1904703435854799306

“It may surprise you, but there are at least three other art styles besides Studio Ghibli. Now is the time for people who actually know something of our collective cultural history, since they can do magic. I wrote this two years ago, more true now. https://x.com/emollick/status/1904925262006939859

“obligatory studio ghibli-fied pfp lol https://x.com/iScienceLuvr/status/1904842046244024540

“using moondream to hide all ghibli posting from the timeline” / X https://x.com/vikhyatk/status/1904972748927246683

OpenAI’s viral Studio Ghibli moment highlights AI copyright concerns | TechCrunch https://techcrunch.com/2025/03/26/openais-viral-studio-ghibli-moment-highlights-ai-copyright-concerns/

“ByteDance just announced InfiniteYou available on Hugging Face Flexible Photo Recrafting While Preserving Your Identity https://x.com/_akhaliq/status/1902937194198700280

ByteDance/InfiniteYou · Hugging Face https://huggingface.co/ByteDance/InfiniteYou

Trending

Discover more from Ethan B. Holland

Subscribe now to keep reading and get access to the full archive.

Continue reading