Image created with Ideogram 3.0. Image prompt: Lower-East-Side street-corner photograph reminiscent of a late-80s album cover: weathered red-brick tenement with exterior fire-escapes, canvas awning shading racks of vintage clothes; above the awning, a hand-painted board reads ‘Local SPORTSWEAR’; a hanging blade sign in cursive script reads ‘Local Boutique’; a deli flyer screaming ‘Shop Local’ is taped to the glass door; warm golden-hour light, subtle 35mm film grain, muted yet punchy color palette, gritty NYC vibe.
thrilled to be partnering with jony, imo the greatest designer in the world. excited to try to create a new generation of AI-powered computers. https://x.com/sama/status/1925242282523103408?s=46&t=b7l37rB6wtbyAh6ah1NpZQ
Details leak about Jony Ive’s new ‘screen-free’ OpenAI device | The Verge https://www.theverge.com/news/672357/openai-ai-device-sam-altman-jony-ive
Introducing Gemma 3n, our multimodal model built for mobile on-device AI. 🤳 It runs with a smaller memory footprint, cutting down RAM usage by nearly 3x – enabling more complex applications right on your phone, or for livestreaming from the cloud. Now available in early https://x.com/GoogleDeepMind/status/1925916216083779774
Apple to Open AI Models to Developers in Bid to Spur New Apps – Bloomberg https://www.bloomberg.com/news/articles/2025-05-20/apple-to-open-ai-models-to-developers-betting-that-it-will-spur-new-apps?embedded-checkout=true
Zed just dropped the fastest Agentic code editor built in Rust. Works with Claude Sonnet 3.7, Gemini 2.5 Pro and local models via Ollama. 100% opensource. https://x.com/Saboo_Shubham_/status/1921754009221906848
Google I/O 2025: Gemini on Android XR coming to glasses, headsets https://blog.google/products/android/android-xr-gemini-glasses-headsets/
Wow, @jandotai is now Apache licensed – big win for on device community! 🔥 Way to go team! https://x.com/reach_vb/status/1925475572219568269
Apple will reportedly open up its local AI models to third-party apps | The Verge https://www.theverge.com/news/670868/apple-intelligence-ai-third-party-developer-access-model
Looking back at the original stats for the still-not-released AFM model for the still-not-released Apple Intelligence, it is striking that Microsoft released (9 months ago!) an open model that runs on an iPhone & which not only beats the phone version of AFM, but the server one! https://x.com/emollick/status/1922847054750974030
Stability AI and Arm Collaborate to Release Stable Audio Open Small, Enabling Real-World Deployment for On-Device Audio Generation — Stability AI https://stability.ai/news/stability-ai-and-arm-release-stable-audio-open-small-enabling-real-world-deployment-for-on-device-audio-control
Stability AI open-sourced Stable Audio Open Small, a text-to-audio AI —341M-parameters —Generates 11s of audio, including drum loops, foley, riffs, and textures —Optimized for Arm-based consumer devices https://x.com/adcock_brett/status/1924133939376996539
Today we’re open-sourcing Stable Audio Open Small, a 341M-parameter text-to-audio model optimized to run entirely on @Arm CPUs. This means 99% of smartphones can now generate music-production samples in seconds, right on-device with no internet required. Built for fast, https://x.com/StabilityAI/status/1922675163411497094
Google’s Gemini expands to cars, TVs, smartwatches, and headsets https://www.therundown.ai/p/googles-gemini-ai-expands-across-devices
google/shieldgemma-2-4b-it · Hugging Face https://huggingface.co/google/shieldgemma-2-4b-it
Gemma keeps delivering! I’m very excited to share with you the most recent LMArena results for the Gemma 3 family💥 Gemma 3 stays as the best open model that can run on a single GPU. And stay tuned, more to come! https://x.com/osanseviero/status/1923159046900548014
Two awesome new MLX + Hugging Face hub integrations. It’s easier than ever to get started running models locally: https://x.com/awnihannun/status/1924512714287939816
From the Hub to your Mac with MLX 🪄”” / X https://x.com/fdaudens/status/1924856908302692362
Analog Foundation Models “”In this work, we introduce a general and scalable method to robustly adapt LLMs for execution on noisy, low-precision analog hardware. Our approach enables state-of-the-art models – including Phi-3-mini-4k-instruct and Llama-3.2-1B-Instruct to – https://x.com/iScienceLuvr/status/1923269433751158884
Sam & Jony introduce io https://x.com/OpenAI/status/1925235156157440438
Avoid Excessive Reasoning Excessive reasoning hurts smaller models on simple tasks. A too-long prompt can negatively impact smaller models on basic reactive tasks, while larger models show more robust behaviour. Adding reflections or plans dilutes key information and often”” / X https://x.com/omarsar0/status/1924182835289620950




