Image created with Flux Pro v1.1 Ultra. Image prompt: Giant “100” as pure white negative‑space cutout dominating the frame; minimalist poster style; image frame, waveform, and text block icons interlinking inside the zeros; prism‑spectrum accents on slate backdrop; high contrast, crisp edges, soft studio light, no other text, no logos
Blue (@heyBlueX) lets you control your phone’s apps by voice so tasks actually get finished, hands-free. It handles messages, email, and actions across apps by tapping and typing as you would. https://x.com/ycombinator/status/1958182627422146811
ByteDance just opensourced a desktop automation AI Agent. This agent can use any desktop app, open files, and browse websites using vision models running locally. 100% Free, Opensource, and Local. https://x.com/unwind_ai_/status/1956538069311500514
Google Translate adds live translation and language learning https://blog.google/products/translate/language-learning-live-translate/
My favorite demo of the new gpt-realtime model from @matthieulc — Shoggoth Mini using Realtime API with image input https://x.com/pbbakkum/status/1961120041799487654
Harvard dropouts to launch ‘always on’ AI smart glasses that listen and record every conversation | TechCrunch https://techcrunch.com/2025/08/20/harvard-dropouts-to-launch-always-on-ai-smart-glasses-that-listen-and-record-every-conversation/
A smarter way to talk to your TV: Microsoft Copilot launches on Samsung TVs and monitors | Microsoft Copilot Blog https://www.microsoft.com/en-us/microsoft-copilot/blog/2025/08/27/a-smarter-way-to-talk-to-your-tv-microsoft-copilot-launches-on-samsung-tvs-and-monitors/
Google’s NotebookLM updates Audio and Video Overviews https://blog.google/technology/google-labs/notebook-lm-audio-video-overviews-more-languages-longer-content/
🎬 Introducing MuseSteamer—an image-to-video model newly launched at Baidu AI Day. Here’s what it offers: – Generation of synchronized visuals, Chinese dialogue, and sound effects from a single image. – Creation of 10-second, 1080P cinematic-quality clips with nuanced facial https://x.com/Baidu_Inc/status/1941151611755393210
🤖 From this week’s issue: Liquid AI released LFM2-VL, their first series of vision-language foundation models. https://x.com/dl_weekly/status/1960387356889928174
Liquid AI Releases LFM2-VL: Super-Fast, Open-Weight Vision-Language Models Designed for Low-Latency and Device-Aware Deployment – MarkTechPost https://www.marktechpost.com/2025/08/20/liquid-ai-releases-lfm2-vl-super-fast-open-weight-vision-language-models-designed-for-low-latency-and-device-aware-deployment/
What if robots could learn new tasks as easily as ChatGPT picks up a new prompt? [📍 Save the project for later] Robots don’t yet have the in-context learning powers of LLMs like ChatGPT or Gemini. Most Vision-Language-Action (VLA) models need retraining or fine-tuning to https://x.com/IlirAliu_/status/1958795845681226024
Tracking drones in the wild is one of the hardest problems in computer vision. 🚁🔥 Small, fast-moving objects. Changing backgrounds. Real-time constraints… Most models break under these conditions. That’s why this open-source project from @chesterzelaya, extended by https://x.com/IlirAliu_/status/1959310919147622741




