Image created with gemini-3.1-flash-image-preview with claude-sonnet-4-5. Image prompt: Using the provided reference image, preserve the exact crate construction with horizontal dark reddish-brown wooden slats, peeling paint, iron hardware, and three-panel layout with hand-painted black stencil lettering; replace the original text with ‘Local’ in the same thick brushstroke style; place the crate on a weathered wooden platform at a rural depot or roadside stop with regional architectural details barely visible in soft background blur, early spring mud and first green shoots around the platform edge, warm raking sunlight emphasizing wood grain texture.

Huh. I am not sure distilling Gemini models to run on phones is going to result in the generally capable agents that people will soon expect, but we shall see.
https://x.com/emollick/status/2036845759283220546

When @karpathy built MenuGen (https://t.co/E6yaFk3hfu), he said: “”Vibe coding menugen was exhilarating and fun escapade as a local demo, but a bit of a painful slog as a deployed, real app. Building a modern app is a bit like assembling IKEA future. There are all these services,
https://x.com/patrickc/status/2037190688950161709?s=20

Is the Future of AI Local? | Tom Bedor’s Blog https://tombedor.dev/open-source-models/

Local AI is free, fast & secure! So today we’re introducing hf-mount: attach any storage bucket, model or dataset from @huggingface as a local filesystem. This is a game changer, as it allows you to attach remote storage that is 100x bigger than your local machine’s disk. This
https://x.com/ClementDelangue/status/2036452081750409383

Now available on Hugging Face: hf-mount 🧑‍🚀 The team really cooked, still wrapping my head everything possible but you can do things like: – mount a 5TB dataset as a local folder and query only the parts you need with DuckDB (✅ works) – browse any model repo with ls/cat like
https://x.com/victormustar/status/2036476453370380416

WebGPU is INSANE! 🤯 Here’s a 24B parameter model running locally in a web browser, at a blazing ~50 tokens/second on my M4 Max. ⚡️ It’s the largest model we’ve ever run with Transformers.js… and we’re not stopping here. Big announcement soon.
https://x.com/xenovacom/status/2036908326462665211

Leave a Reply

Trending

Discover more from Ethan B. Holland

Subscribe now to keep reading and get access to the full archive.

Continue reading