Image created with gemini-2.5-flash-image with claude-sonnet-4-5. Image prompt: Black and white photograph of fast-moving stratus clouds creating horizontal motion blur across entire frame, sharp bold sans-serif title card reading INTERNET anchored in lower third, high contrast between blurred cloud streams and crisp typography, cinematic composition suggesting endless data flow and network connectivity, film grain texture
To measure these capabilities, we’re open-sourcing DeepSearchQA, a new benchmark to evaluate agents on complex web search tasks. Deep Research achieves state-of-the-art performance on this benchmark, as well as on the full Humanity’s Last Exam set (reasoning & knowledge), and https://x.com/GoogleDeepMind/status/1999165706231820297
GLM-4.6V can read my horrendous hand writing and explain the math correctly Really loving this model, how well it does tool calling, how many languages it knows and its visual accuracy. https://x.com/0xSero/status/1998328482930073887
GLM-4.6V is out. This is new vision language model from @Zai_org – it’s a MOE with 12B active parameters and 106B total. – there’s a leaner variant with 9B – context lengths are 128k – it has native multimodal function calling Should be perfect for agentic tasks like browser https://x.com/ben_burtenshaw/status/1998019922664865881
GLM-4.6V just dropped on Hugging Face https://x.com/_akhaliq/status/1998052965597241647
GLM-4.6V: Open Source Multimodal Models with Native Tool Use https://z.ai/blog/glm-4.6v
Shopify Editions | Winter ’26 https://www.shopify.com/editions/winter2026#shopify-simgym-app
Shopify merchants can now sell products through AI chatbots | BetaKit https://betakit.com/shopify-merchants-can-now-sell-products-through-ai-chatbots/
We’re rolling out Product Network: merchants can now sell each other’s products with zero integration work. LLMs analyze storefronts and buyer behavior, find products that fit naturally, and place them right on the page. Crucially, shoppers buy *without* ever leaving your site. https://x.com/MParakhin/status/1998789844794012049
We’ve developed the FACTS Benchmark Suite with @GoogleResearch. 📊 It’s the industry’s first comprehensive test evaluating LLM factuality across four dimensions: internal model knowledge, web search, grounding, and multimodal inputs. https://x.com/GoogleDeepMind/status/1998831084277313539
New Google web ecosystem tools and partnerships https://blog.google/products-and-platforms/products/search/tools-partnerships-web-ecosystem/
Google Online Security Blog: Architecting Security for Agentic Capabilities in Chrome https://security.googleblog.com/2025/12/architecting-security-for-agentic.html
GLM-4.6V has day zero support on MLX-VLM 🚀 Quants uploading to the hub. PS: Install from source because they changed vision model type https://x.com/Prince_Canuma/status/1998024143212851571
I tested multimodal tool calling with GLM-4.6V on HuggingChat, it works very well! https://x.com/mervenoyann/status/1998405366313345295
interesting how some benchmark doesn’t seems to get huge boost between glm-4.6V and the flash version (which is ONLY 9B dense compare to 106B A12B MoE)”” / X https://x.com/eliebakouch/status/1998015034979389563
A visual editor for the Cursor Browser · Cursor https://cursor.com/blog/browser-visual-editor
Bringing More Real-Time News and Content to Meta AI https://about.fb.com/news/2025/12/bringing-more-real-time-news-and-content-to-meta-ai/





Leave a Reply