Image created with Flux Pro v1.1 Ultra. Image prompt: Giant “100” as pure white negative‑space cutout dominating the frame; minimalist poster style; community “model cards” shelves and a subtle friendly‑smile motif inside the zeros; warm yellow backdrop; high contrast, crisp edges, soft studio light, no other text, no logos
Grok 2 from @xai has just been released on @huggingface: https://x.com/ClementDelangue/status/1959356467959439464
Grok-2 has been “”open sourced”” but has one of the worst licenses of any recent major open weights release. Given that it’s already quite outdated by the time they’ve got around to releasing it, combined with the license, this will see little use. It’s dead on arrival. https://x.com/xlr8harder/status/1959490601264533539
Pretty cool that they open sourced the actual full-sized production model. Here’s the Grok 2.5 architecture overview next to a roughly similarly sized Qwen3 model. The MoE residual is quite interesting. Kind of like a shared expert. I don’t think I’ve seen this setup before. https://x.com/rasbt/status/1959643038268920231
The @xAI Grok 2.5 model, which was our best model last year, is now open source. Grok 3 will be made open source in about 6 months. https://x.com/elonmusk/status/1959379349322313920
xAI just released Grok 2 on Hugging Face. This massive 500GB model, a core part of xAI’s 2024 work, is now openly available to push the boundaries of AI research. https://x.com/HuggingPapers/status/1959345658361475564
xai-org/grok-2 · Hugging Face https://huggingface.co/xai-org/grok-2
Grok now has a model card – which is a big step forward! But it is light on details, with unexplained results. Some examples: if the MASK measurement is the same as in the source paper, .43 would be a fairly high level of deception, also the sycophancy score is hard to interpret https://x.com/emollick/status/1959116132096336066
OpenAI just released HealthBench on Hugging Face. This new dataset is designed for rigorously evaluating large language models’ capabilities in improving human health. A vital step for AI in medicine! https://x.com/HuggingPapers/status/1960749923218895332
microsoft is dropping (still uploading) VibeVoice-1.5B model on @huggingface! i love the multi-speaker conversational audio feature for podcasts! https://x.com/MaziyarPanahi/status/1959994276198351145
microsoft/VibeVoice-1.5B · Hugging Face https://huggingface.co/microsoft/VibeVoice-1.5B
VibeVoice A Frontier Open-Source Text-to-Speech Model https://x.com/_akhaliq/status/1960106923191140373
VibeVoice is a framework from @MSFTResearch for generating expressive, long-form, multi-speaker audio conversations. Create podcasts from text. MIT licensed🔥 Synthesize speech up to 90 minutes long with 4 distinct speakers 🤯 https://x.com/Gradio/status/1960023019239133503
Try the model on our platform, deploy it privately, or for research use on @huggingface. Find more details in our blog: https://x.com/cohere/status/1961081787674763525
🔔 Two months ago, we released #IneqMath, which revealed the Soundness Gap: LLMs can guess answers to Olympiad-level inequalities problems, but still struggle to make rigorous proof steps. Since then, it’s been downloaded 4K+ times on HuggingFace! ➡️ https://x.com/lupantech/status/1960384184842879444
microsoft (Microsoft) https://huggingface.co/microsoft
Also first time i hear about this South Korean company, didn’t get the attention it deserve imo 👀 paper https://x.com/eliebakouch/status/1959598956540755984
Wow, pretty cool that they also open sourced a FSDP2 compatible Muon and PolyNorm working with @huggingface kernels! https://x.com/eliebakouch/status/1959652478422536611




