“🚨 AgentStack v0.3.3 ~ New Framework support! 🦙 @llama_index added as an official framework 🔨 @PaymanAI tool added 🥰 Dev experience improvements 🪳 Bug fixes
https://x.com/braelyn_ai/status/1887958100612952080
“Hugging Face has quietly become the biggest AI app store with 400,000 total apps, 2,000 new apps created every day, getting visited 2.5M times every week! Now you can search through any of them with AI or categories. The future of AI will be distributed, have fun everyone!
https://x.com/ClementDelangue/status/1886861567326650526
“uh it might be over… they put r1 in a loop for 15minutes and it generated: “better than the optimized kernels developed by skilled engineers in some cases”
https://x.com/abacaj/status/1889847093046702180
[R] o3 achieves a gold medal at the 2024 IOI and obtains a Codeforces rating on par with elite human competitors” : r/MachineLearning https://www.reddit.com/r/MachineLearning/comments/1io4c7r/r_o3_achieves_a_gold_medal_at_the_2024_ioi_and/
argilla/FinePersonas-v0.1 · Datasets at Hugging Face https://huggingface.co/datasets/argilla/FinePersonas-v0.1
Hugging Face has quietly become the biggest AI app store with 400,000 total apps, 2,000 new apps created every day, getting visited 2.5M times every week! Now you can search through any of them with AI or categories. The future of AI will be distributed, have fun everyone! https://x.com/ClementDelangue/status/1886861567326650526
Many of you asked for code & weights for π₀, we are happy to announce that we are releasing π₀ and pre-trained checkpoints in our new openpi repository! We tested the model on a few public robots, and we include code for you to fine-tune it yourself. https://x.com/physical_int/status/1886822689157079077
Open R1: Update #2 https://huggingface.co/blog/open-r1/update-2
uh it might be over… they put r1 in a loop for 15minutes and it generated: better than the optimized kernels developed by skilled engineers in some cases” https://x.com/abacaj/status/1889847093046702180
Democratize Intelligence https://www.demi.so/
DeepScaleR: Surpassing O1-Preview with a 1.5B Model by Scaling RL https://pretty-radio-b75.notion.site/DeepScaleR-Surpassing-O1-Preview-with-a-1-5B-Model-by-Scaling-RL-19681902c1468005bed8ca303013a4e2
ollama run deepscaler A fine-tuned version of Deepseek-R1-Distilled-Qwen-1.5B that surpasses the performance of OpenAI’s o1-preview with just 1.5B parameters on popular math evaluations. / X https://x.com/ollama/status/1889496833875124735
This is wild – UC Berkeley shows that a tiny 1.5B model beats o1-preview on math by RL! They applied simple RL to Deepseek-R1-Distilled-Qwen-1.5B on 40K math problems, trained at 8K context, then scaled to 16K & 24K. 3,800 A100 hours ($4,500) to beat o1-preview in math! Best https://x.com/Yuchenj_UW/status/1889387582066401461
Introducing deep research | OpenAI https://openai.com/index/introducing-deep-research/
The point of Elon’s offer to buy OpenAI for $97.4B to try to fuck over the non-profit to for-profit conversion A standing offer from many accredited investors to buy the non-profit portion OAI prolly tanks the argument to IRS/states Which then maybe kills future fundraising” / X https://x.com/dylan522p/status/1889128785687236769
This paper is wild – a Stanford team shows the simplest way to make an open LLM into a reasoning model. They used just 1,000 carefully curated reasoning examples & a trick where if the model tries to stop thinking, they append Wait” to force it to continue. Near o1 at math. https://x.com/emollick/status/1887696014829641983
Team Says They’ve Recreated DeepSeek’s OpenAI Killer for Literally $30 https://futurism.com/researchers-deepseek-even-cheaper
We haven’t reached full potential yet. DeepScaleR is a 1.5B parameter model fine-tuned with Reinforcement Learning that surpasses @OpenAI’s O1-preview in math benchmarks, proving RL scaling is effective even for smaller models. Recipe: 0️⃣ Started from https://x.com/_philschmid/status/1889592742088515630
Sam Altman Regrets Ditching Open Source, Says He’s Been on the “Wrong Side of History” https://futurism.com/sam-altman-open-source-wrong-side-history
DeepSeek surpassed OpenAI in GitHub stars for their top 2 projects! 🐋 DeepSeek-R1 cooked openai-cookbook”, in just 3 weeks. A milestone in open-source AI history! https://x.com/Yuchenj_UW/status/1887720580503499181
New open source reasoning model! Huginn-3.5B reasons implicitly in latent space 🧠 Unlike O1 and R1, latent reasoning doesn’t need special chain-of-thought training data, and doesn’t produce extra CoT tokens at test time. We trained on 800B tokens 👇 https://x.com/tomgoldsteincs/status/1888980680790393085
🔥 Video AI is taking over! Out of 17 papers dropped on @HuggingFace today, 6 are video-focused – from Sliding Tile Attention to On-device Sora. The race for next-gen video tech is heating up! 🎬🚀 https://x.com/fdaudens/status/1888858943415361807
@ZyphraAI have launched their first Text to Speech model, Zonos-v0.1 – now the leading open weights Text to Speech model in Artificial Analysis Speech Arena. Zonos-v0.1 is currently scoring a Speech Arena ELO of 1020, just behind proprietary speech models available on Google https://x.com/ArtificialAnlys/status/1889150365913972930
The discussion about Deep Research shows the gaps between fields for what they think “research” is. OpenAI Deep Research is not great at digging up tons of citations or facts (Google’s is better at that) Deep Research is great at digging up arguments, holes & providing analysis https://x.com/emollick/status/1887295987284074580
What we’ve been seeing in the past few weeks is that when you’re really ambitious about AI, when you’re a bit radical in fostering ecosystems with open science and open-source, you can actually play a part in AI —whether you’re in the US, China, Canada, France” — https://x.com/fdaudens/status/1888984421304340604
DeepSeek Gets an ‘F’ in Safety From Researchers https://gizmodo.com/deepseek-gets-an-f-in-safety-from-researchers-2000558645
Daily Wenfeng W: «China’s 3 big telecoms operators rush to integrate DeepSeek models into cloud services, freezing their own LLM projects» -SCMP «If a complete upstream and downstream industrial ecosystem is formed, then there is no need for us to make applications ourselves» https://x.com/teortaxesTex/status/1888812805932875828
The open weights space is already vigorous. Not just Mistral and the various Chinese models, but Meta has been all-in on open models for a long time and seems committed to continuing to hyperscaling new open models at great cost (plus small open models from Microsoft & Google)” / X https://x.com/emollick/status/1888477074064646410
Deepseek is killing it! I’ve long been Team Claude for everything coding but Deepseek blew past Claude in our OSS PR review 81% critical bug to noise ratio with 3.7x more bugs caught! https://x.com/Aiswarya_Sankar/status/1887356821738037742
Arrived in Paris for the AI summit with @IreneSolaiman and team. Let’s push open-source AI! https://x.com/ClementDelangue/status/1888920800528331091
Btw it also beats qwen at MMLU Pro… Wasn’t this complex domain meant to be where decoder models were meant to be required? / X https://x.com/jeremyphoward/status/1889435769959489800
Since launching DeepSeek-R1, we’ve seen a wave of companies looking to deploy reasoning models in production—but scaling them efficiently remains a challenge. Today, we’re expanding beyond our ultra-fast Serverless API with Together Reasoning Clusters: dedicated, secure, https://x.com/togethercompute/status/1889743684977168547
> Why CPU/GPU Hybrid Inference? DeepSeek’s MLA operators are highly computationally intensive. While running everything on CPU is possible, offloading the heavy computations to the GPU results in a massive performance boost. / X https://x.com/teortaxesTex/status/1889531203742466250
DeepSeek: The countries and agencies that have banned the AI company’s tech | TechCrunch https://techcrunch.com/2025/02/03/deepseek-the-countries-and-agencies-that-have-banned-the-ai-companys-tech/
Ai2 says its new AI model beats one of DeepSeek’s best | TechCrunch https://techcrunch.com/2025/01/30/ai2-says-its-new-ai-model-beats-one-of-deepseeks-best/
If You Think Anyone in the AI Industry Has Any Idea What They’re Doing, It Appears That DeepSeek Just Accidentally Leaked Its Users’ Chats https://futurism.com/the-byte/deepseak-leaks-user-chatlogs
🌻 Community Extensions + AutoGen You can create and publish your own open-source extensions for the AutoGen ecosystem, e.g., to implement advanced clients, tools, and teams. Creating extensions is easy. – name the package autogen-” – use common interfaces from core and https://x.com/pyautogen/status/1886676694439993636
What if I told you that, without needing any task specific fine tuning, the encoder-only ModernBERT 0.3b beats Qwen 0.5b at MMLU? This *will* kick off a whole new revolution in language models. / X https://x.com/jeremyphoward/status/1889434481519632505
Let’s goo! Hugging Face just released Open R1 Math – Large scale math reasoning dataset 🔥 > 220K Math problems > Matches DeepSeek R1 7B with less than 25% of the SFT data on Math > 800k raw R1 reasoning traces > Based on Numina Math 1.5 > Generated on 512H100s > Apache 2.0 https://x.com/reach_vb/status/1888994979218915664
News publishers sue Cohere for copyright and trademark infringement https://www.axios.com/2025/02/13/publishers-sue-cohere-ai-copyright
Exposed DeepSeek Database Revealed Chat Prompts and Internal Data | WIRED https://www.wired.com/story/exposed-deepseek-database-revealed-chat-prompts-and-internal-data/
Browser Use UI can do DeepResearch👀 Repo is 100% open source ↓ Thanks @Gradio for powering the UI🔥 https://x.com/gregpr07/status/1887622796337197340
Mistral and Perplexity are both moving to Cerebras. Why? Because we make our customers products 10x faster than their competitors.” / X https://x.com/draecomino/status/1889430107288416340
Open-source DeepResearch – Freeing our search agents https://huggingface.co/blog/open-deep-research
Teortaxes▶️ (DeepSeek 推特🐋铁粉 2023 – ∞) (@teortaxesTex) / X https://x.com/teortaxesTex
🚀 Open-R1’s latest: OpenR1-Math-220k dataset with 220k verified math problems, running on 512 H100s! The Qwen-7B model nears DeepSeek’s performance, while community finds quality reasoning possible with just ~1000 samples. Check it out: https://x.com/fdaudens/status/1889045933490307217
New open source release from Meta FAIR: Audiobox Aesthetics was trained on 562 hours of audio aesthetic data annotated by professional raters across four dimensions to create a model that enables the automatic evaluation of aesthetics for speech, music and sound. https://x.com/AIatMeta/status/1889418249466683449
@rjmacarthy @sama Distribute and run open source models for developers locally! It’s complementary to the hosted openAI models / X https://x.com/ollama/status/1889784880923394257
MobileLLM – a facebook Collection https://huggingface.co/collections/facebook/mobilellm-6722be18cb86c20ebe113e95
Cerebras Launches World’s Fastest DeepSeek R1 Llama-70B Inference – Cerebras https://cerebras.ai/blog/cerebras-launches-worlds-fastest-deepseek-r1-llama-70b-inference
🤔 How’s DeepSeek-R1 usage stacking up against Llama 3 on @huggingface 3 weeks after launch? Just pulled 30-day numbers: DeepSeek: 1K+ derivatives created by the community (10M downloads) 8 original models (5M downloads) Llama 3: 9K+ derivatives created by the community (30M https://x.com/fdaudens/status/1888615746789367915
Introducing OpenR1-Math-220k! https://x.com/_lewtun/status/1889002019316506684




