Image created with gemini-2.5-flash-image with claude-sonnet-4-5. Image prompt: Photorealistic architectural photography of six Ionic limestone columns on Mizzou’s quad topped with a classical entablature, the frieze carved with ‘INTERNATIONAL’ in Roman serif capitals flanked by bas-relief stone panels depicting national flags as classical relief sculpture, late afternoon golden hour light, warm beige limestone texture, red brick buildings and green lawn background, wide landscape composition with clear blue sky.
@Kimi_Moonshot Congratulations to the entire Moonshot team — today is a great day for open source everywhere. We’re excited to continue supporting Kimi models with fast inference on Baseten. https://x.com/basetenco/status/1986494013109903362
@QuixiAI @Kimi_Moonshot a single H200 node is enough😃”” / X https://x.com/vllm_project/status/1986626058897269070
📢 New Model(s) Drop: Kimi K2 Thinking and Kimi K2 Thinking Turbo are now on Yupp! This pair of thinking models from @Kimi_Moonshot specialize in deep reasoning tasks. We explored their capabilities with some prompts on Yupp: https://x.com/yupp_ai/status/1986469027997491422
🚀 Hello, Kimi K2 Thinking! The Open-Source Thinking Agent Model is here. 🔹 SOTA on HLE (44.9%) and BrowseComp (60.2%) 🔹 Executes up to 200 – 300 sequential tool calls without human interference 🔹 Excels in reasoning, agentic search, and coding 🔹 256K context window Built https://x.com/Kimi_Moonshot/status/1986449512538513505
🚨 New Open Source Model Update! Touted for its reasoning and coding strengths, Kimi K2 Thinking by @Kimi_Moonshot is now live for both Text and WebDev in Battle, Side by Side and Direct. Bring your toughest prompts! 💪 The last time Kimi K2 was in the Arena with a new model, https://x.com/arena/status/1986482438768673107
5 Thoughts on Kimi K2 Thinking – by Nathan Lambert https://www.interconnects.ai/p/kimi-k2-thinking-what-it-means
70% on SWE bench verified 30% terminal bench those are two intuitive thresholds for “”actually useful and not frustrating”” coding assistant. Kimi k2 thinking got 71.3% on SWE-Bench Verified 47.1% on Terminal-Bench”” / X https://x.com/andrew_n_carr/status/1986538323876454461
Congrats to the Kimi K2 team on the great numbers on our SWE-bench Verified, SWE-bench Multilingual and SciCode benchmarks!! https://x.com/OfirPress/status/1986475891158040760
Kimi AI – Kimi K2 is Live https://www.kimi.com/
Kimi API is barely alive right now kinda slow ~20tks/s and get quite a few timeouts / network errors when I let the model reason for a long time”” / X https://x.com/scaling01/status/1986476278908920061
Kimi K2 Thinking feels like a big milestone for open-source AI. The first time in a while that open-source gets ahead of proprietary APIs on their big area of focus (agents). Fun to see that it’s happening at a time when the proprietary APIs have the most money/attention”” / X https://x.com/ClementDelangue/status/1986833436607160600
Kimi K2 Thinking https://moonshotai.github.io/Kimi-K2/thinking.html
Kimi K2 Thinking is now available in anycoder https://x.com/_akhaliq/status/1986468663600337125
Kimi K2 Thinking is the new leading open weights model: it demonstrates particular strength in agentic contexts but is very verbose, generating the most tokens of any model in completing our Intelligence Index evals @Kimi_Moonshot’s Kimi K2 Thinking achieves a 67 in the https://x.com/ArtificialAnlys/status/1986911675820446013
Kimi K2 Thinking just launched on Product Hunt! 🥳 Not chasing votes, just using PH as a clean milestone log for our model updates. 🙂 Huge thanks to the helpful team from @ProductHunt https://x.com/crystalsssup/status/1986714377983304137
Kimi-K2 is an exceptional base model GPQA Diamond 77% GPT-4.5 only got 71.4%”” / X https://x.com/scaling01/status/1986112227875954967
Kimi-K2 Reasoning is coming very soon just got merged into VLLM LETS FUCKING GOOOO im so hyped im so hyped im so hyped https://x.com/scaling01/status/1986071916541870399
Kimi-K2 reasoning is landing soon; it just got merged into vLLM https://x.com/cedric_chee/status/1986073808672067725
Kimi-K2 Thinking ranking 19th on SimpleBench improving Kimi-K2s score from 26.3% (rank 33) to 39.6% This makes it the 3rd best open-source model on SimpleBench. Other chinese open-source models like DeepSeek R1 0528 and DeepSeek V3.1 beat it by roughly 1 %. https://x.com/scaling01/status/1986846212050362510
Live in Cline: kimi-k2-thinking https://x.com/cline/status/1986512739490275680
MoonshotAI has released Kimi K2 Thinking, a new reasoning variant of Kimi K2 that achieves #1 in the Tau2 Bench Telecom agentic benchmark and is potentially the new leading open weights model Kimi K2 Thinking is one of the largest open weights models ever, at 1T total parameters https://x.com/ArtificialAnlys/status/1986541785511043536
moonshotai/Kimi-K2-Thinking · Hugging Face https://huggingface.co/moonshotai/Kimi-K2-Thinking
ollama run kimi-k2-thinking:cloud Kimi K2 Thinking is Moonshot AI’s best open-source thinking model. Try it on Ollama’s cloud! https://x.com/ollama/status/1986640693108863271
Our first research paper: custom Mixture-of-Experts (MoE) kernels that make deployment of trillion-parameter models like Kimi K2 viable for the first time on AWS EFA https://x.com/AravSrinivas/status/1986106660386222592
Unsurprisingly, Kimi K2 Thinking is already number one trending on HF. The AI frontier is open-source! https://x.com/ClementDelangue/status/1986827413532057712
🚀 Day 0 support: Kimi K2 Thinking now running on vLLM! In partnership with @Kimi_Moonshot, we’re proud to deliver official support for the state-of-the-art open thinking model with 1T params, 32B active. Easy deploy in vLLM (nightly version) with OpenAI-compatible API: What https://x.com/vllm_project/status/1986455911066706160
It even compares with GPT 5 Pro on some benches. Looks like Kimi’s interpretation of Pro mode is 8 samples + self reflection https://x.com/nrehiew_/status/1986453238552666320
From my tests, Kimi K2 thinking is better than everything Xai, Anthropic, Google has to offer atm. The only thing that is better than this is Gpt 5 codex (at code) and Gpt 5 pro (at high level algorithm design) It beats the SOTA at creative writing by a mile. Good work”” / X https://x.com/karmay007/status/1986454592809529493
IndQA is a new benchmark designed to evaluate how well AI systems understand culture, context and history to answer questions that matter to people in India. With 2278 questions created in partnership with 250+ experts, IndQA dives deep into reasoning about everyday life,”” / X https://x.com/snsf/status/1985719755551158754
“ChatGPT-o1 & DeepSeek-R1, achieved diagnostic accuracy up to 93.75%. For context, this figure approaches the 96% accuracy benchmark reported for primary care physicians on the same vignette set” Except they told folks to get urgent care too often. Not unexpected given alignment”” / X https://x.com/emollick/status/1985164511947682070
China issues 50% electricity subsidies for datacenters. Very clever. Their industrial energy costs are already (generally) below the US and won’t spike due to such trifle as datacenters, but their chips are far less efficient. With this they get to > Hopper levels of FLOPs/$. https://x.com/teortaxesTex/status/1985540154065318157
Microsoft’s $15.2 billion USD investment in the UAE – Microsoft On the Issues https://blogs.microsoft.com/on-the-issues/2025/11/03/microsofts-15-2-billion-usd-investment-in-the-uae/
Telekom and NVIDIA building a $1.1B datacenter in Munich with 10k GPUs including DGX B200 and RTX PRO Servers https://x.com/scaling01/status/1985741851991621712
Nvidia’s Jensen Huang: ‘China is going to win the AI race,’ FT reports | Reuters https://www.reuters.com/world/asia-pacific/nvidias-jensen-huang-says-china-will-win-ai-race-with-us-ft-reports-2025-11-05/
Nvidia’s Jensen Huang: ‘China is going to win the AI race,’ FT reports https://finance.yahoo.com/news/nvidias-jensen-huang-says-china-211900769.html
Ant AQ-Team @AQ_MedAI @TheInclusionAI and SGLang RL Team @sgl_project just helped land Kimi-K2-Instruct RL on slime — fully wired up and running on 256× H20 141GB 🚀 Huge shout-out to @yngao016, @menlzy, @Yonah_x from AQ Team and @Ji_Li_233, @Yefei_RL from the SGLang RL Team for”” / X https://x.com/slime_framework/status/1986811354502906304
Fixed the token generation speed on https://x.com/Kimi_Moonshot/status/1986754111992451337
Here’s the command I ran: “` mlx.launch –hosts first.ip,second.ip –env MLX_METAL_FAST_SYNCH=1 mlx-lm/mlx_lm/examples/pipeline_generate.py –model mlx-community/Kimi-K2-Thinking –prompt “”Write an HTML and JavaScript page implementing space invaders”” -m 16384 “` PR here”” / X https://x.com/awnihannun/status/1986602098017116357
The new 1 Trillion parameter Kimi K2 Thinking model runs well on 2 M3 Ultras in its native format – no loss in quality! The model was quantization aware trained (qat) at int4. Here it generated ~3500 tokens at 15 toks/sec using pipeline-parallelism in mlx-lm: https://x.com/awnihannun/status/1986601104130646266
Gemini 3.0 Ultra or Gemini 3.0 Pro? Which is it and why do you think that? It sounds too big for Pro but too small for Ultra, but since models just get sparser and sparser I believe it’s Pro and very similar to Kimi-K2 1.2T@30B. Also as of right now Ultra is still just a”” / X https://x.com/scaling01/status/1986161974883860486
Team from Ant Group @TheInclusionAI helped land Kimi model @Kimi_Moonshot on @Zai_org’s slime framework! Open AIs help Open AIs ♥️ CN AIs help CN AIs ♥️”” / X https://x.com/bigeagle_xd/status/1986815075785879723
New offices in Paris and Munich expand Anthropic’s European presence \ Anthropic https://www.anthropic.com/news/new-offices-in-paris-and-munich-expand-european-presence
We’re announcing a partnership with Iceland’s Ministry of Education and Children to bring Claude to teachers across the nation. It’s one of the world’s first comprehensive national AI education pilots: https://x.com/AnthropicAI/status/1985612560255893693
We’ve released an early preview of Qwen3-Max-Thinking–an intermediate checkpoint still in training. Even at this stage, when augmented with tool use and scaled test-time compute, it achieves 100% on challenging reasoning benchmarks like AIME 2025 and HMMT. You can try the https://x.com/Alibaba_Qwen/status/1985347830110970027
Amazing work by @RidgerZhu and the ByteDance Seed team — Scaling Latent Reasoning via Looped LMs introduces looped reasoning as a new scaling dimension. 🔥 The Ouro model is now runnable on vLLM (nightly version) — bringing efficient inference to this new paradigm of latent”” / X https://x.com/vllm_project/status/1985695123469209703
ByteDance released BindWeave Subject-Consistent Video Generation via Cross-Modal Integration https://x.com/_akhaliq/status/1986058046876070109
Australians have been promised three free hours of solar power a day. Here’s what you need to know | Energy | The Guardian https://www.theguardian.com/environment/2025/nov/04/australia-free-solar-power-scheme-how-when-houshold-bills
How Russia derailed China’s rise for over a century: new lecture and Q&A w Sarah Paine. IMO the Chinese Civil War is 1 of the top 3 most important events of the 20th century. To understand why it transpired as it did, you need to understand Stalin’s role in the whole thing. https://x.com/dwarkesh_sp/status/1984282136816321021
How much RAM do you need to run tiny models? Jamba Reasoning 3B runs on just 2.25 GiB, the lightest among small models like Qwen (@Alibaba_Cloud), Llama (@Meta), Granite (@IBM), and Gemma (@GoogleDeepMind). 👉 Try Jamba Reasoning 3B yourself: https://x.com/AI21Labs/status/1986439953539076169
Check out Mathis Felardos and Mickaël Seznec’s amazing discussion of how @MistralAI uses P/D disaggregation to optimize their @vllm_project deployment https://x.com/robertshaw21/status/1986868946071327038
#1 on the MTEB multilingual leaderboard”” / X https://x.com/fdaudens/status/1984541314063446191
The Department of Commerce has allowed Microsoft to ship NVIDIA GPU’s to the UAE for the first time. Brad Smith announced this today in Abu Dhabi. He said MS received the license in September, and will spend $7.9 billion on datacenters in the UAE over the next four years. https://x.com/AndrewCurran_/status/1985325278823125483
Individual releases of open AI models only matter in the short term. These models becomes obsolete without continued releases (look at Llama versus newer Chinese models), because the capability/cost improvement curve is steep and you don’t want to use an older model forever. https://x.com/emollick/status/1984993332251263061
🎉 Now you can easily use Qwen3-VL in Jan!”” / X https://x.com/Alibaba_Qwen/status/1985542635373937102
API usage for Qwen3-Max-Thinking-Preview: Model name: qwen3-max-preview Parameter setting: enable_thinking=True https://x.com/Alibaba_Qwen/status/1985586316197937256
Thread by @Alibaba_Qwen on Thread Reader App – Thread Reader App https://threadreaderapp.com/thread/1985347830110970027.html
adding camera control to the list of things Qwen Image Edit is great at + with a specialized multi-angle LoRA it’s even better✨ > rotate the camera > tilt between bird’s-eye and worm’s-eye views > adjust lens (wide, close-up) of course we built a demo for it 🤝📹 https://x.com/linoy_tsaban/status/1986090375409533338
Qwen Image Multiple Angles LoRA is an exquisitely trained LoRA! 📐˚₊‧꒰ა Keep character and scenes consistent, and flies the camera around! Open source got there! One of the best LoRAs I’ve come across lately 🙌 https://x.com/multimodalart/status/1986174924038218087
Qwen3-VL Accuracy Differences on Ollama vs MLX Video: https://x.com/andrejusb/status/1985612661447331981
Hybrid models like Qwen3-Next, Nemotron Nano 2 and Granite 4.0 are now fully supported in vLLM! Check out our latest blog from the vLLM team at IBM to learn how the vLLM community has elevated hybrid models from experimental hacks in V0 to first-class citizens in V1. 🔗 https://x.com/PyTorch/status/1986192579835150436
Reminder that Huawei SuperPoDs with plans to scale to gigawatt-class, 1M-device superclusters by Q4 2027 are dedicated to DeepSeek. Regular thanks to the Whale for inspiring so many. R1 will be one of the most significant events in the history of tech, close to ChatGPT 3.5. https://x.com/teortaxesTex/status/1985567870227460166





Leave a Reply