Image created with gemini-2.5-flash-image with claude-sonnet-4-5. Image prompt: Cinematic wide shot of a futuristic orbital command center interior, thousands of floating green-lit GPU circuit boards arranged in geometric formations like a tactical fleet, glowing data streams connecting them, zero gravity environment, deep space visible through massive windows, cool blue and green lighting with dramatic rim lights, sleek military sci-fi aesthetic, epic scale and isolation, Ender’s Game inspired strategic control room
SoftBank’s Nvidia sale rattles market, raises questions | TechCrunch https://techcrunch.com/2025/11/11/softbanks-nvidia-sale-rattles-market-raises-questions/
Nvidia’s Jensen Huang: ‘China is going to win the AI race,’ FT reports https://finance.yahoo.com/news/nvidias-jensen-huang-says-china-211900769.html
Baseten used @nvidia Dynamo to double inference speed for long-context code generation and increased throughput by 1.6x. Dynamo simplifies multi-node inference on Kubernetes, helping us scale deployments while reducing costs. Read the full blog post below👇”” / X https://x.com/basetenco/status/1989058852789317717
Scaling Large MoE Models with Wide Expert Parallelism on NVL72 Rack Scale Systems | NVIDIA Technical Blog https://developer.nvidia.com/blog/scaling-large-moe-models-with-wide-expert-parallelism-on-nvl72-rack-scale-systems/
☁️ NVIDIA Dynamo is now available across major cloud providers–including @awscloud, @googlecloud, @Azure, and @OracleCloud–to enable efficient multi-node inference on Kubernetes in the cloud. And It’s already delivering results: @basetenco is seeing faster, more cost-effective https://x.com/NVIDIADC/status/1989005718083518866
Besides A100s still being in use, I actually expect H100s to have a much longer lifespan than even the A100s From V100 -> A100 -> H100 we had pretty dramatic changes in each generation to reflect the realities of LLM training. The B200/300s are really nice and definitely push”” / X https://x.com/code_star/status/1988062247818850421
the NVFP4 kernels on Blackwell competition has started on @GPU_MODE!!! the first problem, NVFP4 GEMV is now out and submissions can be made. good luck to all the participants!”” / X https://x.com/a1zhang/status/1987972190898450922
Amid the global consensus discussion about an AI bubble. H100 spot price to increase by 8% in Q4 2025. H200 spot price to increase by 18% in Q4 2025. $NVDA https://x.com/FundaBottom/status/1987905008541831521
How cool is this! @Siemens is breaking new ground with an open-source-first, self-contained LLM platform, optimized by @vllm_project. Learn how they deployed their sustainable AI stack with flexibility, full control, and cost savings at scale: https://x.com/NVIDIAAIDev/status/1987944094883037559
The 4.5 trillion dollar elephant in the room https://stevenadler.substack.com/p/the-45-trillion-dollar-elephant-in
A couple of tier 1 frontier labs are saying that NVIDIA is not taking seriously the potential perf per TCO advantage of MI450X UALoE72 for inference workloads especially when factoring in that AMD is offering up to 10% of AMD shares to OpenAI https://x.com/SemiAnalysis_/status/1988044940149235844
Nvidia presents TiDAR Think in Diffusion, Talk in Autoregression https://x.com/_akhaliq/status/1988963077690438097
🤖 From this week’s issue: A technical blog post explaining how NVIDIA TensorRT-LLM’s Wide Expert Parallelism efficiently scales large Mixture-of-Experts models on GB200 NVL72 systems, achieving significant performance and cost improvements. https://x.com/dl_weekly/status/1987913458654786008





Leave a Reply