A fashion photoshoot of a runway look inspired by computer chips. A large screen displays the word “Chips” –ar 4:3 –style raw

TSMC says it has discussed moving fabs out of Taiwan but such a move impossible | Reuters

Securing Research Infrastructure for Advanced AI | OpenAI

We’re sharing some high-level details on the security architecture of our research supercomputers.

World’s first bioprocessor uses 16 human brain organoids for ‘a million times less power’ consumption than a digital chip | Tom’s Hardware

Even the Raspberry Pi is getting in on AI – The Verge

Amazon

Down for less than four minutes a month: how AWS deploys code

AMD

AMD launches new AI chips to take on leader Nvidia | Reuters

AMD unveils new AI chips to compete with Nvidia – Fast Company

Groq

“Last Week: Groq exceeded 30,000 Tokens / second input rate on Llama3 8B❗️ This Week: Llama3 70B at 40,792 Tokens/s input rate‼️ – FP16 Multiply, FP32 Accumulate – 7989 tokens in – full Llama context length Next Week: …? 😮 

“A GPT-4 level chatbot, available to use completely free, running at over 800 tokens per second on Groq. I’m genuinely mindblown by LlaMA 3. Try it with the link in the next tweet. 

“Put this mind bending achievement in perspective: @GroqInc runs Llama 70b in lossless precision on ~4 Wikipedia articles in quite literally the blink of an eye. – A 70B model in 16-bit precision with 32-bit accumulation (loss-less). – Processing ~8000 tokens in 0.2 seconds (or”

Mamba

“excited to finally release Mamba-2!! 8x larger states, 50% faster training, and even more S’s 🐍🐍 Mamba-2 aims to advance the theory of sequence models, developing a framework of connections between SSMs and (linear) attention that we call state space duality (SSD) w/@tri_dao 

“With @_albertgu, we’ve built a rich theoretical framework of state-space duality, showing that many linear attn variants and SSMs are equivalent! The resulting model, Mamba-2 is better & faster than Mamba-1, and still matching strong Transformer arch on language modeling. 1/ 

State Space Duality (Mamba-2) Part I – The Model | Goomba Lab

NVIDIA

FTC and DOJ reportedly opening antitrust investigations into Microsoft, OpenAI, and Nvidia – The Verge

U.S. Clears Way for Antitrust Inquiries of Nvidia, Microsoft and OpenAI – The New York Times

Nvidia and Salesforce double down on AI startup Cohere in $450 million round, source says | Reuters

“The most underrated announcement from the Nvidia Computex keynote: Project G-Assist It’s like the ChatGPT Mac App, but instead of understanding your screen, it provides context-aware help and personalized responses for gaming.

“Nvidia CEO Jensen Huang just announced next-gen ‘Rubin’ chips, slated for 2026, with ‘Rubin Ultra’ coming a year later. “The days of millions of GPU data centers are coming” says Jenson. It’s Nvidia’s AI world — we’re all just living in it. 

“Nvidia also showed off Project G-Assist. It’s an AI gaming assistant that provides context-aware help and personalized responses for PC games. The way we game right now is going to look like Tetris in 5 years… 

“Computational cost reduced by 350X for datacenter-scale AI by @nvidia over the last 8 years. 🤯 From Jensen Huang Keynote at COMPUTEX 2024 finished just now. 

China’s Nvidia Loophole: How ByteDance Got the Best AI Chips Despite U.S. Restrictions — The Information

Nvidia’s Latest Investment is an AI Startup Focusing on Video Search – Bloomberg

“Thanks Jensen. Now you can use @nvidia NIM directly from the @huggingface hub for Llama3. Very excited to see that Hugging Face is becoming the gateway for AI compute! 

“Optimum-NVIDIA from @nvidia – By changing just a single line of code, you can unlock up to 28x faster inference and 1,200 tokens/second on the NVIDIA platform. 🔥 📌 Optimum-NVIDIA is the first Hugging Face inference library to benefit from the new float8 format supported on 

“At COMPUTEX 2024, @nvidia CEO Jensen showing how Pandas code is now 50x faster on @GoogleColab after integration with RAPIDS cuDF over standard pandas. This is with zero code changes, and is available in the default runtime environment. Just add in %load-ext cudf.pandas over 

Twelve Labs Earns $50 Million Series A Co-led by NEA and NVIDIA’s NVentures to Build the Future of Multimodal AI

Nvidia announces new AI chips as market competition heats up

Nvidia prepping AI PC chip with Arm and Blackwell cores • The Register

Nvidia is now more valuable than Apple at $3.01 trillion – The Verge

Nvidia value hits $3tn, overtaking Apple

NVIDIA at Computex 2024 | June 4-7, 2024 | NVIDIA

“When it became obvious that compute is the bottleneck for scaling to AGI, Nvidia’s stock skyrocketed. Now that it’s becoming obvious that energy/power is the new bottleneck for scaling to AGI, whose stock(s) will skyrocket in a few years/months from now?”

“Earth 2 is easily one of my favorite projects — fusing geospatial imagery, physics simulation, genAI and 3D visualization to help countries & corporations to predict and respond to extreme weather. And now it goes to street level ft. an AI Jensen Huang delivering the narration. https://twitter.com/bilawalsidhu/status/1799238160024580284

This week’s executive overview and top links are here:

AI News #36: Week Ending 06/07/2024 with Executive Summary and Top 40 Links

The post you just read is an deep dive extension of my weekly newsletter, This Week In AI, an executive summary of the top things to know in AI. Each week, I create an accessible overview for laypeople to feel confident they are conversant with the week’s AI developments. I include a curated list of must-click links of the week, to offer everyone a hands-on opportunity to explore the most intriguing updates in artificial intelligence across various categories, including robotics, imagery, video, AR/VR, science, ethics, and more. Beyond the overview, I post these topic-based deeper dives (below). If you haven’t read this week’s overview, I recommend starting there.

Credits/Sources

Most of these weekly links come from just a few prolific oversharing sources. Please follow them, as they work hard to find the news each week and they make it a lot easier for me to compile.

For previous issues, please visit the archives!

Thanks for reading!

One response to “Chips, Hardware, and Infrastructure: Week Ending 06/07/2024”

  1. […] Chips and Hardware AI News of the Week: Most of the chip news is NVIDA usually, yet more and more Meta, Google, and OpenAI are starting toward their own manufacturing. I have to make the call whether to put Meta, Google, and OpenAI’s chip news under this section or their company sections. Lately, I’m putting each company’s chips news into the company category, rather than the chips category. This is the rest of the chips headlines.This week’s latest chips and hardware news: https://ethanbholland.com/2024/06/07/chips-hardware-and-infrastructure-week-ending-06-07-2024/ […]

Leave a Reply

Trending

Discover more from Ethan B. Holland

Subscribe now to keep reading and get access to the full archive.

Continue reading