Ban the H20: Competing in the Inference Age
substack.com · 3,493 words · saved by 1 readers
How inference scaling should change US AI strategy
A guest post by Venkat Somala. He previously served at the Special Competitive Studies Project, the House Select Committee on the Chinese Communist Party, and has held roles in product and data science in industry. TLDR: U.S. export controls targeting China’s AI capabilities focus primarily on limiting training hardware but overlook the growing importance of inference compute as a key driver of AI innovation. Current restrictions don’t effectively limit China’s access to inference-capable hardware (such as NVIDIA’s H20) and don’t account for China’s strong inference efficiency. While…
saved by
related reading
- An Action Plan for American Leadership in AI | IFPifp.org
- The Short Case for Nvidia Stock | YouTube Transcript Optimizeryoutubetranscriptoptimizer.com
- 2028: Two scenarios for global AI leadership \ Anthropicanthropic.com
- Dario Amodei — On DeepSeek and Export Controlsdarioamodei.com
- China's AI Models Are Closing the Gap—but America's Real Advantage Lies Elsewhere | RANDrand.org
- The Inference Shift – Stratechery by Ben Thompsonstratechery.com
- Why does China open source their AI models?firstscattering.com
- Why does China open source their AI models?substack.com
- My picture of the present in AI — LessWronglesswrong.com
- Dario Amodei — On DeepSeek and Export Controlsdarioamodei.com
- Who’s Afraid of Chinese Models? – Stratechery by Ben Thompsonstratechery.com
- No Jensen, Not All Compute is Created Equalchinatalk.media