SemiAnalysis Bridging the gap between the world's most important industry, semiconductors, and business.
-
Most Neoclouds Suck At Security
by Jordan Nanos on August 30, 2026
OpenAI vs HuggingFace, Container Escapes, Kernel Bypass, Network Policies, Security Keys, Multi-tenant Grafana, and a ClusterMAX 3.0 Preview
-
OpenAI Jalapeño: Better Than Nvidia Blackwell
by Bryan Shan on August 25, 2026
OpenAI’s self-designed ASIC compared with Rubin, Jalapeño’s TCO, throughput per MW, and spicy deets
-
AgentX - InferenceXv3: Does CUDA Moat Hold up in Agentic Inferencing?
by Cam Quilici on August 24, 2026
$3 Million USD dataset open sourced, 1 Mil+ Context Length, Multiturn, Sub Agents 95%+ KVCache HitRate, GB300 NVL72, MI355, B200
-
Are Open Models Catching Up?
by Evan Cloutier on August 21, 2026
Comparing open vs. closed models across the eras of frontier models
-
Cerebras's Next Generation CS-4: Fast Just Got Faster
by Myron Xie on August 19, 2026
Double the Performance with Double the Power
-
$12B of US ratepayers' money wasted on a modeling mistake and PJM wants to do it again
by Robert Boswall on August 16, 2026
American Grid design needs an overhaul, Why it is good to be full of cold air.
-
Ultra-High Interactivity on NVIDIA GPUs? - TileRT InferenceX
by Bryan Shan on August 10, 2026
Can TileRT software on NVIDIA GPU compete with Cerebras, Groq LPU, SambaNova? Batch Size 1, Disaggregated engine, high throughput engine Prefill, high interactivity engine decode
-
SpaceX 10GW in 2027 – Why It’s Real, Will Drive $300B ARR for SpaceX, and Why Microsoft Will Be the Largest Offtaker
by Jeremie Eliahou Ontiveros on August 7, 2026
Inference at 100B/GW/year, SpaceX's stellar pace, Microsoft's 10GW 2026 Awakening, Azure Can Grow Tiple-Digits
-
Gemini is Cooked but GCP is Cooking
by Max Kan on August 7, 2026
or why DeepMind's long term failure is GCP's short term gain
-
Kimi K3, The Manos, The Mythos, The Legendos
by Kimbo Chen on August 3, 2026
Kimi K3’s architecture: compressed memory, attention across depth, latent expert routing, and the inference performance
-
The Wild Wild West Of LEGO Datacenters
by Nicolas Bontigui on July 29, 2026
The Labor Problem and Modularization to the Rescue
-
Can AMD break the CUDA Moat? AMD Advancing AI 2026
by Bryan Shan on July 25, 2026
Agentic Kernel Generation, Improvement in Software Quality, Unstable Internal Development Clusters, Helios MI455X Production Ramp Hell, Up to 105% Discounts from Finance Engineering
-
Vera Rubin NVL72 vs GB200 NVL72? Inference TCO & Architecture Analysis
by Alec Ibarra on July 23, 2026
3 bit Rubin LUT Based Tensor Core, SM140 Feynman, Rack Scale, Perf Per MegaWatt, Perf Per Dollar, Software Improvements, Public Rubin Software, PyTorch, vLLM, OpenAI Triton
-
Meta’s Infrastructure Team Needs A Culture Reset
by Wayne Ma on July 22, 2026
Meta Infrastructure has become bloated, with middle managers expending resources on over-engineered technology solutions that lose sight of broader organizational needs.
-
The Future of Meta Superintelligence: A 1 Year Progress Update
by Max Kan on July 9, 2026
A top tier RL environment startup spawns out of thin air, the most aggressive compute ramp we've ever seen, 2000km+ scale-across, and some advice for Google DeepMind
-
Anthropic 3Q26 Profit Over $1B: The Anthropic IPO Financials Sneak Peak
by Joey Brookhart on July 8, 2026
Anthropic’s Opportunity is Theirs to Lose
-
Nvidia GPU Debt Backstop Unleashes the AI Project Trinity: Capital, Offtake and Datacenters
by Daniel Nishball on July 6, 2026
Over 7T AI debt by 2029, There can be no Neoclouds without the Trinity. Nvidia's Backstop Economics Explained. AI Debt Needs Quantified. Nvidia's Objective is to Broaden Compute Access, Develop AI Financing, Grow Neoclouds.
-
Meta Compute: Everyone Wants To Be A Neocloud
by Jeremie Eliahou Ontiveros on July 2, 2026
Zuck Takes Plan B? SpaceX 2.0, Bedrock 2.0, MSL Isn't Giving Up, Scaling RecSys by 10x... ClusterMAX ranking coming soon?
-
EMIB-T Roadmap, Custom HBM, HBM4 Packaging Challenges, Microfluidic Cooling, Photonic Interconnects, and More
by Afzal Ahmad on July 2, 2026
ECTC 2026 Roundup, Intel, TSMC, SK Hynix, Samsung, Micron, Marvell, Lightmatter, Microsoft
-
TokenBudgeting: Our Conversations with Enterprises on Token Spend
by Crystal Huang on June 30, 2026
Was Widespread TokenMaxxing Ever Really Here?




