Podcast cover
Jordan Nanos, Doug O'Laughlin · Business

SemiAnalysis Weekly

Everything semiconductors and AI Covering the spectrum

Episodes

Episode cover

Ep. 035 - Tech DD’s, Performance Projections, Benchmarks, Supply Chain, Investment Thesis (Consulting)| Abhilash Jain, Jordan Nanos

05 Oct 2026
34m
AI processed

SemiAnalysis consulting bridges the gap between technical hardware expertise and financial strategy, serving institutional investors, hyperscalers, and semiconductor companies. The practice focuses on custom engagements, ranging from technical due diligence on AI infrastructure—such as evaluating neocloud performance a...

Episode cover

Ep. 034 - Engrams: How DeepSeek Offloads KV Cache to DRAM and SSD (Core Research) | Jordan Nanos, Cam Quilici, Alec Ibarra, Bryan Shan

02 Oct 2026
1h 3m
AI processed

Architectural co-design and hardware optimization drive significant gains in AI inference efficiency. Engrams enable the offloading of model components to DRAM or SSD, effectively bypassing HBM bandwidth constraints by leveraging token ID-based retrieval. The AgentX benchmark demonstrates that agentic coding workloads ...

Episode cover

Ep. 033 - ClusterMAX 3.0 Is Here! Neoclouds Ranked (Neoclouds, GPUs) | Sam Harshe, Pratt Bhatt, Jordan Nanos

23 Sep 2026
1h 8m
AI processed

The neocloud industry faces significant challenges in reliability, performance, and security as providers scale to meet massive GPU demand. Current testing reveals that many managed cluster providers struggle with basic health checks, often implementing flawed systems that interfere with workloads rather than remediati...

Episode cover

Ep. 032 - 300 Datacenter Bans, 3 Projects Delayed: Moratoriums Explained (Datacenter, Energy) | Maya Barkin, Reyk Knühtsen, Jordan Nanos

18 Sep 2026
53m
AI processed

Data center moratoriums across the United States function primarily as political signaling rather than substantial obstacles to infrastructure development. While over 300 local jurisdictions have enacted pauses, these measures impact only three major projects, as most ordinances target areas lacking active development ...

Episode cover

Ep. 031 - EMERGENCY EPISODE: Are We Doomed? | Jordan Nanos, Doug O'Laughlin, Max Kan, Joey Brookhart

16 Sep 2026
1h 2m
AI processed

Anthropic’s "We Must Pace the Frontier" proposal signals a shift toward prioritizing safety, yet this commitment likely accelerates rather than reduces compute demand. Implementing robust safety controls, such as chain-of-thought monitoring and extensive QA for RL environments, requires significant computational resour...

Episode cover

Ep. 030 - Long Live the Short King: Why 4-HI HBM Wins (Memory) | Myron Xie, Jordan Nanos

14 Sep 2026
38m
AI processed

The AI hardware industry is pivoting from maximizing HBM capacity to prioritizing bandwidth and supply chain efficiency, as evidenced by the redesign of NVIDIA’s Rubin Ultra from 1TB to 192GB of HBM. This shift toward 4-high and 8-high HBM stacks addresses critical manufacturing yield losses and severe memory supply sh...

Episode cover

Ep. 029 - Modular Data Centers Cut Build Time to 12 Months (Datacenter, Energy) | Nico Bontigui, Jordan Nanos, Nigel Chiang, Eric Wen

11 Sep 2026
46m
AI processed

Modular data centers, often termed "Lego data centers," represent a shift from traditional on-site construction to off-site prefabrication, significantly reducing build times and addressing critical labor shortages. By moving mechanical and electrical assembly into factory environments, developers achieve greater predi...

Episode cover

Ep. 028 - Most Neoclouds Suck At Security: How Agents Hacked Hugging Face (Neoclouds, Security) | Doug O'Laughlin, Sam Harshe, Jordan Nanos

02 Sep 2026
50m
AI processed

Neocloud providers often lack the rigorous security standards of hyperscalers, creating significant risks for companies relying on them for GPU-intensive workloads. Many providers fail to implement basic protections like tenant isolation, updated software, and Kubernetes admission controllers, leaving systems vulnerabl...

Episode cover

Ep. 027 - OpenAI Jalapeño: Better Than Nvidia Blackwell (Accelerators)

30 Aug 2026
1h 3m
AI processed

OpenAI’s custom ASIC, Jalapeno, redefines inference efficiency by outperforming Nvidia’s Vera Rubin on a performance-per-megawatt and total cost of ownership basis. Despite having lower theoretical specifications, the chip achieves superior throughput by utilizing a microarchitecture that minimizes data movement and em...

Episode cover

Ep. 026 - PJM's $12B Modeling Mistake Is Hitting Ratepayers Again (Datacenter, Energy) | Robert Boswell, Jordan Nanos

20 Aug 2026
43m
AI processed

PJM’s current capacity auction design forces ratepayers to absorb $12 billion in avoidable costs due to systemic modeling errors and inefficient market structures. By failing to account for the increased efficiency of gas turbines in winter and relying on backward-looking reliability data, the grid operator consistentl...

Episode cover

Ep. 25 - DYLAN IS HERE, LIVE! | Dylan Patel & Jordan Nanos

17 Aug 2026
38m
AI processed

Operational dynamics and strategic challenges define the growth of an AI-focused firm like SemiAnalysis. Initial capital expenditures on AI tools and cloud resources spike during R&D phases, yet steady-state costs remain manageable as workflows mature. AI-driven efficiency often defies traditional cost-cutting metrics,...

Episode cover

Ep. 024 - SpaceX's 10GW Plan Drives $300B ARR by 2027 (Datacenter, Energy) | Reyk Knuhtsen, Jeremie Eliahou Ontiveros, Jordan Nanos

09 Aug 2026
50m
AI processed

SpaceX’s ambition to deploy 10 gigawatts of data center capacity by 2027 hinges on a high-margin business model that prioritizes speed over traditional infrastructure constraints. By leveraging existing warehouses and gas pipelines, the company bypasses standard permitting delays to meet the urgent compute needs of fro...

Episode cover

Ep. 023 - Everyone Leaves Google, Elon Forecasts 1T ARR, Reflecting On GPT-5, Building Personalized Software (Roundtable) | Jon Y, Doug O'Laughlin, Jordan Nanos

07 Aug 2026
44m
AI processed

Google’s recent loss of key technical talent, including Jeff Dean and other Gemini leads, signals a significant shift in the company’s ability to maintain frontier AI dominance. While Google functions as a massive, bureaucratic research lab, agile startups like OpenAI and Anthropic are better positioned to execute on r...

Episode cover

Ep. 022 - Market Drawdown, Historic Bubbles, Funding The Buildout, AI Politics (Doug is Back)

29 Jul 2026
49m
AI processed

The recent semiconductor and AI stock market drawdown reflects a correction following a historic period of rapid growth, exacerbated by high leverage and cooling expectations for memory price increases. While demand for AI compute remains robust, driven by the proliferation of coding agents and evolving model capabilit...

Episode cover

Ep. 021 - The AI Project Trinity: Capital, Offtake, Data Center (Datacenter, Energy) | Dan Nishball, Jordan Nanos, Zane Fong, Kang Wen Cheang

23 Jul 2026
52m
AI processed

AI infrastructure financing requires $7.1 trillion in debt to support $11 trillion in cumulative capital expenditure through 2029. To bridge the gap between speculative NeoCloud projects and traditional lending requirements, NVIDIA acts as a central bank, providing GPU backstops that guarantee a floor price for compute...

Episode cover

Ep. 020 - Anthropic vs OpenAI Usage, Margins, Meta Compute, Future of MSL (Tokenomics) | Crystual Huang, Max Kan, Joey Brookhart, Jordan Nanos

18 Jul 2026
50m
AI processed

The current AI landscape is shifting from aggressive token consumption to strategic austerity as enterprises refine budgets while prioritizing high-ROI coding tasks. Anthropic maintains a competitive edge through its enterprise-focused, high-margin API business, though OpenAI is rapidly regaining market share with the ...

Episode cover

[Emergency Episode] Moonshot’s Kimi K3 has Arrived! China has a Frontier Model

18 Jul 2026
31m
AI processed

Kimi K3 establishes itself as the world’s third-best AI model, signaling a significant shift in the competitive landscape previously dominated by OpenAI, Anthropic, and Google. With a 2.8 trillion parameter architecture, the model demonstrates that frontier performance is achievable at scale, suggesting that leading cl...

Episode cover

Ep. 019 - Inside the STEEL Lab: From Package to Transistor (Teardown Lab) | Afzal Ahmad, Andrew Wagner, Jordan Nanos

16 Jul 2026
36m
AI processed

The Steel teardown lab at SemiAnalysis provides critical insights into semiconductor manufacturing and chip architecture by deconstructing consumer devices down to the transistor level. Through advanced techniques like focused ion beam milling, transmission electron microscopy, and chemical mechanical polishing, the te...

Episode cover

Ep. 018 - Stop Saying Half of 2026 US Datacenter Capacity Is Canceled (Datacenter, Energy) | Jeremie Eliahou Ontiveros, Reyk Knuhtsen, Ellie Holbrook, Jordan Nanos

09 Jul 2026
50m
AI processed

Media reports claiming that half of projected 2026 US data center capacity is canceled misinterpret early-stage project volatility as systemic failure. In reality, the industry is experiencing a strategic pivot toward behind-the-meter power generation to bypass grid interconnection bottlenecks, with forecasts projectin...

Episode cover

Ep. 017 - DeepSeek V4 and Huawei Ascend NPU Performance (InferenceX) | Kimbo Chen, Cam Quilici, Bryan Shan, Jordan Nanos

01 Jul 2026
34m
AI processed

DeepSeek V4’s transition to a 1-million context length architecture relies on aggressive innovations in sparse attention and Mega MoE, which reduce KV cache memory requirements by approximately 100x compared to standard models. Achieving day-zero inference performance on new hardware requires complex engineering, speci...

Page 1

Follow this podcast in Podwise

Sign in to get AI summaries, transcripts and mind maps for any episode, including new ones.

Open in Podwise