Podcast cover
SemiAnalysis · Technology

SemiAnalysis

Bridging the gap between the world's most important industry and business.

Episodes

Episode cover

Ep. 034 - The Fight for Fast Tokens, TPU v7, Vera Rubin, and Engrams (AI Supply Chain, InferenceX)

02 Oct 2026AI processed

AI inference optimization hinges on architectural co-design and strategic memory management to overcome HBM bandwidth limitations. Engrams enable efficient model serving by offloading KV cache to DRAM or SSD, effectively reducing parameter memorization overhead. The AgentX benchmark, which replicates real-world agentic...

Episode cover

Ep. 033 - ClusterMAX 3.0 Is Here! Neoclouds Ranked (Neoclouds, GPUs)

23 Sep 2026AI processed

The ClusterMAX 3.0 review evaluates 77 cloud providers, highlighting a shift in market dynamics where GPU scarcity drives high margins regardless of infrastructure quality. Reliability remains the primary differentiator, yet many providers struggle with basic health checks, networking configurations, and storage stabil...

Episode cover

Ep. 033 - 300 Data Center Bans, 3 Projects Delayed: Moratoriums Explained (Datacenter, Energy)

18 Sep 2026AI processed

Data center moratoriums, while frequently appearing in headlines as a widespread threat to industry growth, currently impact only a negligible fraction of the U.S. project pipeline. These temporary legal pauses, primarily enacted at the local level, often serve as political signaling or a mechanism for municipalities t...

Episode cover

Ep. 031 - EMERGENCY EPISODE: Are We Doomed? | Jordan Nanos, Doug O'Laughlin, Max Kan, Joey Brookhart

16 Sep 2026AI processed

The industry's shift toward "pacing the frontier" of AI development reflects a growing consensus on the necessity of safety, though this commitment will likely accelerate rather than diminish compute demand. Implementing robust safety controls, such as embedded third-party evaluators and continuous chain-of-thought mon...

Episode cover

Ep. 030 - Long Live the Short King: Why 4-hi HBM Wins (Memory)

14 Sep 2026AI processed

The AI hardware industry is shifting from high-density HBM stacks toward 4-high and 8-high configurations to address critical memory supply shortages. While previous generations prioritized increasing capacity to accommodate larger model parameters, modern inference and post-training workloads prioritize bandwidth over...

Episode cover

AI is running out of Power

06 Sep 2026AI processed

The AI industry’s rapid expansion is colliding with a physical bottleneck in the electrical grid, forcing companies to adopt "Behind-the-Meter" power generation to maintain growth. As AI datacenters evolve into industrial-scale intelligence factories, they increasingly rely on "Bring Your Own Generation" (BYOG) strateg...

Episode cover

Ep. 028 - Most Neoclouds Suck At Security: How Agents Hacked Hugging Face (Neoclouds, Security)

02 Sep 2026AI processed

Neocloud providers often lack the enterprise-grade security standards of hyperscalers, leaving critical infrastructure vulnerable to exploitation. The recent Hugging Face and OpenAI incidents highlight how AI agents leverage mundane security failures—such as outdated kernels, misconfigured Kubernetes clusters, and poor...

Episode cover

Ep. 027 - OpenAI Jalapeño: Better Than Nvidia Blackwell (Accelerators)

29 Aug 2026AI processed

OpenAI’s custom ASIC, "Jalapeno," significantly outperforms Nvidia’s Rubin architecture in inference throughput per megawatt and total cost of ownership. By leveraging HBM4 memory and AI-assisted microarchitecture design, the chip achieves superior efficiency despite having lower raw specifications on paper. The develo...

Episode cover

Ep. 25 - DYLAN IS HERE, LIVE! | Dylan Patel & Jordan Nanos

17 Aug 2026AI processed

SemiAnalysis operations and the evolving role of AI agents dominate this discussion. The conversation centers on the company's internal AI spending, which has transitioned from high-growth R&D to a more stable state as employees integrate agents into their daily workflows. A significant focus is placed on the emergence...

Episode cover

Ep. 022 - Market Drawdown, Historic Bubbles, Funding The Buildout, AI Politics (Doug is Back)

29 Jul 2026AI processed

The current stock market drawdown, particularly within the AI and memory sectors, stems from a combination of technical unwinding after a historic rally and concerns regarding the sustainability of memory price increases. While memory manufacturers face pressure due to shifting long-term agreements and slowing rates of...

Episode cover

Why next-gen AI scale-up needs CPO

03 Jun 2026AI processed

Modern AI data centers face a critical bottleneck as copper interconnects reach their physical limits for high-speed data transmission. While copper remains the standard for rack-internal "scale-up" networks due to its low latency and native compatibility with semiconductor signals, its reach is restricted to approxima...

Episode cover

How Makora Generates CUDA Kernels That Beat Hand-Tuned Code | Researcher Conversations at GTC

27 May 2026
26m
Episode cover

Why Anthropic is renting Colossus 1 from SpaceX

18 May 2026
5m
Episode cover

Analog's Mechatronics Engineer: Hardware Powering AI Inference | Researcher Conversations at GTC

14 May 2026
9m
Episode cover

AWS Trainium: How Amazon Built Their Own AI Chips | Researcher Conversations at GTC

30 Apr 2026
7m
AI processed

AWS is expanding its cloud infrastructure by adding one million GPUs this calendar year, bringing its total footprint to three million units to meet surging generative AI demand. This expansion leverages a 15-year partnership with NVIDIA, including the upcoming deployment of Rubin-generation systems. Beyond third-party...

Episode cover

Why Positron AI is Choosing LPDDR over HBM for Next-Gen LLM | Researcher Conversations at GTC

16 Apr 2026
10m
Episode cover

Claude Code Psychosis: How SemiAnalysis Is Token Mogging Meta | Ep. 008

10 Apr 2026
34m
Episode cover

Bryan Shan x Cameron Quilici | Researcher Conversations at GTC

07 Apr 2026
9m
Episode cover

Waleed Atallah (Makora) x Dylan Patel | Researcher Conversations at GTC

30 Mar 2026
8m
Episode cover

AI Chip & Silicon Round-up 2026

12 Mar 2026
13m
Page 1

Follow this podcast in Podwise

Sign in to get AI summaries, transcripts and mind maps for any episode, including new ones.

Open in Podwise