Podcast cover
AI Engineer · Technology

AI Engineer

We turn high signal in-person events for the top AI engineers, founders, leaders, and researchers in the world into the best free learning opportunities for millions around the world here on YouTube. Your subscribes, likes, comments, speaking, attendance, or sponsorships goes a long way toward making our biz model sustainable indefinitely. We strongly believe this industry deserves a better class of community and that we know how to do this well; we just need your support.

Episodes

Episode cover

How Autoresearch is changing ML research — Zhengyao Jiang, Weco

16 Jul 2026
16m
AI processed

Autonomous research agents are transforming machine learning engineering by shifting the focus from manual implementation to high-level system design. Aiden, an autonomous agent developed by Weco, demonstrated this potential by securing seven leaderboard records in OpenAI’s Parameter Golf competition, outperforming hum...

Episode cover

Computer-Use 2.0: Agents Just Got Multi-Cursor — Francesco Bonacci, Cua

15 Jul 2026
16m
AI processed

Computer-using agents are evolving from screen-takeover models to background-driven automation. The CUA driver enables AI agents to interact with operating systems—macOS, Windows, and Linux—using accessibility trees and pixel-level observation without disrupting the user. Evaluating these agents requires rigorous bench...

Episode cover

Recursive Model Improvement — Lee Robinson, Cursor, SpaceXAI

15 Jul 2026
20m
AI processed

Cursor accelerates AI model development through a recursive improvement loop that integrates real-world user feedback, rigorous evaluations, and massive compute resources. The training process utilizes an outer loop of agent usage data and an inner loop of complex, automated engineering tasks designed to push model cap...

Episode cover

Claude Fable, Claude Tag, and Anthropic's Culture — Cat Wu & Thariq Shihipar ft Simon Willison

15 Jul 2026
51m
AI processed

The integration of AI agents like Claude Code and Claude Tag is fundamentally transforming software engineering by shifting the focus from manual implementation to high-level product strategy and business sense. These tools allow engineers to delegate menial tasks, enabling faster iteration cycles and more ambitious pr...

Episode cover

WTF Is the Context Layer? The Missing Infrastructure for Production Agents — Prukalpa Sankar

14 Jul 2026
20m
AI processed

Context is the critical missing component for transforming AI from a high-IQ tool into a high-performance business asset. While cognitive intelligence has grown exponentially, AI remains limited by a lack of situated knowledge, expertise, and organizational norms. Effective AI deployment requires a "context layer" that...

Episode cover

Don't Ship Skills Without Evals — Philipp Schmid, Google DeepMind

14 Jul 2026
21m
AI processed

AI agents require rigorous evaluation to ensure performance and reliability, as shipping skills without testing leads to unpredictable behavior and negative outcomes. Skills—defined as modular instructions for models—should be treated as code, requiring clear directives rather than verbose descriptions. Developers shou...

Episode cover

Forward Deployed Engineering at Cursor — Pauline Brunet

14 Jul 2026
20m
AI processed

Forward Deployment Engineering (FDE) functions serve as a strategic bridge between product innovation and customer success by embedding highly technical experts directly into client organizations. Effective FDE teams avoid the pitfalls of staff augmentation by focusing exclusively on high-impact, project-based work tha...

Episode cover

"The engineer of the future is the person who is able to choose what is worth doing." — Addy Osmani

14 Jul 2026
18m
AI processed

Engineering in the era of AI agents requires a fundamental shift from task execution to human-led accountability and strategic judgment. As AI scales production capabilities, the engineer’s primary value moves toward owning the "verdict"—deciding what is worth building and taking responsibility for the resulting system...

Episode cover

"I've never seen anything scarier than an LLM with tool calls." — Erik Meijer aka @HeadinTheBox

13 Jul 2026
21m
AI processed

AI agents pose significant security risks because they operate in uncontrolled loops capable of executing harmful side effects, such as deleting files or accessing private data. Current alignment strategies, which attempt to bake safety directly into model weights, are insufficient and frequently bypassed by jailbreaki...

Episode cover

Stop Evaluating Models Like It's the 50s - Alejandro Vidal, Mindmakers

13 Jul 2026
23m
AI processed

Item Response Theory (IRT) offers a superior framework for evaluating Large Language Models, moving beyond the limitations of simple accuracy counts inherent in Classical Test Theory. By assigning difficulty and discrimination parameters to individual test items, developers can derive more precise intelligence estimate...

Episode cover

From fork() to Fleet: Designing an Agent Sandbox Cloud — Abhishek Bhardwaj, OpenAI

13 Jul 2026
44m
AI processed

Securely executing untrusted code requires robust sandbox infrastructure to protect host systems from potential exploits while maintaining high performance. AI agents, when granted code execution capabilities, effectively solve verifiable problems like math and programming, but this necessitates isolated environments. ...

Episode cover

Modern Post-Training: A Deep Dive — Will Brown, Prime Intellect

13 Jul 2026
46m
AI processed

Modern post-training infrastructure empowers researchers to refine AI models through modular, open-source tools that bridge the gap between evaluation and large-scale training. Prime Intellect’s "open superintelligence stack" centers on the Verifiers library, which decouples tasks, harnesses, and runtimes to provide a ...

Episode cover

RLM: Recursive Language Models for Large Codebases - Shashi, Superagentic AI

12 Jul 2026
17m
Episode cover

The AI bugpocalypse is here. Now what? - Jack Cable, Corridor

12 Jul 2026
19m
Episode cover

Semantic Blindness: 500,000 Sensors Confused an LLM - Raahul Singh & Vanč Levstik, Phaidra

12 Jul 2026
16m
Episode cover

The Agentic Web and the Bazaar Era of AI - Ramesh Raskar, MIT Media Lab

12 Jul 2026
12m
Episode cover

A Song of Types and Agents - Roberto Stagi, Ratel

12 Jul 2026
14m
Episode cover

ReviewDebt: a practical framework for scoring every pull request — Sachin Gupta, Ebay

12 Jul 2026
25m
Episode cover

remobi.app: Don't change your terminal workflow for mobile

12 Jul 2026
9m
Episode cover

What Does Done Even Mean? Agents and Paperclip's Liveness Model - Dotta, Paperclip

12 Jul 2026
7m
Page 18

Follow this podcast in Podwise

Sign in to get AI summaries, transcripts and mind maps for any episode, including new ones.

Open in Podwise