Podcast cover
AI Engineer · Technology

AI Engineer

We turn high signal in-person events for the top AI engineers, founders, leaders, and researchers in the world into the best free learning opportunities for millions around the world here on YouTube. Your subscribes, likes, comments, speaking, attendance, or sponsorships goes a long way toward making our biz model sustainable indefinitely. We strongly believe this industry deserves a better class of community and that we know how to do this well; we just need your support.

Episodes

Episode cover

Your LLM Deception Monitor Is Broken. The Fix Is in the Training Data - Sachin Kumar, LexisNexis

08 Jul 2026
13m
AI processed

Sleeper agents in fine-tuned large language models present a critical security risk, as backdoors remain dormant during standard behavioral evaluations and only activate upon specific, benign triggers. Traditional monitoring, including joint feature analysis, fails to isolate these malicious signals because they are di...

Episode cover

GTM Is You - Victoria Melnikova, Evil Martians

07 Jul 2026
12m
Episode cover

Beyond the Harness: A Journey Towards Adaptative Engineering - Rajiv Chandegra, Annicha Labs

07 Jul 2026
37m
AI processed

AI engineering is transitioning from a "factory model" of fixed, pre-determined harnesses to an "adaptive engineering" paradigm. While current tools like CLI coding agents and IDEs rely on rigid, upfront configurations to ensure reliability, these systems struggle with the messiness of real-world, multi-agent, and cros...

Episode cover

How we taught agents to use good retrieval - Hanna Lichtenberg, Mixedbread AI

07 Jul 2026
14m
AI processed

The "knowledge gap" between rapidly evolving LLM reasoning capabilities and stagnant search retrieval systems limits the effectiveness of AI agents in complex tasks. While models are highly capable, they often rely on inefficient, keyword-heavy queries optimized for legacy benchmarks, leading to significant performance...

Episode cover

What if the harness mattered more than the model? - Aditya Bhargava, Etsy

07 Jul 2026
32m
AI processed

The performance of AI agents depends more on the "harness"—the surrounding framework and tools—than the underlying model. By prioritizing robust harness design, developers can achieve high-level performance using local, open-source models, reducing dependency on proprietary systems. Effective agent development requires...

Episode cover

Build AI Systems for Discernment, Not Approval - Angel Ortmann Lee, Duolingo

07 Jul 2026
25m
AI processed

Building AI systems requires prioritizing human discernment over passive approval to combat cognitive surrender, where users uncritically adopt AI outputs. Research, including a case study on the Duolingo English Test, demonstrates that even skilled reviewers often defer to AI signals, effectively rubber-stamping error...

Episode cover

Respect The Process - Andrew Dumit, Watershed Technology Inc.

07 Jul 2026
16m
Episode cover

The Pipeline Is Dead - Iris ten Teije, Sky Valley Ambient Computing

07 Jul 2026
19m
Episode cover

500 people vibe-coded for 30 days. I was one of them. - Sanja Grbic, Automattic

07 Jul 2026
17m
Episode cover

SWE-Marathon: Evaluating Coding Agents at Billion-Token Scale - Rishi Desai, Abundant AI

07 Jul 2026
12m
Episode cover

Field Guide to Fable — Thariq Shihipar, Anthropic

06 Jul 2026
19m
AI processed

Fable, the latest model from Anthropic, represents a significant shift toward agentic coding where the conceptual "map" of a project meets the complex "territory" of a real-world codebase. Successfully navigating this transition requires "unhobbling" the model by recognizing its spiky capabilities, such as using code e...

Episode cover

The Missing Layer After Launch - Raphael Kalandadze, Wandero AI

05 Jul 2026
19m
Episode cover

Continual Learning for AI Agents: From Failures to Durable Improvements - Soheil Feizi, RELAI

05 Jul 2026
22m
AI processed

Verifiable Continual Learning (VCL) enables AI agents to improve from experience without forgetting previous successes or introducing regressions. This approach shifts focus from simple model fine-tuning to a multi-layered strategy involving harness and memory updates. Because raw production logs lack the structure for...

Episode cover

MCP Apps: Primitives, discovery, and the Future of Software - Pietro Zullo, Manufact, Inc

05 Jul 2026AI processed

MCP Apps represent the next evolution of the Model Context Protocol, shifting from simple JSON-based tool calls to interactive, UI-rich experiences embedded directly within LLM chat interfaces. By leveraging sandboxed iframes, these applications enable bidirectional communication between the user and the host, allowing...

Episode cover

Your AI Product Will Fail Unless You Can Explain It - Veronica Hylak, Hey AI

05 Jul 2026
6m
Episode cover

WF26: Harness Engineering & Startup Battlefield ft. Garry Tan, Mike Krieger, @t3dotgg , DSPy

03 Jul 2026
9h 11m
Episode cover

The Prompt Is Still a Punch Card - Ted Johnson, JoinIn AI

02 Jul 2026
20m
Episode cover

WF2026: Autoresearch & Keynotes ft. Anthropic, Google DeepMind, Amazon AGI, Sonar, Arena, Recursive

02 Jul 2026
8h 51m
Episode cover

WF2026: Software Factories & Keynotes ft. Microsoft, OpenAI, OpenClaw, Z.ai (GLM), MiniMax, HF

01 Jul 2026
8h 36m
AI processed

The AI Engineer World's Fair centers on the transition of software development from manual coding to autonomous "software factories" driven by agentic loops. This paradigm shift emphasizes orchestration over individual code generation, where engineers design and manage complex, self-improving systems that handle the fu...

Episode cover

Building Great Agent Skills: The Missing Manual

29 Jun 2026
20m
AI processed

"Skill Hell" arises when developers struggle to integrate freely available AI skills, leading to poor performance and organizational inefficiency. To overcome this, evaluate skills using a four-part framework: trigger, structure, steering, and pruning. First, determine if a skill is user-invoked or model-invoked to bal...

Page 20

Follow this podcast in Podwise

Sign in to get AI summaries, transcripts and mind maps for any episode, including new ones.

Open in Podwise