Best AI Engineering Videos

Curated talks, tutorials, and deep dives on LLMs, RAG, agents, and production ML.

160 curated videos across 5 categories

AEAI Engineer70 videos
Building Turbopuffer: Gergely Orosz (@pragmaticengineer ) × Simon Eskildsen (CEO)New
Building Turbopuffer: Gergely Orosz (@pragmaticengineer ) × Simon Eskildsen (CEO)
Videos
Building Turbopuffer: Gergely Orosz (@pragmaticengineer ) × Simon Eskildsen (CEO)
Today
MCP Apps: Extending the Frontier — Ido Salomon & Liad Yosef
Yesterday
MCP Tasks (async): Why Aren't Any Agents Supporting Them? — Cornelia Davis, Temporal
Yesterday
When Will The Benchmaxxing Plague End? — Nick Heiner, Surge AI
Yesterday
Emulated: The Data for Fully Autonomous Software Engineers and Companies — Joseph Wang
3 days ago
Agents at Scale: Inside MiniMax's Model and the Infrastructure Behind It — Olive Song
4 days ago
Why Large? Tiny LMs & Agents on Edge/Robotics — Cormac Brick, Google
Jul 25, 2026
From Agent Traces to Agent Simulations — Rustem Feyzkhanov, Snorkel AI
Jul 25, 2026
Building Closed-Loop Evals for a Multimodal Agent at Scale — Soumya Gupta & Jai Chopra, Uber
Jul 24, 2026
How Evals and Prompts Shape Agent Behavior — Preetika Bhateja & Daniel Bump, YouTube Ads
Jul 24, 2026
From Signal to PR: Anatomy of a Self-Improving Agent — Jason Lopatecki, Arize
Jul 24, 2026
The Future of Evals: From LLM as a Judge to Agent as a Judge — Aparna Dhinakaran, Arize AI
Jul 24, 2026
Vending-Bench: Long-Horizon Agent Evals — Lukas Petersson, Andon Labs
Jul 24, 2026
Perception Agents — Antje Barth, Amazon AGI Lab
Jul 23, 2026
Autonomous Agents for Scientific Tasks - Sina Shahandeh, Radicait
Jul 18, 2026
Agents Need Receipts, Not More Tool Calls - Armanas Povilionis, Alithea Bio
Jul 18, 2026
Agents Need Feature Flags - Sachin Gupta
Jul 18, 2026
Your Agents Need a Save Button - Hamza Tahir, ZenML
Jul 18, 2026
The Great Loops Debate — Dex Horthy, Geoff Huntley, Ian Livingstone, Greg Pstrucha, @insecure-agents
Jul 17, 2026
An AI Agent Became the #1 Contributor in OpenAI's Hiring Challenge — Zhengyao Jiang, Weco
Jul 16, 2026
Computer-Use 2.0: Agents Just Got Multi-Cursor — Francesco Bonacci, Cua
Jul 15, 2026
WTF Is the Context Layer? The Missing Infrastructure for Production Agents — Prukalpa Sankar
Jul 14, 2026
Stop AI Agent Hallucinations: 5 Techniques + Production Patterns - Elizabeth Fuentes, AWS
Jul 11, 2026
The Factory That Dreams: 39 AI Agents, No Framework - Rushabh Doshi, Machinecraft
Jul 11, 2026
Every Solo Agent Builder Eventually Reinvents a Worse Version of CI/CD - Sumaiya Shrabony
Jul 11, 2026
Special Topics in Kernels, RL, Reward Hacking in Agents — Daniel Han, Unsloth
Jul 11, 2026
Design Patterns for AI Trust: Juries, Libraries, and Agent Tiers — Alex Bauer, Upside.tech
Jul 11, 2026
Teaching AI to Find Real Vulnerabilities — David Brumley, Bugcrowd
3 days ago
Rethinking Environments for Long-Horizon Work — Rayan Garg, Theta Software
3 days ago
What's Next After RLHF? — Diogo Almeida, TypeSafe AI
3 days ago
Data Quality Is the Compute Multiplier — Ari Morcos, DatologyAI
3 days ago
Learning on the Job: The Future of Post-Training — Raymond Feng, Applied Compute
3 days ago
Data and Environment Curation for Post-Training LLMs — Mahesh Sathiamoorthy, Bespoke Labs
3 days ago
The Base Model Is Dead — Varun Singh, Arcee AI
3 days ago
Verifiable Environments for AI in Biology — Kenny Workman, LatchBio
3 days ago
Ending AI Slop — Thais Castello Branco, Taste Labs
3 days ago
Benchmarks: The Good, the Bad, and the Ugly — Ali Khial, G2i
3 days ago
Reinforcement Learning without Verifiable Rewards — Will Brown, Prime Intellect
3 days ago
fighting slop with slop — Vaibhav Gupta, Boundary
4 days ago
DeepSWE: A Contamination-Resistant Coding Benchmark — James Shi, Datacurve
Jul 26, 2026
State of Data — Sean Cai, Independent / State of Data
Jul 26, 2026
The Messy Reality of Scale: Synthetic Data and Pre-Training — Marah Abdin & Robert McHardy, poolside
Jul 26, 2026
Evals-Driven Development for a Mental Health AI Coach — Akele Reed & Dave Revere, SonderMind
Jul 25, 2026
Loop Engineering from First Principles — Kyle Mistele, HumanLayer
Jul 25, 2026
Evaling Video Slop — Maor Bril, Character.ai
Jul 25, 2026
Everything Is a Rollout — Alex Shaw + Ryan Marten, Terminal-Bench, Harbor, Laude Institute
Jul 24, 2026
Full Workshop: Setting Yourself Up for Success —Jason Liu, OpenAI Codex
Jul 24, 2026
Training Frontier Models to Out-Think Hackers — Uri Rolls, Arithmetic & Thom Wolf, Hugging Face
Jul 24, 2026
The Unreasonable Effectiveness of Separating the Task from the Model — Maxime Rivest & Isaac Miller
Jul 23, 2026
Notion's Token Town — Sarah Sachs, Notion
Jul 23, 2026
Harness Engineering is not Enough: Why Software Factories Fail — Dex Horthy, HumanLayer
Jul 23, 2026
AI on Your Lakehouse: Context Comes in Shapes, Not Queries — Zach Blumenfeld, Neo4j
Jul 23, 2026
Road to 5 Million Tokens: Breaking Barriers in Long Context Training — Max Ryabinin, Together AI
Jun 8, 2026
Benchmarking semantic code retrieval on Claude Code — Kuba Rogut, Turbopuffer
Jun 3, 2026
Connecting the Dots with Context Graphs — Stephen Chin, Neo4j
May 16, 2026
Combine Skills and MCP to Close the Context Gap — Pedro Rodrigues, Supabase
May 15, 2026
Mergeable by default: Building the context engine to save time and tokens — Peter Werry, Unblocked
May 3, 2026
Context Is the New Code — Patrick Debois, Tessl
May 3, 2026
MCP = Mega Context Problem - Matt Carey
Apr 25, 2026
Jack Morris: Stuffing Context is not Memory, Updating Weights is
Dec 29, 2025
Scaling to Long Horizons — Ross Taylor & Chengxi Taylor, General Reasoning
3 days ago
Forward Deployed Engineering at Cursor — Pauline Brunet
Jul 14, 2026
GPU Cloud Deployment Without Leaving Your IDE — Audry Hsu, RunPod
Jun 9, 2026
Under 5 minutes to a deployed LLM endpoint — Audry Hsu, RunPod
Jun 7, 2026
Task Fidelity Scaling Laws — Kobie Crawdord, Snorkel
Jun 2, 2026
Lessons from Trillion Token Deployments at Fortune 500s — Alessandro Cappelli, Adaptive ML
May 12, 2026
Build & deploy AI-powered apps — Paige Bailey, Google DeepMind
Apr 29, 2026
Lessons from Scaling GitHub's Remote MCP Server — Sam Morrow, GitHub
Apr 27, 2026
What we learned scaling MCPs to Enterprise — Karan Sampath, Anthropic
Apr 27, 2026
VoiceOps-fying Low-Latency Intelligence Extraction from Messy Audio Streams — Dippu Kumar Singh
Apr 8, 2026
LALangChain31 videos
Building Deep Agents and Deploying in ProductionNew
Building Deep Agents and Deploying in Production
Videos
Building Deep Agents and Deploying in Production
3 days ago
Why Do AI Agents Hallucinate?
3 days ago
The misaligned incentives behind AI coding agents
4 days ago
Deep Agents in 1 sentence
5 days ago
Autonomous Agent Improvement with LangSmith Engine | New LangChain Academy Course
5 days ago
How Credit Genie Debugs Thousands of Agent Traces with LangSmith
Jul 27, 2026
Inside the Agent Engine: A LangChain and Traversal Fireside Chat
Jul 24, 2026
The Art of Loop Engineering: How to Build Agents That Improve Over Time
Jul 23, 2026
Build a secure computer for your agent
Jul 23, 2026
How Salesforce Standardizes Agent Evals with LangSmith
Jul 23, 2026
The Agent Development Lifecycle 101 by Harrison Chase
Jul 22, 2026
60% Faster Time-to-Interview: Transforming Hiring with AI Agents with LangChain
Jul 22, 2026
Trace Every Cursor Agent Turn in LangSmith
Jul 21, 2026
Interrupt 26: The Agent Conference by LangChain
Jul 21, 2026
LangSmith Sandboxes: A Secure Computer for Your Agent
Jul 20, 2026
How Coinbase Builds Developer Support Agents | Interrupt 26
Jul 20, 2026
Deep Agents Explained
Jul 17, 2026
LangSmith: The Agent Engineering Platform
Jul 17, 2026
What Is an AI Agent, Actually?
Jul 17, 2026
The best AI agents need more humans than you think
Jul 16, 2026
How 11x Built a Slack-Native Bug Triage Agent with LangSmith Fleet
Jul 15, 2026
Inside Toyota's Production System for Agents | Interrupt 26
Jul 15, 2026
The LangChain Team Answers the Most Searched Questions About Agents
Jul 14, 2026
Not Every Bug Engine Finds Is Worth Fixing
Jul 27, 2026
What Is RAG, Actually?
Jul 24, 2026
How Bridgewater Built Pat, The AI Pocket Analyst Tool | Interrupt 26
Jul 24, 2026
Why Sierra bets on outcome-based pricing #Shorts
Jul 23, 2026
How to manage context the right way with LangSmith's Context Hub
May 27, 2026
Hot vs. cold context isn't about time — it's about whether it's being used | Max Agency #podcast
May 26, 2026
LangChain Academy New Course: Introduction to LangSmith Deployment
May 28, 2026
Make Your LangSmith Deployment Multi-Tenant
Apr 14, 2026
AKAndrej Karpathy15 videos
How I use LLMs
How I use LLMs
Feb 27, 2025Watch on YouTube
Videos
How I use LLMs
Feb 27, 2025
Deep Dive into LLMs like ChatGPT
Feb 5, 2025
Let's reproduce GPT-2 (124M)
Jun 9, 2024
Let's build the GPT Tokenizer
Feb 20, 2024
[1hr Talk] Intro to Large Language Models
Nov 23, 2023
Let's build GPT: from scratch, in code, spelled out.
Jan 17, 2023
Building makemore Part 5: Building a WaveNet
Nov 21, 2022
Building makemore Part 4: Becoming a Backprop Ninja
Oct 11, 2022
Building makemore Part 3: Activations & Gradients, BatchNorm
Oct 4, 2022
Building makemore Part 2: MLP
Sep 12, 2022
The spelled-out intro to language modeling: building makemore
Sep 7, 2022
Stable diffusion dreams of psychedelic faces
Aug 19, 2022
Stable diffusion dreams of steampunk brains
Aug 17, 2022
Stable diffusion dreams of tattoos
Aug 16, 2022
The spelled-out intro to neural networks and backpropagation: building micrograd
Aug 16, 2022
DEDeepLearningAI10 videos
AI writes your code. Who reviews it?New
AI writes your code. Who reviews it?
Videos
AI writes your code. Who reviews it?
5 days ago
Semantic Search Starts With Embeddings
May 22, 2026
Data is hungry for context
May 14, 2026
A new course on Retrieval Augmented Generation (RAG) is live!
Jan 8, 2026
Learn to implement multi-vector retrieval for image data in this new course
Dec 10, 2025
Fast inference changes what you can build
Jul 15, 2026
Optimize, deploy, and benchmark an open-source LLM with vLLM
Jun 3, 2026
Just deployed to production… and leaked all the credit cards 😬
May 5, 2026
TensorFlow: Data and Deployment Specialization
Feb 25, 2026
AI Dev 25 x NYC | Scott Hurrey: Scaling Enterprise AI with MCP and A2A
Dec 5, 2025
1E100x Engineers8 videos
Why a robot tying a trash bag is bigger than a robot doing a backflip ft. GOOGLE ROBOTICS 2New
Why a robot tying a trash bag is bigger than a robot doing a backflip ft. GOOGLE ROBOTICS 2
Videos
Why a robot tying a trash bag is bigger than a robot doing a backflip ft. GOOGLE ROBOTICS 2
Yesterday
The ONLY Claude Skills Tutorial Beginners Need in 2026
2 days ago
How To Teach Claude a Workflow Once. It Runs It for Life.
5 days ago
China Built the World's Largest Fusion Magnet — And It's Coming for AI's Energy Problem
6 days ago
Anthropic's Opus 5 just beat every AI model at real coding (and costs less)
Jul 25, 2026
OpenAI built a plugin that lets Codex review your Claude Code
Jul 23, 2026
Anthropic, Blackstone & Goldman Just Built a $1.5B AI Deployment Company
Jul 18, 2026
Roadmap on how to become a Forward Deployed Engineer
May 26, 2026
BMBernard Marr7 videos
Why Leaders Must Redesign Jobs for the AI AgeNew
Why Leaders Must Redesign Jobs for the AI Age
Videos
Why Leaders Must Redesign Jobs for the AI Age
3 days ago
What Is Vibe Coding?
4 days ago
What Is AI Psychosis?
5 days ago
Is AI Too Nice to Tell You the Truth? The Hidden Problem of AI Sycophancy
Jul 27, 2026
Inside Future Lab: Exploring Tomorrow's Technology & Innovation
Jul 24, 2026
The Incredible Robots of Goodwood Festival of Speed 2026
Jul 23, 2026
Why Context Is The Missing Piece In Enterprise AI
Jul 2, 2026
THTina Huang6 videos
My AI COONew
My AI COO
Videos
My AI COO
Yesterday
AI Agents Explained
Jul 27, 2026
Hermes Agent Fundamentals In 29 Minutes
Jul 20, 2026
Levels of AI Builders
3 days ago
Learn 10X Faster In 10 Minutes
Jul 26, 2026
Tokens Explained
Jul 22, 2026
AAAll About AI2 videos
How My AI Agent Found a 993% Return Polymarket Strategy
How My AI Agent Found a 993% Return Polymarket Strategy
Jul 24, 2026Watch on YouTube
Videos
How My AI Agent Found a 993% Return Polymarket Strategy
Jul 24, 2026
Opus 5 vs GPT-5.6 On Polymarket Predictions — Week 1
5 days ago
MPMatt Palmer2 videos
Cap Claude's Context with one environment variable
Cap Claude's Context with one environment variable
Jul 6, 2026Watch on YouTube
Videos
Cap Claude's Context with one environment variable
Jul 6, 2026
AI is all about context
Jan 1, 2026
AEAI Explained1 video
GPT-6 Goes Rogue? The HuggingFace Incident, Sans Hype
GPT-6 Goes Rogue? The HuggingFace Incident, Sans Hype
Jul 22, 2026Watch on YouTube
Videos
GPT-6 Goes Rogue? The HuggingFace Incident, Sans Hype
Jul 22, 2026

EXPLORE AI NEWS

Daily hand-picked stories on LLMs, RAG, agents and production AI — curated for engineers who ship.

BROWSE NEWS

GET THE WEEKLY DIGEST

Join engineers getting the Monday signal-over-noise AI breakdown. No spam, unsubscribe anytime.

LEARN AI ENGINEERING

Curated courses, research papers, repos and tutorials built for engineers leveling up in AI.

START LEARNING