How to Debug, Evaluate, and Ship Reliable AI Agents with LangSmith

​Learn the foundations for understanding, improving, and confidently deploying AI agents. Get practical steps on how to debug non-deterministic agent behavior, iterate on performance, and ship reliably across any frameworks and models. ​What you’ll learn: ​- How to get end-to-end visibility into agent behavior with tracing for each agent step, tool calls, conversation, latency, error, and token count. ​- Practical ways to debug and improve agents using production traces, insights, and iterative prompt/tool refinement. ​- How to set up evaluations (datasets, experiments, and subject matter expert annotations) to measure quality and prevent regressions.

The Agent Development Lifecycle: Build, Test, Deploy, Monitor | Interrupt 26
▶︎

The Agent Development Lifecycle: Build, Test, Deploy, Monitor | Interrupt 26

RL for Agents Workshop - Deep Dive on Training Agents with RL and Open Source
▶︎

RL for Agents Workshop - Deep Dive on Training Agents with RL and Open Source

You’re Thinking About Agent Frameworks Wrong
▶︎

You’re Thinking About Agent Frameworks Wrong

Stop Confusing LangChain, LangGraph, and LangSmith | Full Breakdown
▶︎

Stop Confusing LangChain, LangGraph, and LangSmith | Full Breakdown

💻Beyond ChatGPT and Claude Building Business AI with Microsoft Foundry🤖 with MVP Lewis Prince
▶︎

💻Beyond ChatGPT and Claude Building Business AI with Microsoft Foundry🤖 with MVP Lewis Prince

Building Better AI Agents: Artificial Intelligence Observability How To
▶︎

Building Better AI Agents: Artificial Intelligence Observability How To

LangChain vs LangGraph: A Tale of Two Frameworks
▶︎

LangChain vs LangGraph: A Tale of Two Frameworks

Urgent Update- AI Sputnik Moment: Kimi K3 Released w/ Emad Mostaque | Ep. 272
▶︎

Urgent Update- AI Sputnik Moment: Kimi K3 Released w/ Emad Mostaque | Ep. 272

Helios Is AMD’s First AI System To Rival Nvidia Vera Rubin — We Got An Exclusive, First Look
▶︎

Helios Is AMD’s First AI System To Rival Nvidia Vera Rubin — We Got An Exclusive, First Look

Why Netflix is betting on systems thinkers—not specialists—in the AI era | Elizabeth Stone (CPTO)
▶︎

Why Netflix is betting on systems thinkers—not specialists—in the AI era | Elizabeth Stone (CPTO)

How to Systematically Setup LLM Evals (Metrics, Unit Tests, LLM-as-a-Judge)
▶︎

How to Systematically Setup LLM Evals (Metrics, Unit Tests, LLM-as-a-Judge)

Introducing Omnigent: an open meta-harness – Matei Zaharia, Co-founder and CTO, Databricks
▶︎

Introducing Omnigent: an open meta-harness – Matei Zaharia, Co-founder and CTO, Databricks

Don't learn AI Agents without Learning these Fundamentals
▶︎

Don't learn AI Agents without Learning these Fundamentals

AI Agents Full Course 2026: Master Agentic AI (2 Hours)
▶︎

AI Agents Full Course 2026: Master Agentic AI (2 Hours)

CLAUDE CODE ADVANCED FULL COURSE (3 HOURS)
▶︎

CLAUDE CODE ADVANCED FULL COURSE (3 HOURS)

Why Your AI Agents Keep Forgetting (And How To Fix That) - Vasilije Markovic (Cognee)
▶︎

Why Your AI Agents Keep Forgetting (And How To Fix That) - Vasilije Markovic (Cognee)

Observing & Evaluating Deep Agents Webinar with LangChain
▶︎

Observing & Evaluating Deep Agents Webinar with LangChain

Building OpenCode with Dax Raad
▶︎

Building OpenCode with Dax Raad

Anthropic's Boris Cherny: Why Coding Is Solved, and What Comes Next
▶︎

Anthropic's Boris Cherny: Why Coding Is Solved, and What Comes Next

You Can Learn AI Agent Harness & Loop Engineering In 19 Min | LLM Ops, Eval, Tracing, RAG
▶︎

You Can Learn AI Agent Harness & Loop Engineering In 19 Min | LLM Ops, Eval, Tracing, RAG