Evaluating Agents in Production: Traces, LLM-as-Judge, and Prompt Management at Wonder
LangChain · 2:56
Kartik Arora of Wonder (an AI meal-planning app) describes how LangSmith became core to building and debugging their agent. It moved them from manual log digging to traces, LLM-as-judge evals, prompt management and MC...