Evaluating Agents in Production: Traces, LLM-as-Judge, and Prompt Management at Wonder

LangChain · 2:56

Kartik Arora of Wonder (an AI meal-planning app) describes how LangSmith became core to building and debugging their agent. It moved them from manual log digging to traces, LLM-as-judge evals, prompt management and MC...

Read the full summary on tuber

Redirecting...