LangChain · AI
Evaluating Agents in Production: Traces, LLM-as-Judge, and Prompt Management at Wonder
Kartik Arora, AI Engineer at Wonder, explains how his team builds a generative AI application that takes every aspect of your life into account and plans all 21 of your meals for the week, then gets them delivered. He walks through what changed when they moved off manual log debugging to LangSmith traces, evaluators, and LLM-as-judge scoring of how their agent responds to different users, and how prompt management took a feedback loop that used t

Introductie van de bron.
LangChain