When the expert fixes your agent's output, where does that fix actually go?
An evals practitioner asks where expert corrections actually land: fixes die in Slack threads and the agent repeats the same mistake weeks later. The discussion centers on closing the feedback loop so human corrections become durable agent behavior, and what to do when two experts disagree.