Why single-turn evaluation is insufficient
Why is evaluating only single-turn interactions a mistake for agent testing?
This exercise is part of the course
Building AI Agent Harnesses with Strands Agents
Hands-on interactive exercise
Turn theory into action with one of our interactive exercises
Start Exercise