Skip to main content
Score agent output on a 1-10 scale with Agent as Judge, using an on_fail callback to handle evaluation failures.
1

Add the following code to your Python file

agent_as_judge_basic.py
2

Set up your virtual environment

3

Install dependencies

4

Export your OpenAI API key

5

Run the example