Skip to main content
Demonstrates that eval model metrics can be accumulated into the original agent’s run_output using the run_metrics parameter on evaluate_answer.
accuracy_eval_metrics.py

Run the Example

1

Set up your virtual environment

2

Install dependencies

3

Export your OpenAI API key

4

Run the example

Save the code above as accuracy_eval_metrics.py, then run:
Full source: cookbook/09_evals/accuracy/accuracy_eval_metrics.py