Commit graph

2 commits

Author SHA1 Message Date
Andreas Köpf
bfa5f8078b
Eval N completions per prompt (#374)
* feat: Add support for generating multiple completions per prompt
* feat: Track best and mean scores for multiple completions per prompt
* feat: Add checkpoint and resume functionality to evaluation script
2025-03-15 16:39:36 +01:00
Andreas Koepf
4109b5b72c update eval yaml config files 2025-03-10 00:48:32 +01:00
Renamed from eval/yaml/openai-o3.yaml (Browse further)