Commit graph

10 commits

Author SHA1 Message Date
Zafir Stojanovski
58a641e59f lint 2025-02-12 11:21:46 +01:00
Zafir Stojanovski
3d84816f95 system prompt for structured output, and parse such outputs 2025-02-12 10:44:42 +01:00
rishabhranawat
d7b69190ba commit formatting 2025-02-10 22:05:45 -08:00
rishabhranawat
1dc7af587f [eval-v1] benchmark with 50 samples 2025-02-10 22:05:09 -08:00
rishabhranawat
9e4870125d [eval-v1] pre commit formatting 2025-02-10 21:50:22 -08:00
rishabhranawat
df5438498e [eval-v1] add timer 2025-02-10 21:48:44 -08:00
rishabhranawat
247464a47d [eval-v1] async to speed up inference/evaluation 2025-02-10 21:35:46 -08:00
rishabhranawat
0657222a8f [eval-basic] remove large results files, add gitignore, only leave summary 2025-02-09 22:52:10 -08:00
rishabhranawat
c214724a46 [eval-basic] run precommit formatting 2025-02-09 22:40:45 -08:00
rishabhranawat
75cfd31ec2 [eval-basic] initial scripts for evaluating models on reasoning gym 2025-02-09 22:36:27 -08:00