4 Commits (55f34b8503a3d7413b365b69311a7ca51d19bb5e)

Author SHA1 Message Date
automaticcat 56911a73f9 Update README.md
3 months ago
thinhlpg 37730095a9 feat: add eval scripts that compare base model performance with the grpo trained model
3 months ago
thinhlpg a58722e16f feat: add initial project structure and core functionality
3 months ago
Thinh Le bf32fdd897 Initial commit
3 months ago