3 Commits (31dcbf5d8af3dd6fb2ee0eb026c62b48a282e03d)

Author SHA1 Message Date
thinhlpg 37730095a9 feat: add eval scripts that compare base model performance with the grpo trained model
1 month ago
thinhlpg a58722e16f feat: add initial project structure and core functionality
2 months ago
Thinh Le bf32fdd897 Initial commit
2 months ago