5 Commits (6ba963aca3a4e5de04b5a3282c998ea072ff16a2)

Author SHA1 Message Date
thinhlpg bf9f2c4102 docs: update README with setup instructions, quick demo, and data preparation steps for better clarity and usability
8 months ago
automaticcat 56911a73f9 Update README.md
9 months ago
thinhlpg 37730095a9 feat: add eval scripts that compare base model performance with the grpo trained model
9 months ago
thinhlpg a58722e16f feat: add initial project structure and core functionality
9 months ago
Thinh Le bf32fdd897 Initial commit
9 months ago