4 Commits (1a18cd7bfdffe9971714887012b7a9b4e7725e81)

Author SHA1 Message Date
thinhlpg d2f03b96ab feat: enhance evaluation script and remove deprecated shell script
1 month ago
thinhlpg c90c03267e feat: change user prompt template to search-r1 inspried format
1 month ago
thinhlpg 04593fa8fd style: change line length to 119, organize imports
1 month ago
thinhlpg 37730095a9 feat: add eval scripts that compare base model performance with the grpo trained model
1 month ago