Commit Graph

4 Commits (3910ef343a4df9e41f7692f67bcce1334d15241d)

Author SHA1 Message Date
thinhlpg 31dcbf5d8a feat: refactor whole code base, add logic for training R1 distil base models, change some template and reward logics 2 months ago
thinhlpg da79e986b6 feat: add new script and functionality in train script to save model in 16 bit format 2 months ago
thinhlpg 04593fa8fd style: change line length to 119, organize imports 2 months ago
thinhlpg 3c2deaced9 refactor: restructure code base, better centralize logging logic 3 months ago