PaTaRM PaTaRM is a Generative Reward Model (GRM) for RLHF alignment. AIJian/PaTaRM-8B Text Generation • 0.5B • Updated 14 days ago • 847 AIJian/PaTaRM-data Preview • Updated 15 days ago • 35 AIJian/PaTaRM-14B Text Generation • 0.5B • Updated 14 days ago • 1.34k AIJian/PaTaRM Updated 15 days ago
PaTaRM PaTaRM is a Generative Reward Model (GRM) for RLHF alignment. AIJian/PaTaRM-8B Text Generation • 0.5B • Updated 14 days ago • 847 AIJian/PaTaRM-data Preview • Updated 15 days ago • 35 AIJian/PaTaRM-14B Text Generation • 0.5B • Updated 14 days ago • 1.34k AIJian/PaTaRM Updated 15 days ago