Regression Report: 124M
[['?we=openrlbenchmark&wpn=lm-human-preferences&ceik=task_id&cen=task.value.policy.initial_model&metrics=ppo/objective/score&metrics=ppo/objective/kl', '124M']]
Created on June 23|Last edited on June 23
Comment
openrlbenchmark/lm-human-preferences/124M ({})
40
openrlbenchmark/lm-human-preferences/124M ({})
41
Add a comment