Preference Model
Reinforcement-learning environments that help AI labs train models to solve tasks without exploiting flaws in the grading system.
$16MSeed
ParticipantActivity from announced rounds
Reinforcement-learning environments that help AI labs train models to solve tasks without exploiting flaws in the grading system.