Preference Model
Reinforcement-learning environments that help AI labs train models to solve tasks without exploiting flaws in the grading system.
$16MSeed
ParticipantActivity from announced rounds
Reinforcement-learning environments that help AI labs train models to solve tasks without exploiting flaws in the grading system.
Develops AI-native security platform for software vulnerability detection and fix recommendations
Pitching Julian Schrittwieser? Send the deck from a free RoundOS data room and see which pages they read.
Create a free room