EleutherAI Releases OLMo-3-7B Models for Reward-Hacking Research
EleutherAI releases two OLMo-3-7B models trained via GRPO to study reward-hacking dynamics and validate the hack-igni ...
EleutherAI Releases Bergson Leaderboard Baseline GPT-2 Model
EleutherAI has released the baseline model EleutherAI/bergson-wikitext-gpt2-leaderboard for the training data attribu ...