We turned @OpenAI's math problems into open-source RL environments on @huggingface!
Early and rough (the verifier only accepts the exact original formalization, so an equivalent proof can still score 0), but this is what it looks like when a research release becomes executable training infrastructure for everyone.
This is an exciting direction imo: turning open research into open executable environments that anyone can build on to train better open models!
