We turned @OpenAI's math problems into open-source RL environments on @huggingface!
Early and rough (the verifier only accepts the exact original formalization, so an equivalent proof can still score 0), but this is what it looks like when a research release becomes executable training infrastructure for everyone.
This is an exciting direction imo: turning open research into open executable environments that anyone can build on to train better open models!
https://huggingface.co/datasets/FineEnvs/openai-math

