Clément Delangue
@ClementDelangue
We turned @OpenAI's math problems into open-source RL environments on @huggingface!
Early and rough (the verifier only accepts the exact original formalization, so an equivalent proof can still score 0), but this is what it looks like when a research release becomes executable https://t.co/PFPtqBPun6