I've heard some complaints about the frontier models still be bad at explaining math and was thinking of an RL environment that would help might be to: -Take very hard math problem with a verifiable answer -Have frontier model explain to a tiny model like (0.

Source: [Hacker News](https://news.ycombinator.com/item?id=49326422)

Sponsored