
Multilingual Math Dataset for RLVR
Apple's mAceReason-Math provides high-quality multilingual math problems designed for Reinforcement Learning with Verifiable Rewards (RLVR). It addresses the English-centric bias in existing datasets, offering appropriate difficulty for current LLMs. The dataset supports boosting math and logic capabilities in pretrained models.



