Announcement_15

🚀🚀🚀 We are pleased to release Magenta, an end-to-end formal-informal self-correction pipeline that achieves 100% mathematical benchmarks, marking the first time a fully open-source model has achieved a perfect score. In addition, when using a 7B model for informal reasoning, the pipeline achieved a perfect score on IMO 2026, making it the smallest model to date to do so.

For more details, please see our recent work, Magenta: Closing the Loop Between Mathematical Reasoning and Lean Verification🔥.