ended6월 12일· 1 sources

AI Achieves Gold-Medal Level on IMO 2025 Through Population-Level Proof Scaling

M3 모델, IMO 2025에서 인간 금메달 수준의 수학 증명 달성

Why it matters

MaxProof demonstrates that AI can now match human mathematicians on competition-level proof problems by using a novel test-time scaling approach that treats a single model as a generator, verifier, and refiner across a population of candidate proofs. This breakthrough reveals that scaling inference-time resources—rather than training parameters—may be a critical path for advancing AI reasoning on complex formal problems.

1
Sources
+0
24h
Growth
100d
Active
Mathematical ProofGenerative-Verifier RLTest-Time ScalingIMO 2025M3Population Search

Sources

Related Issues