We share our AI model’s proof attempts for the First Proof math challenge, testing research-grade reasoning on expert-level problems.
Advances in neural theorem provers have been impressive, but the successes obscure a broader vision of what AI can do for mathematics and how mathematicians can engage with AI. This essay advances a m...
arXiv:2606. 18119v1 Announce Type: new Abstract: To assess the ability of current AI systems to correctly solve research-level mathematics problems, we tested several AI systems on a set of ten problems in a broad range of mathematical fields; these problems arose naturally in the research process of the contributors.
By Mohammed Abouzaid, Nikhil Srivastava, Rachel Ward, Lauren Williams
arXiv:2608.23218v1 Announce Type: new
Abstract: Advances in neural theorem provers have been impressive, but the successes obscure a broader vision of what AI can do for mathematics and how mathemati...
By Jeremy Avigad
The paper introduces InternGeometry, a large language model agent that achieves medalist-level performance on International Mathematical Olympiad geometry problems. It overcomes traditional heuristic limitations by iteratively proposing and verifying auxiliary constructions with a symbolic engine, supported by a dynamic memory mechanism that allows over 200 interactions per problem. Using Complexity-Boosting Reinforcement Learning, InternGeometry trains on only 13,000 examples—0.004% of the data used by AlphaGeometry 2—and solves 44 of 50 IMO geometry problems, surpassing the average gold medalist score.
By Haiteng Zhao, Junhao Shen, Yiming Zhang, Songyang Gao, Kuikun Liu, Tianyou Ma, Fan Zheng, Dahua Lin, Wenwei Zhang, Kai Chen