Mathematics in the age of AI
- Published
- Source
- arXiv
- Paper number
- 940
- Field
- Research
- arXiv ID
- 2608.16753
Key points
- Taking the premise that AI can perform research-level mathematics without debating AI capability itself, the paper argues that the mathematics community should explicitly state the goals and values it has long advanced implicitly.
- In the second batch of the First Proof project, seven of ten problems received an effectively perfect passing assessment from at least one AI system.
- The compute cost of these evaluations ranged from tens to hundreds of dollars per problem.
- The paper warns that AI is adept at finding gaps between measurements and their intended targets, making the risk of Goodhart's law especially acute: once a measure becomes a target, it ceases to be a good measure.
- In an age of abundance in which proofs are plentiful, the authors propose shifting cultural emphasis from generating proofs to digesting them through explanation, review, and incorporation into the canon.
Paper links
External research summaries. These are not HDATF publications or measured product results.