Mathematics in the age of AI

Published
Source
arXiv
Paper number
940
Field
Research
arXiv ID
2608.16753

Key points

  • Taking the premise that AI can perform research-level mathematics without debating AI capability itself, the paper argues that the mathematics community should explicitly state the goals and values it has long advanced implicitly.
  • In the second batch of the First Proof project, seven of ten problems received an effectively perfect passing assessment from at least one AI system.
  • The compute cost of these evaluations ranged from tens to hundreds of dollars per problem.
  • The paper warns that AI is adept at finding gaps between measurements and their intended targets, making the risk of Goodhart's law especially acute: once a measure becomes a target, it ceases to be a good measure.
  • In an age of abundance in which proofs are plentiful, the authors propose shifting cultural emphasis from generating proofs to digesting them through explanation, review, and incorporation into the canon.

Paper links

External research summaries. These are not HDATF publications or measured product results.

Read original (opens in a new tab)