papersSEP 11 04:00 UTC
Paper Measures AI Progress Toward Mathematical Discovery with Automatic Verification
A revised arXiv preprint introduces a method that uses automatic verification to track how well language models reason about unsolved mathematical problems. The author notes that although large language models now handle sophisticated math and science reasoning, whether they can contribute genuinely new research remains contested and thinly studied. The work aims to give a measurable way to assess progress on that question.