LiveMathematicianBench: A Live Benchmark for Research-Level Mathematical Reasoning with Proof Sketches

arXiv:2604.01754v2 Announce Type: replace-cross Abstract: Mathematical reasoning is a hallmark of human intelligence, and whether large language models (LLMs) can meaningfully perform it remains a central question in artificial intelligence and cognitive science. As LLMs are increasingly integrated…

aiscience

Sources

LiveMathematicianBench: A Live Benchmark for Research-Level Mathematical Reasoning with Proof Sketches · TechNews