Short proofs in combinatorics and number theory

Boris Alexeev, Moe Putterman, Mehtaab Sawhney, Mark Sellke, Gregory Valiant · 2026 · arXiv 2603.29961

5 Pith papers cite this work. Polarity classification is still indexing.

5 Pith papers citing it

read on arXiv browse 5 citing papers

citation-role summary

background 1

citation-polarity summary

background 1

representative citing papers

Soohak: A Mathematician-Curated Benchmark for Evaluating Research-level Math Capabilities of LLMs

cs.CL · 2026-05-09 · unverdicted · novelty 8.0 · 2 refs

Soohak is a 439-problem mathematician-curated benchmark where frontier LLMs reach at most 30.4% on research math challenges and no model exceeds 50% on refusal for ill-posed problems.

Unbounded logarithmic limsup in Erd\H{o}s problem 684

math.NT · 2026-04-26 · unverdicted · novelty 8.0

f(n) exceeds (C-o(1)) log n for any fixed C>1 and infinitely many n, so limsup f(n)/log n is infinite.

AI co-mathematician: Accelerating mathematicians with agentic AI

cs.AI · 2026-05-07 · unverdicted · novelty 7.0

An interactive AI workbench for mathematicians achieves 48% on FrontierMath Tier 4 and helped solve open problems in early tests.

Advancing Mathematics Research with AI-Driven Formal Proof Search

cs.AI · 2026-05-21

Short Proofs in Algebraic and Enumerative Combinatorics

math.CO · 2026-05-19 · 2 refs

citing papers explorer

Showing 5 of 5 citing papers.

Soohak: A Mathematician-Curated Benchmark for Evaluating Research-level Math Capabilities of LLMs cs.CL · 2026-05-09 · unverdicted · none · ref 3 · 2 links
Soohak is a 439-problem mathematician-curated benchmark where frontier LLMs reach at most 30.4% on research math challenges and no model exceeds 50% on refusal for ill-posed problems.
Unbounded logarithmic limsup in Erd\H{o}s problem 684 math.NT · 2026-04-26 · unverdicted · none · ref 1
f(n) exceeds (C-o(1)) log n for any fixed C>1 and infinitely many n, so limsup f(n)/log n is infinite.
AI co-mathematician: Accelerating mathematicians with agentic AI cs.AI · 2026-05-07 · unverdicted · none · ref 40
An interactive AI workbench for mathematicians achieves 48% on FrontierMath Tier 4 and helped solve open problems in early tests.
Advancing Mathematics Research with AI-Driven Formal Proof Search cs.AI · 2026-05-21 · unreviewed · ref 3
Short Proofs in Algebraic and Enumerative Combinatorics math.CO · 2026-05-19 · unreviewed · ref 1 · 2 links

Short proofs in combinatorics and number theory

citation-role summary

citation-polarity summary

fields

years

verdicts

roles

polarities

representative citing papers

citing papers explorer