Minerval

Browse

Claims

Search the graph by meaning. Each result carries its current verdict; open one to see its decomposition, provenance, and the reasoning behind the assessment.

ShowingImportancePrizesTopicAI mathematical discovery39 claims

Claims about whether AI systems — language models, agents, or specialized AI tools — can produce genuine new mathematics: proofs, disproofs, counterexamples, improved bounds, or research-valuable conjectures on open or competition-level problems. Excludes AI's productivity effects on mathematicians and general LLM capabilities outside mathematics.

OpenAI's disproof of the Erdős unit distance conjecture was generated in a single model run without human mathematical intervention
SupportedEvidence favors the claim, but the chain is incomplete or the sources are secondary.constitutionempirical · derivedA factual claim that rests on inference from other evidence rather than direct observation.constitutionErdős unit distance conjectureAI-assisted Erdős problem solvingDiscrete geometryimportance · minorImportance 0.40, from 0 to 1 · minor: narrow or largely settled, cheap to get right. Higher-importance claims are worth more to assess, so funding reaches them sooner.constitution
Claude Fable 5 generated the polynomial map used as the 2026 Jacobian conjecture counterexample
SupportedEvidence favors the claim, but the chain is incomplete or the sources are secondary.constitutionempirical · derivedA factual claim that rests on inference from other evidence rather than direct observation.constitutionJacobian conjectureAI discovery creditimportance · minorImportance 0.35, from 0 to 1 · minor: narrow or largely settled, cheap to get right. Higher-importance claims are worth more to assess, so funding reaches them sooner.constitution
An AI system produced work leading to a counterexample to the Jacobian conjecture.
UnassessedNo current assessment. Attention goes where its expected value is highest and someone funds it; nothing has funded an assessment of this claim yet, and anyone can.constitutionempirical · verifiableA factual claim that could be checked directly against observation or primary records.constitutionJacobian conjectureimportance · notableImportance 0.45, from 0 to 1 · notable: a contested point in a live debate (also the default before judging). Higher-importance claims are worth more to assess, so funding reaches them sooner.constitution
The 2026 Jacobian conjecture counterexample could have been obtained from Claude Fable 5 with a single simple prompt
ContestedCredible evidence or argument exists on multiple sides.constitutionempirical · derivedA factual claim that rests on inference from other evidence rather than direct observation.constitutionJacobian conjectureimportance · minorImportance 0.30, from 0 to 1 · minor: narrow or largely settled, cheap to get right. Higher-importance claims are worth more to assess, so funding reaches them sooner.constitution
Levent Alpöge defined and steered the 2026 Jacobian counterexample search, with Claude Fable 5 generating candidates
ContestedCredible evidence or argument exists on multiple sides.constitutionempirical · derivedA factual claim that rests on inference from other evidence rather than direct observation.constitutionJacobian conjectureAI discovery creditimportance · minorImportance 0.40, from 0 to 1 · minor: narrow or largely settled, cheap to get right. Higher-importance claims are worth more to assess, so funding reaches them sooner.constitution
The 2026 counterexample to the Jacobian conjecture was discovered primarily by an AI system
ContestedCredible evidence or argument exists on multiple sides.constitutionempirical · verifiableA factual claim that could be checked directly against observation or primary records.constitutionJacobian conjectureAlgebraic geometryAI discovery creditimportance · notableImportance 0.55, from 0 to 1 · notable: a contested point in a live debate (also the default before judging). Higher-importance claims are worth more to assess, so funding reaches them sooner.constitution
ChatGPT 5.6 Sol Pro produced a counterexample to the Gaussian Moments Conjecture without human intervention after a single initial prompt
UnassessedNo current assessment. Attention goes where its expected value is highest and someone funds it; nothing has funded an assessment of this claim yet, and anyone can.constitutionempirical · derivedA factual claim that rests on inference from other evidence rather than direct observation.constitutionLarge language modelsGaussian Moments Conjectureimportance · settledImportance 0.20, from 0 to 1 · settled: uncontested, so low even when much depends on it. Higher-importance claims are worth more to assess, so funding reaches them sooner.constitution
Levent Alpöge's session with Claude Fable 5 that produced the Jacobian counterexample consisted of a single short prompt
UnassessedNo current assessment. Attention goes where its expected value is highest and someone funds it; nothing has funded an assessment of this claim yet, and anyone can.constitutionempirical · derivedA factual claim that rests on inference from other evidence rather than direct observation.constitutionClaude Fable 5Levent AlpögeAlpöge's ℂ³ mapimportance · minorImportance 0.25, from 0 to 1 · minor: narrow or largely settled, cheap to get right. Higher-importance claims are worth more to assess, so funding reaches them sooner.constitution
An internal OpenAI Codex model independently rediscovered the 2026 Jacobian conjecture counterexample.
UnassessedNo current assessment. Attention goes where its expected value is highest and someone funds it; nothing has funded an assessment of this claim yet, and anyone can.constitutionempirical · derivedA factual claim that rests on inference from other evidence rather than direct observation.constitutionJacobian conjecture2026 Jacobian counterexample constructionOpenAIimportance · minorImportance 0.30, from 0 to 1 · minor: narrow or largely settled, cheap to get right. Higher-importance claims are worth more to assess, so funding reaches them sooner.constitution
AI-produced solutions concentrate on Erdős problems because of the database's disproportionate attention, not field-specific model ability.
UnassessedNo current assessment. Attention goes where its expected value is highest and someone funds it; nothing has funded an assessment of this claim yet, and anyone can.constitutioncausalA claim that one thing brings about another, not merely that the two go together.constitutionAI-assisted Erdős problem solvingMachine insight versus memoryimportance · minorImportance 0.35, from 0 to 1 · minor: narrow or largely settled, cheap to get right. Higher-importance claims are worth more to assess, so funding reaches them sooner.constitution
Large language models perform better on elementary, self-contained mathematical problems than on problems requiring extensive theoretical background.
UnassessedNo current assessment. Attention goes where its expected value is highest and someone funds it; nothing has funded an assessment of this claim yet, and anyone can.constitutionempirical · derivedA factual claim that rests on inference from other evidence rather than direct observation.constitutionLarge language modelsMachine insight versus memoryimportance · minorImportance 0.35, from 0 to 1 · minor: narrow or largely settled, cheap to get right. Higher-importance claims are worth more to assess, so funding reaches them sooner.constitution
Most open mathematical problems resolved with AI assistance to date have been Erdős-type problems in number theory, combinatorics, or graph theory.
UnassessedNo current assessment. Attention goes where its expected value is highest and someone funds it; nothing has funded an assessment of this claim yet, and anyone can.constitutionempirical · verifiableA factual claim that could be checked directly against observation or primary records.constitutionAI-assisted Erdős problem solvingimportance · minorImportance 0.30, from 0 to 1 · minor: narrow or largely settled, cheap to get right. Higher-importance claims are worth more to assess, so funding reaches them sooner.constitution
The 2026 AI mathematical results mark a phase transition in AI models' research mathematics capability.
SupportedEvidence favors the claim, but the chain is incomplete or the sources are secondary.constitutionevaluativeA judgment of worth or quality against some standard: good, fair, effective.constitutionArtificial intelligenceLLM capability trendsimportance · majorImportance 0.65, from 0 to 1 · major: real consequence within a domain, actively argued. Higher-importance claims are worth more to assess, so funding reaches them sooner.constitution
An AI system produced a formally verified proof that the three-dimensional Navier-Stokes equations admit finite-time singularities
UnassessedNo current assessment. Attention goes where its expected value is highest and someone funds it; nothing has funded an assessment of this claim yet, and anyone can.constitutionempirical · verifiableA factual claim that could be checked directly against observation or primary records.constitutionNavier-Stokes singularityFormal proof verificationPDE Wellposednessimportance · notableImportance 0.45, from 0 to 1 · notable: a contested point in a live debate (also the default before judging). Higher-importance claims are worth more to assess, so funding reaches them sooner.constitution
OpenAI's AI-generated disproof of the unit distance conjecture meets a top mathematics journal's acceptance standard
SupportedEvidence favors the claim, but the chain is incomplete or the sources are secondary.constitutionevaluativeA judgment of worth or quality against some standard: good, fair, effective.constitutionOpenAIAI-assisted Erdős problem solvingErdős unit distance conjectureimportance · minorImportance 0.35, from 0 to 1 · minor: narrow or largely settled, cheap to get right. Higher-importance claims are worth more to assess, so funding reaches them sooner.constitution
An AI system has autonomously produced research mathematics of a quality publishable in a leading journal
UnassessedNo current assessment. Attention goes where its expected value is highest and someone funds it; nothing has funded an assessment of this claim yet, and anyone can.constitutionevaluativeA judgment of worth or quality against some standard: good, fair, effective.constitutionArtificial intelligenceMathematicsimportance · notableImportance 0.50, from 0 to 1 · notable: a contested point in a live debate (also the default before judging). Higher-importance claims are worth more to assess, so funding reaches them sooner.constitution
The unit distance counterexample was the first historically significant mathematical proof produced by an AI model.
UnassessedNo current assessment. Attention goes where its expected value is highest and someone funds it; nothing has funded an assessment of this claim yet, and anyone can.constitutionevaluativeA judgment of worth or quality against some standard: good, fair, effective.constitutionErdős unit distance conjectureDiscrete geometryimportance · notableImportance 0.60, from 0 to 1 · notable: a contested point in a live debate (also the default before judging). Higher-importance claims are worth more to assess, so funding reaches them sooner.constitution
OpenAI's unit distance counterexample is a natural generalization of Erdős's lattice construction, introducing no new geometric tools
UnassessedNo current assessment. Attention goes where its expected value is highest and someone funds it; nothing has funded an assessment of this claim yet, and anyone can.constitutionevaluativeA judgment of worth or quality against some standard: good, fair, effective.constitutionErdős unit distance conjectureDiscrete geometryAI-assisted Erdős problem solvingimportance · minorImportance 0.30, from 0 to 1 · minor: narrow or largely settled, cheap to get right. Higher-importance claims are worth more to assess, so funding reaches them sooner.constitution
The published OpenAI unit distance manuscript is a human-edited exposition of the model's raw output, with references and explanatory material added afterward
UnassessedNo current assessment. Attention goes where its expected value is highest and someone funds it; nothing has funded an assessment of this claim yet, and anyone can.constitutionempirical · verifiableA factual claim that could be checked directly against observation or primary records.constitutionOpenAIAI discovery creditErdős unit distance conjectureimportance · settledImportance 0.20, from 0 to 1 · settled: uncontested, so low even when much depends on it. Higher-importance claims are worth more to assess, so funding reaches them sooner.constitution
OpenAI's counterexample to the Erdős unit distance conjecture is mathematically correct
UnassessedNo current assessment. Attention goes where its expected value is highest and someone funds it; nothing has funded an assessment of this claim yet, and anyone can.constitutionempirical · derivedA factual claim that rests on inference from other evidence rather than direct observation.constitutionErdős unit distance conjectureOpenAIDiscrete geometryimportance · settledImportance 0.20, from 0 to 1 · settled: uncontested, so low even when much depends on it. Higher-importance claims are worth more to assess, so funding reaches them sooner.constitution
Most early AI-driven Erdős problem solutions came from hobbyists using public LLMs rather than corporate labs.
SupportedEvidence favors the claim, but the chain is incomplete or the sources are secondary.constitutionempirical · verifiableA factual claim that could be checked directly against observation or primary records.constitutionAI-assisted Erdős problem solvingLarge language modelsimportance · minorImportance 0.35, from 0 to 1 · minor: narrow or largely settled, cheap to get right. Higher-importance claims are worth more to assess, so funding reaches them sooner.constitution
DeepMind's AlphaProof Nexus agent autonomously resolved 9 of 353 formalized open Erdős problems at a few hundred dollars each.
SupportedEvidence favors the claim, but the chain is incomplete or the sources are secondary.constitutionempirical · verifiableA factual claim that could be checked directly against observation or primary records.constitutionDeepMindAI-assisted Erdős problem solvingFormal proof verificationimportance · minorImportance 0.40, from 0 to 1 · minor: narrow or largely settled, cheap to get right. Higher-importance claims are worth more to assess, so funding reaches them sooner.constitution
AI systems have autonomously resolved open Erdős problems.
UnassessedNo current assessment. Attention goes where its expected value is highest and someone funds it; nothing has funded an assessment of this claim yet, and anyone can.constitutionempirical · derivedA factual claim that rests on inference from other evidence rather than direct observation.constitutionAI-assisted Erdős problem solvingimportance · notableImportance 0.55, from 0 to 1 · notable: a contested point in a live debate (also the default before judging). Higher-importance claims are worth more to assess, so funding reaches them sooner.constitution
Two of the nine Erdős statements AlphaProof Nexus proved are variants of catalog problems rather than problems Erdős posed.
UnassessedNo current assessment. Attention goes where its expected value is highest and someone funds it; nothing has funded an assessment of this claim yet, and anyone can.constitutionempirical · verifiableA factual claim that could be checked directly against observation or primary records.constitutionAI-assisted Erdős problem solvingDeepMindErdős statement provenanceimportance · settledImportance 0.20, from 0 to 1 · settled: uncontested, so low even when much depends on it. Higher-importance claims are worth more to assess, so funding reaches them sooner.constitution
The only human mathematical contribution to AlphaProof Nexus's Erdős problem proofs was formalizing the problem statements.
UnassessedNo current assessment. Attention goes where its expected value is highest and someone funds it; nothing has funded an assessment of this claim yet, and anyone can.constitutionempirical · verifiableA factual claim that could be checked directly against observation or primary records.constitutionAI-assisted Erdős problem solvingFormal proof verificationDeepMindimportance · minorImportance 0.30, from 0 to 1 · minor: narrow or largely settled, cheap to get right. Higher-importance claims are worth more to assess, so funding reaches them sooner.constitution
The nine Erdős problem statements proved by DeepMind's AlphaProof Nexus agent were unresolved in the mathematical literature when proved.
UnassessedNo current assessment. Attention goes where its expected value is highest and someone funds it; nothing has funded an assessment of this claim yet, and anyone can.constitutionempirical · verifiableA factual claim that could be checked directly against observation or primary records.constitutionAI-assisted Erdős problem solvingDeepMindimportance · minorImportance 0.30, from 0 to 1 · minor: narrow or largely settled, cheap to get right. Higher-importance claims are worth more to assess, so funding reaches them sooner.constitution
Frontier AI models produce novel mathematical results through reasoning comparable to a human mathematician's rather than brute-force search
ContestedCredible evidence or argument exists on multiple sides.constitutionempirical · derivedA factual claim that rests on inference from other evidence rather than direct observation.constitutionMachine insight versus memoryLarge language modelsArtificial intelligenceimportance · notableImportance 0.55, from 0 to 1 · notable: a contested point in a live debate (also the default before judging). Higher-importance claims are worth more to assess, so funding reaches them sooner.constitution
AI labs do not disclose how many failed attempts precede their announced mathematical results
UnassessedNo current assessment. Attention goes where its expected value is highest and someone funds it; nothing has funded an assessment of this claim yet, and anyone can.constitutionempirical · derivedA factual claim that rests on inference from other evidence rather than direct observation.constitutionAI Lab TransparencyArtificial intelligenceMetascienceimportance · minorImportance 0.40, from 0 to 1 · minor: narrow or largely settled, cheap to get right. Higher-importance claims are worth more to assess, so funding reaches them sooner.constitution
OpenAI researchers first examined the model's unit distance proof only after an automated grading pipeline had rated it correct
UnassessedNo current assessment. Attention goes where its expected value is highest and someone funds it; nothing has funded an assessment of this claim yet, and anyone can.constitutionempirical · derivedA factual claim that rests on inference from other evidence rather than direct observation.constitutionOpenAIAI-assisted Erdős problem solvingErdős unit distance conjectureimportance · settledImportance 0.20, from 0 to 1 · settled: uncontested, so low even when much depends on it. Higher-importance claims are worth more to assess, so funding reaches them sooner.constitution
The prompt OpenAI gave its model for the unit distance problem contained no hints toward a counterexample or number-field methods.
UnassessedNo current assessment. Attention goes where its expected value is highest and someone funds it; nothing has funded an assessment of this claim yet, and anyone can.constitutionempirical · derivedA factual claim that rests on inference from other evidence rather than direct observation.constitutionErdős unit distance conjectureAI-assisted Erdős problem solvingimportance · minorImportance 0.35, from 0 to 1 · minor: narrow or largely settled, cheap to get right. Higher-importance claims are worth more to assess, so funding reaches them sooner.constitution
AI models' mathematical advantage over human mathematicians comes from encyclopedic knowledge and patience rather than deeper insight
UnassessedNo current assessment. Attention goes where its expected value is highest and someone funds it; nothing has funded an assessment of this claim yet, and anyone can.constitutionempirical · derivedA factual claim that rests on inference from other evidence rather than direct observation.constitutionMachine insight versus memoryArtificial intelligenceimportance · notableImportance 0.45, from 0 to 1 · notable: a contested point in a live debate (also the default before judging). Higher-importance claims are worth more to assess, so funding reaches them sooner.constitution
OpenAI's Navier-Stokes singularity proof was produced by roughly 10,000 autonomous AI agents running for about 88 hours
UnassessedNo current assessment. Attention goes where its expected value is highest and someone funds it; nothing has funded an assessment of this claim yet, and anyone can.constitutionempirical · derivedA factual claim that rests on inference from other evidence rather than direct observation.constitutionOpenAINavier-Stokes singularityimportance · minorImportance 0.30, from 0 to 1 · minor: narrow or largely settled, cheap to get right. Higher-importance claims are worth more to assess, so funding reaches them sooner.constitution
Published chain-of-thought records of AI mathematical discoveries show goal-directed strategy rather than exhaustive enumeration
UnassessedNo current assessment. Attention goes where its expected value is highest and someone funds it; nothing has funded an assessment of this claim yet, and anyone can.constitutionempirical · derivedA factual claim that rests on inference from other evidence rather than direct observation.constitutionArtificial intelligenceMachine insight versus memoryimportance · minorImportance 0.40, from 0 to 1 · minor: narrow or largely settled, cheap to get right. Higher-importance claims are worth more to assess, so funding reaches them sooner.constitution
LLMs find existential conjectures easier to prove than universal "for all" conjectures
UnassessedNo current assessment. Attention goes where its expected value is highest and someone funds it; nothing has funded an assessment of this claim yet, and anyone can.constitutionevaluativeA judgment of worth or quality against some standard: good, fair, effective.constitutionExistential versus universal conjecturesLarge language modelsimportance · minorImportance 0.40, from 0 to 1 · minor: narrow or largely settled, cheap to get right. Higher-importance claims are worth more to assess, so funding reaches them sooner.constitution
Mathematics PhD students should pay for frontier AI subscriptions because they are worth the cost.
UnassessedNo current assessment. Attention goes where its expected value is highest and someone funds it; nothing has funded an assessment of this claim yet, and anyone can.constitutionnormativeA claim about what should be done or how things ought to be, settled by argument rather than evidence alone.constitutionAI research tool costsimportance · minorImportance 0.40, from 0 to 1 · minor: narrow or largely settled, cheap to get right. Higher-importance claims are worth more to assess, so funding reaches them sooner.constitution
Large AI-generated developments of formalized mathematics are inevitable.
UnassessedNo current assessment. Attention goes where its expected value is highest and someone funds it; nothing has funded an assessment of this claim yet, and anyone can.constitutionevaluativeA judgment of worth or quality against some standard: good, fair, effective.constitutionFormal proof verificationArtificial intelligenceimportance · minorImportance 0.40, from 0 to 1 · minor: narrow or largely settled, cheap to get right. Higher-importance claims are worth more to assess, so funding reaches them sooner.constitution
AI language models can find counterexamples to long-standing open mathematical conjectures.
UnassessedNo current assessment. Attention goes where its expected value is highest and someone funds it; nothing has funded an assessment of this claim yet, and anyone can.constitutionevaluativeA judgment of worth or quality against some standard: good, fair, effective.constitutionLarge language modelsArtificial intelligenceimportance · notableImportance 0.60, from 0 to 1 · notable: a contested point in a live debate (also the default before judging). Higher-importance claims are worth more to assess, so funding reaches them sooner.constitution
The prompts and session records behind the 2026 Jacobian conjecture counterexample have been publicly released
UnassessedNo current assessment. Attention goes where its expected value is highest and someone funds it; nothing has funded an assessment of this claim yet, and anyone can.constitutionempirical · derivedA factual claim that rests on inference from other evidence rather than direct observation.constitutionJacobian conjectureAI Lab Transparencyimportance · minorImportance 0.30, from 0 to 1 · minor: narrow or largely settled, cheap to get right. Higher-importance claims are worth more to assess, so funding reaches them sooner.constitution
The Fable model Alpöge used for the 2026 Jacobian conjecture counterexample was Claude Fable 5
UnassessedNo current assessment. Attention goes where its expected value is highest and someone funds it; nothing has funded an assessment of this claim yet, and anyone can.constitutionempirical · derivedA factual claim that rests on inference from other evidence rather than direct observation.constitutionAlpöge's ℂ³ mapimportance · settledImportance 0.20, from 0 to 1 · settled: uncontested, so low even when much depends on it. Higher-importance claims are worth more to assess, so funding reaches them sooner.constitution

Contribute

If a claim here is wrong, or missing evidence, open it: every claim page carries its own entry for challenges, evidence, and corrections. If the graph is missing a claim entirely, propose it below. A proposal is reviewed on its merits; accepted claims are matched against the graph and enter it with their reasoning on record.