arXiv AI

Categorizing Mathematical Concepts with LLM Voting Ensembles in Mathswitch

arXiv:2606. 28815v1 Announce Type: cross Abstract: Mathswitch is an open-source project that imports mathematical concept records from sources such as Wikidata, Wikipedia, MathWorld, Encyclopedia of Mathematics, nLab, ProofWiki, and Agda-Unimath, and links records that refer to the same concept.

arXiv AI
Sep 25

Learning to Discover Interesting Mathematics

The paper introduces a method for evaluating the intrinsic interestingness of mathematical theorems by comparing the length of their proofs to the length of their statements. It trains a 27B language model to predict proof difficulty, enabling the generation and selection of more interesting theorems while significantly reducing overlap with existing Mathlib. The approach allows iterative expansion of a self‑building, machine‑verified mathematical library guided by quantifiable metrics.

By Niket Patel, Ahmad Rammal, Amaury Hayat, Remi Munos, Julia Kempe
arXiv Computation and Language
Aug 24

Beyond Gold Standards: Epistemic Ensemble of LLM Judges for Formal Mathematical Reasoning

The paper introduces an epistemically and formally grounded ensemble (EFG) of large language model judges to evaluate autoformalization tasks in formal mathematics. It defines four criteria—logical preservation, mathematical consistency, formal quality, and formal validity—to provide a transparent, multi‑granular assessment. Experiments show that this ensemble outperforms coarse‑grained models, offering a scalable and interpretable proxy for evaluating formal mathematical reasoning.

By Lan Zhang, Marco Valentino, Jordan Meadows, Andre Freitas
arXiv AI
Sep 18

WiCleanData: Guaranteeing the Type Consistency of Wikidata by Taxonomy Refinement and Constraint Enforcement

WiCleanData is a refined version of Wikidata that eliminates type inconsistencies and constraint violations. The authors built an automated pipeline that cleans the taxonomy with language‑model assistance, aggregates type constraints hierarchically, and filters facts to ensure no type violations remain. The resulting knowledge graph is publicly available through a web interface for easy exploration and downstream use.

By Yiwen Peng (IP Paris), Marc Jeanmougin (IP Paris), Thomas Bonald (IP Paris)
arXiv AI
Sep 24

Math Reasoning in LLMs is Organized by Approach, Not Topic

The paper argues that large language models (LLMs) organize their internal mathematical reasoning by reusable reasoning approaches rather than by the benchmark topics they are tested on. Using a generation‑replay protocol, the authors extract activation‑importance signatures from eight models across five math sources, cluster these signatures, and find that the resulting groups align more closely with reasoning approaches than with topics. The study shows that changing the requested reasoning approach shifts cluster assignments, while paraphrasing the prompt does not, underscoring the primacy of approach over topic in LLM reasoning.

By Sajad Goudarzi, Samaneh Zamanifard, Moloud Nasiri, Hamed Rahimian
arXiv AI
Sep 2

The zbMATH Open Knowledge Graph: Tracing Centuries of Mathematical Research

The zbMATH Open Knowledge Graph is a large-scale RDF knowledge graph that spans more than 250 years of mathematical scholarship. It goes beyond traditional bibliographic metadata by incorporating expert-curated semantic content such as reviews, keywords, subject classifications, software references, and disambiguated authorship. With 34 million entities and 168 million RDF triples, the graph enables fine-grained, historically grounded exploration of mathematical concepts, research fields, and scholarly relationships over time.

By Yuni Susanti, Moritz Schubotz