Mathematicians Create Nonprofit Archive for AI-Generated Proofs arXiv Declined to Host

October 9, 2026:

Mathematicians Create Nonprofit Archive for AI-Generated Proofs arXiv Declined to Host
ChatGPT Image Oct 7, 2026, 08_16_51 PM
Hexagonmath.org

When a research group at Northwestern clicked a button last summer and discovered a new result about the derived invariance of Hodge numbers of fourfolds in characteristic zero, lead mathematician Benjamin Antieau faced an awkward problem: the finding was real, but he was not willing to put his name to it. “In the past, this would be a result I would have submitted to an excellent, but not top, journal,” Antieau wrote. Now, because a large language model had produced it without any mathematician claiming genuine understanding, the best he could do was post about it on his blog — what he called a brittle way to aggregate mathematics.

That gap — between AI-generated mathematical knowledge and the infrastructure capable of housing it — is what Antieau and a coalition of elite mathematicians have now moved to fill. On October 6, 2026, Fields Medalist Terence Tao cross-posted on his widely-read blog What’s new an announcement from Antieau introducing Hexagon math repository, a new repository for research-level works in the mathematical sciences and theoretical computer science that explicitly accepts submissions where the contributor claims no human understanding whatsoever.

The launch signals something larger than a new preprint server. Hexagon’s willingness to archive results that no mathematician will claim as their own means the community has, in effect, decided that LLMs are now producing genuine mathematical knowledge — not merely assisting human reasoning — and that the field needs its own infrastructure to receive it, on the community’s terms rather than on those of any AI company.

Math’s New Publishing Problem

Hexagon was built to solve a problem that emerged gradually and then all at once. LLMs running on frontier models have begun surfacing mathematical results — not as the primary goal of a computation but as side effects, adjacent discoveries that bubble up in the course of working through a related problem. These results may be real, may be verifiable, may be of genuine interest to researchers in algebraic geometry, combinatorics, or theoretical computer science. But under the norms that have governed mathematical publishing since at least 1991, when Paul Ginsparg launched arXiv at Los Alamos, there is no appropriate home for them.

arXiv has never conducted peer review — it uses human moderation — but it has always required that submissions reflect human intellectual contribution. Submitting entirely LLM-generated content is now explicitly prohibited, and on October 1, 2026, arXiv went further: it announced that all authors face new submission caps, limited to two submissions per calendar month, with a cap of three total active submissions at any given time, in response to what it described as a “watershed moment” in AI-enabled paper volumes. The platform received 40,363 submissions in September 2026 alone — more than four times the volume of the same month a decade earlier — and its volunteer moderators were overwhelmed.

When the Hexagon founding group approached arXiv leadership during development, the answer was clear. “From the arXiv we learned that they were not going to provide the home for these results we were looking for,” Antieau wrote. “Instead they are focused on maintaining their project as it is today, and I am happy for that.”

What Hexagon Actually Does Differently

Hexagon is not merely a permissive arXiv. Its technical architecture and submission policies are deliberately designed to draw a sharp distinction between what it is, what it isn’t, and who is accountable for what.

The platform distinguishes between two roles. Contributors are humans or organizations responsible for creating a submission — this includes AI labs that produce mathematical results as a byproduct of training or inference. Submitters are the humans with ORCID-validated accounts who actually file the submission form and serve as the point of contact for revisions and queries. ORCID (Open Researcher and Contributor ID) is an open, nonprofit researcher identifier that assigns each researcher a persistent 16-digit code — essentially a social security number for scientists — that prevents name ambiguity in academic records and ensures there is always a human accountable for what was filed.

Every submission on the live Hexagon math site carries a pair of disclosure tags that do not exist at arXiv: one indicating the nature of the text (“Primarily AI-generated text,” “A mix of human-written and AI-generated text,” or “Primarily human-written text”) and one indicating the level of human understanding claimed (“Human understanding: all parts,” “some parts,” or “no parts”). Browsing the newest works on October 7, 2026, this reporter observed results carrying the combination “Primarily AI-generated text / Human understanding: no parts” — submissions where a researcher has, in effect, vouched for the result’s potential interest to the community while explicitly disclaiming comprehension of it.

The platform was built using OpenAI’s Codex and hosted on commercial cloud infrastructure. Antieau notes that the engineering is not the difficult part — “This is rather easy nowadays” — and that most of the effort went into deciding exactly what features and policies the community actually needed. Rate limits begin at one submission per day and grow based on a submitter’s prior moderation track record; moderators can grant increases in justified cases. Both read and write API access are provided — Hexagon explicitly invites bulk integrations from “large labs working on language models.”

Who Is Backing It

The Hexagon Mathematics Foundation is a Delaware-registered nonprofit currently applying for 501(c)(3) tax-exempt status. Its governance roster is notably serious for a project assembled over the course of a few months.

The Board of Directors consists of Mohammed Abouzaid (Columbia), François Charles (Université Paris-Saclay), Bryna Kra (Northwestern), David Savitt (Johns Hopkins), and Lauren Williams (Harvard), with Antieau serving as the first Executive Director.

The Advisory Board includes Kevin Buzzard (Imperial College London), Akhil Mathew (University of Chicago), Johannes Schmitt (University of Zürich), Steinn Sigurðsson (Penn State), Nikhil Srivastava (UC Berkeley), Ravi Vakil (Stanford), and Rachel Ward (UT Austin).

The group convened in early August 2026, brought together by overlapping social and professional networks around the shared question of what to do with AI-generated mathematical results. They built and tested the platform through August and September, consulting with arXiv leadership along the way, and opened it to public submissions on September 28, 2026. Funding so far has come entirely from the founding members, with no outside investment or commercial relationships beyond cloud infrastructure costs.

Tao’s Endorsement and the Broader Debate

Terence Tao’s decision to host the Hexagon announcement on What’s new — his blog with a large international readership among mathematicians — carries real institutional weight. Tao has been among the most visible mathematicians to engage publicly with AI as a research tool, and in a paper presented at the 2026 International Congress of Mathematicians in Philadelphia — published as Tao’s ICM paper (arXiv:2608.16753) — he described the discipline’s emerging challenge as one of “proof indigestion”: AI systems are now capable of generating proofs faster than the community can read, verify, and canonicalize them. Hexagon’s place in this framework is explicit: Antieau describes it as where mathematical ideas “emerge from the water for the first time,” with the expectation that results of merit will later be “written up in forms suitable for arXiv submission, publication, and canonicalization.” The reference to Tao’s framework is direct — Antieau cites it in the announcement.

Not all mathematicians are ready to treat AI-generated results as mathematical knowledge, even with better infrastructure. On June 2, 2026, the Leiden Declaration on Mathematics was released, endorsed by the International Mathematical Union and signed by more than 2,654 researchers — including Tao himself. The declaration identified five properties it considers foundational to trustworthy mathematics: proof as a basis for certainty and understanding, clear attribution and accountability, transparency for independent verification, shared community standards for significance, and researcher autonomy. It explicitly stated that AI should not be listed as an author, and that “credit and responsibility should remain with humans belonging to the mathematical community.” Hexagon does not contradict that principle — it preserves the human Submitter as the accountable party — but it does allow results where that accountable human explicitly disclaims understanding, something the Leiden Declaration’s framework does not directly address.

The tension between these two positions — building community infrastructure for AI-generated knowledge vs. protecting the epistemic integrity of mathematics as a human practice — defines where the field sits in October 2026. Hexagon’s emergence as a serious institutional response, with a Delaware nonprofit, ORCID validation, human moderation, rate limits, and API access, suggests the community is moving toward the former: building, not waiting.

Why “Hexagon”?

The name comes from Jorge Luis Borges’s short story “The Library of Babel,” in which an infinite library contains all possible books, arranged in hexagonal galleries. The founding group initially considered names involving “Babel” but concluded that name carried connotations they preferred to avoid. They settled on Hexagon — a reference to the shape of the library’s halls — as the platform committed to housing all possible AI-generated mathematical results. “There was no personal problem, no world problem, whose eloquent solution did not exist — somewhere in some hexagon,” the announcement opens, quoting Borges’s infinite library.

The allusion to Borges is fitting. Like the Library of Babel, Hexagon does not guarantee the truth of what it contains, only that it will be preserved, organized, and retrievable. Unlike the Library, it has human moderators, rate limits, ORCID-validated submitters, and an explicit expectation that the most valuable results will eventually earn their way into more rigorous venues.

How the Platform Currently Works

Hexagon accepts submissions in the standard mathematical subject categories that mirror arXiv’s classification: from math.AG (Algebraic Geometry) through math.NT (Number Theory), plus a range of theoretical computer science subjects. Formalization is not required — Hexagon is not a formalization project, though it links to proof-verification systems including Mathlib, Palomar, prove2.me, and TauCeti for submitters who want their results formally verified.

The platform explicitly welcomes what Antieau calls “micro-results” — small, incremental improvements that are of interest to the mathematics and TCS communities but do not rise to the level of significance or novelty that a journal submission requires. This is a meaningful departure from the norms that have governed preprint publishing. A result that is real, verifiable, and useful to someone working in the field, but insufficiently novel to justify even an arXiv submission under the current quality bar, now has a permanent citable home. Its identifier (e.g., hexagon:2609.00104) can be referenced in future work.

The first result Antieau chose to archive there — the Hodge numbers fourfold finding discovered by button-click — carries the notation hexagon:2609.00104v1. The live Hexagon math site as of October 7, 2026 shows dozens of submissions across algebraic geometry, representation theory, quantum algebra, and combinatorics, with active submissions arriving daily.

What This Means for AI Labs

Among the least-discussed elements of the Tao-hosted Hexagon announcement is this line: “Hexagon welcomes submissions by large labs working on language models.”

This is a direct invitation to OpenAI, Google DeepMind, Anthropic, and others whose systems produce mathematical results as a byproduct of development and evaluation. Until now, those results have had nowhere to go — they cannot be submitted to arXiv under current policy, they cannot be submitted as journal papers without human authorship, and posting them to GitHub or social media provides no lasting citability. Hexagon is proposing that they flow instead into community-maintained infrastructure, where mathematicians retain moderation rights over what appears.

Whether the labs will take up the invitation — and whether the community will be satisfied with the resulting volume of submissions — remains to be seen. But the structural logic is clear: if AI labs are going to generate mathematical knowledge anyway, Hexagon would rather that knowledge be preserved where mathematicians can access it, evaluate it, and build on it, than scattered across model cards and blog posts.


Frequently Asked Questions

Can I submit a result I found using an AI without understanding the proof?

Yes, that is precisely what Hexagon is designed for. The platform accepts submissions tagged “Primarily AI-generated text / Human understanding: no parts” — you are not required to understand the result, only to believe in good faith that it is correct, that it is of potential interest to the mathematics or theoretical computer science community, and that it accurately cites prior work that should be cited. You do need an ORCID-validated account to file the submission, and you serve as the point of contact for questions and revisions. The result receives a permanent identifier (e.g., hexagon:2610.XXXXX) and can be cited in future work. Further details are available on the Hexagon about page.

How is Hexagon different from just posting on arXiv?

arXiv requires submissions to reflect human authorship and intellectual contribution; since May 2026, it has banned content where authors did not verify what the AI produced, and since October 1, 2026, it limits authors to two monthly submissions. Hexagon has no human-authorship requirement, accepts fully AI-generated results, and sets its own rate limits per submitter (starting at one per day) rather than a universal cap. It also explicitly welcomes small, incremental results that do not clear the significance threshold arXiv expects. The two platforms are meant to serve different moments in the lifecycle of a mathematical result: Hexagon math repository is the first stop; arXiv, peer-reviewed journals, and formal canonicalization come later, for the results that earn it.

Does this mean AI is now considered a co-author in mathematics?

Not under current law, and not under Hexagon’s framework. Under copyright law and the COPE guidelines that govern academic publishing, authorship requires a human agent who can bear accountability — AI cannot be an author. Hexagon navigates this by assigning accountability to the human Submitter, who is ORCID-validated and serves as the point of contact. The platform’s contribution categories distinguish between Contributors (who produced the work, which may include AI systems or organizations) and Submitters (the humans responsible for filing and following up). What Hexagon does accept that no existing venue does is a result where the human Submitter explicitly disclaims understanding of the work they are submitting — a category that raises real questions about what verification and accountability mean when human comprehension is optional.

What is the Leiden Declaration, and does Hexagon conflict with it?

The Leiden Declaration on Mathematics, published June 2, 2026, and signed by more than 2,654 mathematicians including Terence Tao, calls on the mathematical community to protect proof integrity, clear attribution, transparency, shared standards, and researcher autonomy in the face of AI-generated results. It states that AI should not be listed as an author and that credit and responsibility must remain with humans. Hexagon does not list AI as an author and preserves human accountability through its Submitter system. Where Hexagon goes further than the Leiden Declaration contemplates is in its acceptance of results where the human Submitter explicitly claims no understanding — a philosophical stance the declaration’s framework does not directly address. The two are not in direct conflict, but they represent different emphases: the declaration is primarily protective; Hexagon is primarily infrastructural.

Source link