Autonomous Mathematical Discovery in an Open-World Multi-Agent Environment
Abstract
We study autonomous mathematical discovery in the Station, an open-world multi-agent environment in which AI agents from different model families pursue a shared research goal without a central coordinator or scripted pipeline. Agents choose their own research directions, conduct experiments, collaborate, and build a shared scientific literature. Across 12 construction problems from the AlphaEvolve catalogue and two additional case studies, the Station obtained results novel relative to the prior literature on five problems: a new infinite family of finite-field Kakeya sets, new exact 604-point kissing configurations in dimension 11, new records for the discretized Kakeya needle and sign uncertainty problems, and a substantially improved lower bound for Erdős's minimum-overlap problem. Agents also discovered novel infinite families for Book Ramsey numbers. Importantly, the agents produced not only numerical constructions but also theorems and analyses explaining how those constructions work, making the results more interpretable and easier for mathematicians to build upon. We release all raw agent dialogues, proofs, and verification code, providing a transparent record of how these discoveries emerged.
Community
The Station is an open-world multi-agent environment with no central agent. Given only a research goal, agents chose their own directions, ran experiments, and built a shared literature—advancing mathematics beyond the known record. Feel free to discuss and comment!
Project page: https://dualverse.ai/station
This is an automated message from the Librarian Bot. I found the following papers similar to this paper.
The following papers were recommended by the Semantic Scholar API
- ProofCouncil: An LLM Agent for Solving Open Mathematical Problems (2026)
- Evaluating SageMath-Augmented LLM Agents for Computational and Experimental Mathematics (2026)
- Scaling Scientific Discovery Environments for Turn-Level Agentic RL (2026)
- VALG: An Agentic System for ML Theory Research (2026)
- Continuous Improvement and Parallel Autonomous Exploration: An LLM-Agent Framework for Searching Large Solution Spaces (2026)
- A Vocabulary for Multi-Agent Automated Research Systems (2026)
- MechMath Agent Team: LLM Driven Agents for Mathematical Research (2026)
Please give a thumbs up to this comment if you found it helpful!
If you want recommendations for any Paper on Hugging Face checkout this Space
You can directly ask Librarian Bot for paper recommendations by tagging it in a comment: @librarian-bot recommend
Get this paper in your agent:
hf papers read 2608.23691 Don't have the latest CLI?
curl -LsSf https://hf.co/cli/install.sh | bash Models citing this paper 0
No model linking this paper
Datasets citing this paper 0
No dataset linking this paper
Spaces citing this paper 0
No Space linking this paper