AI Agents in Open-World 'Station' Discover New Math Results
Autonomous Mathematical Discovery in an Open-World Multi-Agent Environment
Researchers from DualVerse AI and collaborators introduced the Station, an open-world multi-agent environment where AI agents from different model families autonomously pursue shared research goals without central coordination. Across 12 construction problems and two case studies, the agents achieved novel results on five problems, including new infinite families of Kakeya sets, exact kissing configurations, and improved bounds for several mathematical problems. Notably, they produced theorems and analyses explaining their constructions, enhancing interpretability. All raw dialogues, proofs, and code are publicly released.
Agents also discovered novel infinite families for Book Ramsey numbers.
- robotresearcher
I have two thoughts simultaneously about the anthropomorphisation of these systems:
1. we should do it less, because it distorts our ability to think about them properly. Calling these processes 'thinking', 'holidays', etc invites the reader to bring along ideas and expectations that aren't justified by what's happening in the system.
2. it's good to keep doing it, because repeated use reduces the specialness or magic that people seem to reserve for our own behavior ("It's not really intelligent/thinking/reasoning/creative") without any justification for that position beyond feelings.
I'm leaning towards the second.
- dash2
> Agents were also periodically
given holidays, during which they set aside their ongoing work and received random prompts designed to
encourage open-ended thought.
What a world we live in. These guys have reinvented the Cambridge Senior Common Room for AI.
- anigbrowl
If you haven't read Greg Egan's Permutation City, the fact that you clicked on this discussion means you'll get get a lot out of it.
- StrauXX
Reminds me a lot of this LW piece. https://www.lesswrong.com/posts/znbfRXHq285nS7NAh/the-terrar...
- daxfohl
That's exciting, and kind of makes sense in retrospect. Sometimes a "fresh pair of eyes" on a problem can be all you need. Someone who comes in with a different background and can understand the problem in different terms and work on it from a different angle. It doesn't even have to be them doing the work, just a "that kind of reminds me of ... did you think about trying something like that?" that can get a team unstuck after thinking about it in the same way and never making progress.