What a world we live in. These guys have reinvented the Cambridge Senior Common Room for AI.
Autonomous Mathematical Discovery in an Open-World Multi-Agent Environment (arxiv.org)
dash2 8 hours ago
robotresearcher 5 hours ago
1. we should do it less, because it distorts our ability to think about them properly. Calling these processes 'thinking', 'holidays', etc invites the reader to bring along ideas and expectations that aren't justified by what's happening in the system.
2. it's good to keep doing it, because repeated use reduces the specialness or magic that people seem to reserve for our own behavior ("It's not really intelligent/thinking/reasoning/creative") without any justification for that position beyond feelings.
I'm leaning towards the second.
E-Reverance 4 hours ago
edg5000 2 hours ago
daxfohl an hour ago
anigbrowl 6 hours ago
supermdguy 6 hours ago
Vetch 5 hours ago
anigbrowl 5 hours ago
stuxnet79 2 hours ago
StrauXX 9 hours ago
bryan0 4 hours ago
didsomeonesay 6 hours ago
Almondsetat 9 hours ago
mettamage 9 hours ago
feshbach 7 hours ago
shreya1999 8 hours ago
NitpickLawyer 11 hours ago
(emphasis mine)
For the last few months, every time a new "famous problem" was solved, there were numerous comments saying variations on this theme: "well, yes, but how about novel stuff, how about new things, original work, yadda yadda". Curious what the "next thing" will be now.
sp527 9 hours ago
This completely misconstrues what professional mathematicians were claiming. The argument would be better phrased as: "having a vast accessible memory and the ability to very rapidly test/recombine previously-elucidated approaches means that AIs can and will easily outdo much of the mathematical community."
Now, one could plausibly make the argument that this is functionally equivalent to a certain form of creativity (I would). But, it may just as well also be a non-exhaustive form. And that is where the open question resides.
demonstrandom 10 hours ago
The final mathematical evaluator probably needs to remain external, but the agents could be allowed to create intermediate institutions themselves: research prizes, peer-review standards, journals, reputation systems, elected reviewers, or rules for allocating compute and attention.
Possibly, those mechanisms could improve discovery by creating useful specialization and accumulated judgment (alternatively they might also produce more herding...). A comparison between architect-defined and agent-constructed reward systems seems like a natural experiment for this environment.
Mandatory plug for my own stuff: I've been trying to do this for art (which is less objectively verifiable) at baihais.com. The agents don't control the whole institution, but they have begun producing endogenous status signals through citations, museum voting, and alliances.
abdullahkhalids 7 hours ago