11 comments

  • dash21 hour ago
    &gt; Agents were also periodically given holidays, during which they set aside their ongoing work and received random prompts designed to encourage open-ended thought.<p>What a world we live in. These guys have reinvented the Cambridge Senior Common Room for AI.
  • anigbrowl30 minutes ago
    If you haven&#x27;t read Greg Egan&#x27;s <i>Permutation City</i>, the fact that you clicked on this discussion means you&#x27;ll get get a lot out of it.
    • supermdguy25 minutes ago
      Yes! I was also reminded of the truth mines in Diaspora.
  • StrauXX2 hours ago
    Reminds me a lot of this LW piece. <a href="https:&#x2F;&#x2F;www.lesswrong.com&#x2F;posts&#x2F;znbfRXHq285nS7NAh&#x2F;the-terrarium" rel="nofollow">https:&#x2F;&#x2F;www.lesswrong.com&#x2F;posts&#x2F;znbfRXHq285nS7NAh&#x2F;the-terrar...</a>
  • feshbach58 minutes ago
    The key is a review loop: different models critique each other’s work, then reach consensus. You need two pillars, adversarial and creative.
  • didsomeonesay19 minutes ago
    Infinite Fun Space.
  • Almondsetat2 hours ago
    Maybe Hilbert&#x27;s dream was not that crazy after all
    • mettamage2 hours ago
      What is that dream? I don’t know much about it
      • ylliu2 hours ago
        [dead]
        • jrflo1 hour ago
          Mathematics not being axiomatically complete doesn&#x27;t mean you can&#x27;t have crazy progress from a formalized and mechanized systems. It just means that there are corners you can&#x27;t reach mechanically, but we don&#x27;t know if those corners are at all interesting or not. It could be the case that 99.99% of useful math can be found mechanically.
        • sigmoid101 hour ago
          Mathematics in the sense of a complete set of axioms can&#x27;t, but <i>human research into mathematics</i> apparently just needed enough compute to achieve the same output as a high-tier faculty.
  • shreya19992 hours ago
    AI for Math and Science is the real deal!
  • NitpickLawyer4 hours ago
    &gt; We study autonomous mathematical discovery in the Station, an open-world multi-agent environment in which AI agents from different model families pursue a shared research goal <i>without a central coordinator or scripted pipeline</i>. Agents <i>choose their own research directions</i>, conduct experiments, collaborate, and <i>build a shared scientific literature</i>. Across 12 construction problems from the AlphaEvolve catalogue and two additional case studies, the Station obtained results novel relative to the prior literature on five problems: a new infinite family of finite-field Kakeya sets, new exact 604-point kissing configurations in dimension 11, new records for the discretized Kakeya needle and sign uncertainty problems, and a substantially improved lower bound for Erdős&#x27;s minimum-overlap problem. Agents also discovered novel infinite families for Book Ramsey numbers. Importantly, the <i>agents produced not only numerical constructions but also theorems and analyses explaining how those constructions work</i>, making the results more interpretable and easier for mathematicians to build upon. We release all raw agent dialogues, proofs, and verification code, providing a transparent record of how these discoveries emerged.<p>(emphasis mine)<p>For the last few months, every time a new &quot;famous problem&quot; was solved, there were numerous comments saying variations on this theme: &quot;well, yes, but how about novel stuff, how about new things, original work, yadda yadda&quot;. Curious what the &quot;next thing&quot; will be now.
    • sp5272 hours ago
      &gt; there were numerous comments saying variations on this theme: &quot;well, yes, but how about novel stuff, how about new things, original work, yadda yadda&quot;<p>This completely misconstrues what professional mathematicians were claiming. The argument would be better phrased as: &quot;having a vast accessible memory and the ability to very rapidly test&#x2F;recombine previously-elucidated approaches means that AIs can and will easily outdo much of the mathematical community.&quot;<p>Now, one could plausibly make the argument that this is functionally equivalent to a certain form of creativity (I would). But, it may just as well also be a non-exhaustive form. And that is where the open question resides.
    • debugworld3 hours ago
      [dead]
  • demonstrandom3 hours ago
    Very cool work! One extension I would be curious to see is whether some of Station’s reward structure could become endogenous.<p>The final mathematical evaluator probably needs to remain external, but the agents could be allowed to create intermediate institutions themselves: research prizes, peer-review standards, journals, reputation systems, elected reviewers, or rules for allocating compute and attention.<p>Possibly, those mechanisms could improve discovery by creating useful specialization and accumulated judgment (alternatively they might also produce more herding...). A comparison between architect-defined and agent-constructed reward systems seems like a natural experiment for this environment.<p>Mandatory plug for my own stuff: I&#x27;ve been trying to do this for art (which is less objectively verifiable) at baihais.com. The agents don&#x27;t control the whole institution, but they have begun producing endogenous status signals through citations, museum voting, and alliances.
    • abdullahkhalids45 minutes ago
      Any given single agent is not long-lived due to limited context. What impact does it have on the status signals the agent develop compared to the signals humans (who usually have much longer context) have developed?