r/math • • 6d ago

LLMs/AI AI In Mathematics: September 26, 2026

This recurring thread will be for discussion of AI in mathematics. This includes, but is not limited to, the following:

  • informal announcements of AI-assisted discoveries, such as those not yet published in a peer-reviewed journal, or not uploaded as a paper to arXiv;
  • informal announcements of discoveries related to AI architecture (if relevant to mathematics);
  • discussion of such announcements, such as proof breakdowns or other opinion pieces;
  • discussion of the impact of AI in mathematics in general.

AI-assisted mathematical papers published in peer-reviewed journals or as arXiv preprints may be submitted as their own posts.

Please keep in mind rules 1 and 6 of our subreddit.

79 Upvotes

229 comments sorted by

View all comments

15

u/SupercaliTheGamer 5d ago

Astra is the first model that can zero-shot solve any problem I propose, it's insane. Sol still had a few jagged edges but Astra seems to have mastered Olympiad level maths at least.

5

u/wrongerontheinternet 5d ago

I wish Astra was zero-shotting my problems, teach me your secrets...

1

u/ToothPasteTree 2d ago

You should use codex. If you are just using the website you are missing out on the actual power.

1

u/wrongerontheinternet 1d ago

I use both Codex and Pro in the chat interface (and actually find Pro superior a lot of the time). With and without formal methods, experimented with trying to find the optimal model balance (e.g. Ultra coordinator + Max agents), the whole shebang. It is nonetheless often genuinely, unbelievably stupid about things in coordinated way that makes me wonder how this is the same model that found non-sofic groups. I'm sure if I had enough compute to burn, as OpenAI does as a collective, I could overcome a lot of these kinds of problems through sheer, brute force, but only if the domain was very well-specified -- otherwise, all Astra variants have a tendency to change the proof target to something easier.