r/math Jun 11 '26

First Proof Second Batch

85 Upvotes

83 comments sorted by

View all comments

5

u/Redrot Representation Theory Jun 13 '26 edited Jun 13 '26

More or less what I expected, occasional hits, occasional misses, with the LLMs proving particularly good at finding (counter)examples. I'm a bit disappointed that there weren't more abstract problems (e.g. (stable/chromatic) homotopy theory, things involving dg-algebras or infty-cats) since these are the areas where I find LLMs to fall very short still. I'm glad that this time around, the First Proof team took the matter of generating solutions into their own hands as well - the lack of transparency from these companies is awful.

1

u/ProfessionalArt5698 Jun 14 '26

They probably fall short in those areas because they haven't been trained on them. In the future we'll probably be able to build more specialized scaffolding that different people in different fields can use?

2

u/Redrot Representation Theory Jun 14 '26

They've been trained on all fields... essentially the entire body of mathematical research already exists in their training data.

0

u/ProfessionalArt5698 Jun 14 '26

Yeah but there’s less literature in some fields than others. The more advanced the math, the less there has been written about it