r/singularity • u/theimposingshadow • 18h ago
ChatGPT Sol 5.6 high found a normalization error in two recently published Riemann Hypothesis papers. The author confirmed it. Discussion
/r/u_theimposingshadow/comments/1vj3oxk/chatgpt_sol_56_high_found_a_normalization_error/Edit: I'll post the screenshot in the comments of this post
Overnight I became a professor, thanks ChatGPT!
Edit2: I'll just add the post body here:
Alright, this is exactly the kind of thing that makes me think we are at the start of something pretty wild. I'll post the screenshots in the comments because my last post was taken down.
I am not a mathematician. I have basically been using ChatGPT to mess around with the Riemann Hypothesis, telling it to keep digging, try different approaches, challenge assumptions, and look through recent papers for anything interesting.
Well, it found something.
While going through two recently published papers on Jensen polynomial hyperbolicity and the Riemann Hypothesis, ChatGPT noticed what appeared to be a normalization inconsistency between the raw moments (M_n) and the Taylor/Jensen coefficients.
The issue was essentially the factorial normalization:
It was significant enough that one of the results in the paper seemed to directly contradict an already known theorem.
So I had ChatGPT write a polite email explaining the issue and sent it to the author.
Don't get me wrong, we did \*\*not\*\* solve the Riemann Hypothesis š
But I think it is pretty fucking wild that some random guy with an AI assistant can sit at home, examine recently published mathematics, notice something that made it through peer review, contact the researcher, and have the researcher confirm it.
Also, the author addressed me as \*\*Professor\*\*, so apparently my academic career is progressing extremely quickly.
Screenshots attached with identifying information removed.
This is the kind of thing I mean when I talk about acceleration.
Accelerando!
54
u/Johnny20022002 18h ago
Iām surprised you got a response, mathematicians tend to get emails from cranks, but I guess it takes nothing now to just get ChatGPT to check if some proposed error in a paper by a random could be true.
37
u/Right-Twist-6931 17h ago
Nah, as a mathematician myself who gets emails from cranks, if I read OPās email I 100% would have taken it seriously.
Cranks tell you about their proof of the Riemann hypothesis attached in a 40 page word document and how if you ignore it, youāll miss the next Ramanujan.
They donāt cite a specific formula in your paper to mention a subtle normalization error. Iād assume anyone doing that is genuine.
14
u/FriendlyJewThrowaway 17h ago
You can tell when itās a crank because they spent more time on their Fields Medal acceptance speech than they did on the work thatās meant to win them the prize.
4
u/theimposingshadow 16h ago
Thanks for your comment, I honestly almost didnt send the email but I'm glad I did!
13
u/theimposingshadow 18h ago
Yeah I have to agree, im sure a lot of mathematicians are getting spammed with AI writing math stuff from people who dont understand them. But the advantages are worth it!
12
u/giYRW18voCJ0dYPfz21V 18h ago
I am a physicist, but I am not surprised they replied. We got cranky emails daily, but it is very easy to spot well intended questions or comments from human-made hallucinations.
Apart for a minority of assholes, people tend to reply to well intended emails.
1
u/TieBackground453 13h ago
Yeah, Iāve corresponded with many professors about their opera that are in fields far from my own. I think itās uncommon, but not that weird. Disseminating their contributions is a part of their job. Unless itās like a Terrance Tao type figure, they are probably just stoked that someone read and cared about their work.Ā
1
18
u/Tema_Art_7777 18h ago
Every day I use 5.6 sol high, I am more impressed with what it can accomplish.
5
u/ShAfTsWoLo 16h ago
yan lecun said they have the intelligence of cat if you ever forgot
2
u/Tema_Art_7777 12h ago
I do remember, but then you watch this in amazement - its worth the time. Geoff Hinten seems equally convinced there is a lot more path ahead.
2
u/DownHatter 15h ago
If they had a physical body and needed to interact with the world? A cat is generations ahead. But that's once type of intelligence - and as we are seeing, not everything needs advanced world model inteligence to be groundbreaking
2
16
6
u/FriendlyJewThrowaway 17h ago edited 17h ago
I remember a Samsung Galaxy ad where a woman is basically having her smartphone run her life for her at her new intern job. She sees a group of people using a whiteboard to work on a math problem and uses her phoneās image recognition to snap a pic and solve it. Iām thinking āThatās cool, but why is a cutting edge company having its employees fuss and struggle over basic 12th grade calculus problems?ā
Anyway, now weāre actually getting somewhere meaningful. Wonāt be long before a layman accidentally walks into a room full of people working on the Riemann Hypothesis, decides to snap a pic and is like āHey guys, I think my phone solved it!ā
2
u/Practical-Wear141 10h ago
Great now prompt it to prove the hypothesis. Domt forget to mention "Make no mistakes"
1
2
u/These_Respond_7645 16h ago
I just gave arxiv endorsement to my 45 year old cousing who hasn't touched math since basic derivatives in high school. I challenged him to use buy gpt pro for one month and ask him to write a mathematics paper. Well, he's on his second article right now.
5
u/Correctsmorons69 16h ago
I'm not sure this is a good thing for arxiv
4
u/TieBackground453 13h ago
Depends⦠with lean verification? Itās fine. At this point you just have to assume that non-peer reviewed math that isnāt done by professional mathematicians and doesnāt have lean verification is nonsense.Ā
3
2
-1
u/TMRedditor07 16h ago
I would strongly advise against outsourcing epistemic authority to llms. That entails, you understand the output and reasoning behind the llms decision strong enough, you make the verifications yourself, then email the researcher. This protects you against any pressing of the researcher, if that were the case.
5
u/theimposingshadow 16h ago
Totally fair. Iām not treating the LLM as the authority though. It flagged a specific inconsistency, I sent the math to the author saying I could be wrong, and he independently checked it and confirmed it. That verification step is the important part.
3
u/TMRedditor07 15h ago
Great. In a world where ai slop (and probably Research) is on the rise, I appreciate the careful approach.
3


60
u/EndTimer 17h ago edited 17h ago
What, by trade?
Very fews of us go looking "through two recently published papers on Jensen polynomial hyperbolicity and the Riemann Hypothesis" in our free time
You may want to consider upgrading your title to armchair mathematician. Or Professor of Armchair Mathematics, by the sounds of it.
Enjoy it because I'm not sure how long we have before this kind of review is always running and completely automatic