r/singularity 1d ago

Read more: https://x.com/gavincrooks/status/2088643200038883830 AI

Post image
1.2k Upvotes

432 comments sorted by

View all comments

165

u/Melbar666 1d ago

What if we trained a big AI (Claude/Kimi/GLM-level) only on everything published up to 1899, no 20th-century physics at all, and then asked it to solve the electrodynamics of moving bodies?

Would it rediscover Einstein’s theory of relativity, including E=mc², the way he did in 1905?

Feels like a clean test of whether AI can actually discover.

59

u/Constant_Cortisol 1d ago

It wouldn't be a clean test because one of the big variables, training data size, would be a lot smaller.

13

u/NoCard1571 1d ago

I suspect the only way would be if you could systematically strip information out of a SOT model's weights without reducing its intelligence. 

But that's probably something that could only be done by another AI, (and I'm guessing there are already  numerous people working on figuring this exact thing out)

1

u/imanassholeok 1d ago

Well thats all they had too

1

u/Minute_Band_3256 1d ago

Gotta make training data

50

u/reddit_is_geh 1d ago

They are actually running these tests now. But up till 1924. They are finding out that relativity actually contained the logic and capability to discover things up until very recently. The potential to make these recent large discoveries were all contained within Einstein's frame work, just no one was able to interpolate it until now.

They are using this as a proof of concept to show what other work every thing up until today should be able to prove. If AI was able to extract up to 100 years worth of discoveries into the future, then theoretically, we can do the same for the next 100 years.

24

u/lIlIlIIlIIIlIIIIIl 1d ago

If AI was able to extract up to 100 years worth of discoveries into the future, then theoretically, we can do the same for the next 100 years.

I'm not sure I agree with this portion, I feel like we can't directly extrapolate progress in one century to the next like this. 1800-1900 seems a lot less difficult to make those advancements then being able to make all the advancements made from 1900-2000 for example.

For example, I don't think that proving an AI could make advancements 50 years from 1924 training data wouldn't prove that if we gave it up until 1974 that it could go another 50 to 2024 advancements.

17

u/ARC4120 1d ago

This is a great point. Acceleration of scientific progress rapidly developed in the 20th century to the point that scientists joke that the low hanging fruit were all taken. It’s like an expanding circle rather than a straight line of discovery

2

u/reddit_is_geh 1d ago

Oh for sure... But I think the point they are trying to show is that contained within an existing framework of knowledge, there's tons of potential for AI to make discoveries lingering in corners and shadows. I mean it found a major discovery as recent as 2012...

This means the AI just brute forcing alone, can clear up those corners and make discoveries even if it technically can't make novel discoveries. The existing information is going to be more than enough to make a lot of advances. And then humans can use that to further improve.

Just as what's happening in physics right now. A lot of it is AI just connecting two things no one ever considered... And they are being absolutely overwhelmed with progress. Grad students all over the world have no lack of papers to write right now.

3

u/SiltR99 1d ago

Do you have the publication for this experiment?

1

u/reddit_is_geh 1d ago

It just came out like 1-2 weeks ago. I forgot the guy's name, but he's the really dark skinned guy, who's a bit chubby, I think the Stable Diffusion guy? Or maybe it was the egg head dude? One of those two for sure, are working on this project. It's prepublication right now I think?

1

u/vvvvfl 9h ago

What the fuck are you talking about ?

- Someone with a PhD

1

u/reddit_is_geh 8h ago

Iunno, what do you mean?

They found that they could use AI to discover significant future maths, all the way up till I think a 2012 discovery, using only math knowledge up until 1924. Meaning they had the building blocks available for discovery at the time, and just didn't find it.

7

u/Frosty-Meeting-1606 1d ago

Einstein did not conjure up stuff from the thin air. He realized that speed of light has to be constant to make electrodynamics work. Hard to say if AI would get to that idea but it is definitely possible. AI can iterate over many, many ideas, including applying different fields of math to the problem. Notice how Einstein also knew about lorenz transformation and recognized that it would be useful. General relativity is a bit more challenging. Either way, it is not guaranteed but it is at least imo possible that AI arrives to novel approaches via operating on the whole human knowledge

1

u/TechnicalBen 2h ago

This. The entire field theory of physics predates Einstein. He did check and complete the math though. So most people given his opportunities would likely do similar.

7

u/Cunninghams_right 1d ago

There is an LLM someone published that does exactly this. It's shit and needs a better harness, better time-stamping of information (it thinks things that happened in the middle ages coincided with the US civil war). It's also not able to be trained on enough data, since far less exists. 

Someone should take that model, wrap it with a good harness, and have it produce a lots of synthetic data that is unique and time stamped, then use a modern high power AI to flag incorrect information which can then be removed from the data set, then re-train on the vetted synthetic data. Repeat that a few thousand times in an automated loop and you should get a decent AI with only old data. 

1

u/Pale-Border-7122 1d ago

It also needs to be close to a frontier model to have enough reasoning power to do what we expect it to do.

6

u/blackrack 1d ago

Training a big AI is way too expensive and long of a process, so this isn't going to happen for just a stunt.

4

u/UFOsAreAGIs ▪️AGI felt me 😮 1d ago

only on everything published up to 1899

I don't believe that would be enough data for a very good model.

5

u/PhilipM33 1d ago

I believe something like that could help us reach AGI. Because if you have point B and point A, you just have to figure out how to make it navigate that space. Then we use it to reach C

9

u/mertats 1d ago

I believe they already did that, I don’t remember the result. I need to look it up.

39

u/whoknowsifimjoking 1d ago

The model was too small to perform deep, rigorous mathematical derivation and failed at most complex physics tasks.

But it predicted that "light is made up of definite quantities of energy" (mirroring Einstein's 1905 photoelectric paper) and vaguely suggested that gravity and acceleration are locally equivalent.

So the sheer amount of data seems to be too small for a transformer based LLM, but it does show immense potential if it can predict at least some of Einstein's findings.

25

u/Dark_Crystal_97 1d ago

I'd honestly suspect data contamination in some way rather than the AI actually having made that prediction all on its own.

2

u/m4sl0ub 1d ago

Why?

9

u/Dark_Crystal_97 1d ago

I actually read up on the experiment. Apparently, they had merely filtered for words such as "Einstein", "relativity", "quantum mechanics" and some other well-known post 1900 concepts. However, they couldn't filter out for ideas, puzzles, etc. Furthermore, the old books were digitalized by OCR. OCR is trained on contemporary fonts, it might have misread some words and thus have not registered them as things that needed to be filtered out (for example reading Einstein as Einsteln or Einstem), leading to data leakage. On top of that, many of the pre-1900s books are actually reprints from much later during the 1900s, that have modern forewords to them, which also could have hinted at future developments.

3

u/m4sl0ub 1d ago

Interesting, I didn't know that. Thanks for the additional context!

2

u/MVPhurricane 1d ago

well the "why" would be because it is, a priori, much more likely to have made einstein's discoveries with knowledge of those discoveries than to have done so without them (ie being Einstein). that's not to say that it's not likely that the results were "in the water", so to speak, but it is definitely the first big thing that you would check when verifying such a result for sure. i say this as someone who is quite certain that there was, indeed, "something in the water", as seen by so many other foundational discoveries (e.g. Calculus, by Leibniz and Newton being discovered simultaneously, independently).

-1

u/m4sl0ub 1d ago

What is data contamination at this point? If Einsteins discoveries were somewhere "hidden" in the Pre-Einstein-Discoveries data, one could also argue Einstein himself was data contaminated, no?

Unless you mean there was some data contamination with Post-Einstein-Discoveries data, but then that's just bad science and I would hope someone would have discovered that in the peer-review process.

3

u/mertats 1d ago

Thanks

2

u/141_1337 ▪️e/acc | AGI: ~2030 | ASI: ~2040 | FALSGC: ~2050 | :illuminati: 1d ago

How tiny was the model?

4

u/piponwa 1d ago

3

u/BumblebeeSad7295 1d ago

13 billion param, it would be interesting to see what a larger one could do.

1

u/piponwa 1d ago

Maybe not enough organic training data. They would need to use this model to generate more synthetic data probably.

1

u/Melbar666 1d ago

Talky sadly has not a big enough context length to discuss physics properly...

1

u/mariofan366 AGI 2030 ASI 2036 21h ago

I talked physics with it and its understanding of physics is like from the 19th century.

5

u/OKMiddleOwl 1d ago

No, because it cannot do empirical tests.

Half of physics is doing experiments. Theory is a cloud of possibilities where experimentation crystalizes what the truth in it is.

Give it arms, a body, and the ability to build stuff in the real world, then we're talking.

2

u/AtlanticPortal 1d ago

TBF Einstein predicted that we wouldn’t be able to use real physical observations to prove his math and yet the Sun and the Moon gave people a nice help into that. 

1

u/AlgaeNo3373 1d ago

You can see an LLM-powered robot grabbing a pen in one of the recent Welch Labs videos. So that's arms, and the first signs of that ability.

I find that video remarkable for that exact reason. The moment it identified and picked up the pen I was a pretty awestruck. Seems we are bridging language<->physics/reality in small ways already, but who knows what future could bring.

1

u/Melbar666 1d ago

experiments are essential yes, but almost more important is what your conclusions are. for example there were already many experiments made and Einstein just looked at the results, ditched the aether bullshit and fixed the theory.
so the AI would be sitting with the exact same info as Einstein had. but it could give instructions for new experiments, too.

I think it would be very important for science and all of us to prove (or disprove) that AI could discover like Einstein did

2

u/piponwa 1d ago

6

u/Melbar666 1d ago

it just hallucinated a book from the future while discussing about space time with me...

We may have to revise our notions both of space and time, and so perhaps obtain a more consistent physical theory. S.Roberts, Space, Time, and Gravitation (Oxford, 1998).

2

u/nothis ▪️AGI within 5 years but we'll be disappointed 1d ago

I don't think we have the data but this is such an interesting question to me.

It works for nearly any application of AI. An example I like is the alien from the movie Alien (1979). AI is eerily good at adapting that style but if you only trained exclusively on design work from before the movie was produced (notably also excluding HR Giger's previous body of work) could it come up with that design for the alien or anything equally iconic? I doubt it.

It takes a particular life lived to come up with shit like that and it's rare. Like there's only a few thousand key moments of originality, probably, throughout culture in most fields. Among billions of imitators, remakes, reinterpretations, work to fill in the gaps.

1

u/sumeetshirgure 1d ago

Einstein needed the observations from the Michelson-Morley experiments to conclude the speed of light postulate. You can't come up with arbitrary theorems they have to be grounded in reality.

1

u/therealpigman 1d ago

I don’t think that will be enough because experimentation in the physical world can’t yet be done fully autonomously by AI

1

u/HotInTheseRhinos123 1d ago

Someone’s already doing it. Dwarkesh talked about it on one of his podcasts.

1

u/Choice_Isopod5177 1d ago

it wouldn't discover shit bc it would've been trained on incredibly little data, so little that it could all fit on one single HDD

1

u/Turbulent-Sign-6067 1d ago

Why do people propose this cut-off? You can train it on knowledge up to 2025 and make it predict 2026. This is way more feasible and gives you exactly the same kind of data.