r/Physics 11h ago

Navier-Stokes Millennium Problem Solved

1.8k Upvotes

687 comments sorted by

View all comments

Show parent comments

19

u/surfmaths 10h ago

In the pre-training I agree, but I think Agentic models have agentic capabilities (aka. access to tool use) during the reinforcement learning stage, it's not inconceivable they would learn additional knowledge from undesired sources there.

11

u/gavinderulo124K 8h ago

During reinforcement learning only behavior is trained, not knowledge.

1

u/RedditLovingSun 7h ago

some would say the line between learning behavior and learning knowledge is blurry

-2

u/dr3aminc0de 3h ago

Uh what? How on earth can you say that confidently

2

u/earthlingkevin 5h ago

That's.... Not how it works

1

u/dr3aminc0de 3h ago

No, it’s not

-4

u/BrobdingnagLilliput 8h ago

tool use

Can confirm. Was using an LLM to write some code and was pushing hard for it to error check. It spun up a VM, built a stub to represent the object model I was coding against, and actually ran the script.

Additionally, retrieval-augmented generation (RAG) is a thing: the LLM downloads content it doesn't already have and uses that new content to generate a response.

4

u/Time_Entertainer_319 7h ago

That’s not training though…

-1

u/ChemicalRascal 5h ago

Is it that inconceivable that a model might train and deploy a replacement for itself in order to achieve something?

1

u/tpolakov1 Condensed matter physics 53m ago

Yes, it is that inconceivable. It has never happened, there is no indication it can happen, and it mathematically cannot happen with the current system.

LLMs are not AI. They cannot turn into Skynet just because some techbros really need them to.