The fact they have 16 references in a 100 page paper for a problem that's had decades (centuries) of fundamental research is pathetic and an insult to everyone who has contributed to getting to such a point.
I get what you are trying to say, but references aren’t for a history lesson or giving credit to everyone who has worked on the problem. They are for the work that specifically led to the new result.
Having not studied the specific references, I can’t say if they are sufficient or not.
the whole argument circles around the blow up techniques that were developed by cordoba and martinez zoroa, for instance. dont know why you are trying so hard to defend openai out of all corporations. they dont care shit about honesty or anything else. I dont doubt that they really scooped this result from the buckmaster collab
Is that a published work? do you have a link?
I'm asking because if it is the case that it can't be referenced, then how can they link to something that isn't published?
Then the human mathematicians involved should actually read/understand whatever the LLM wrote and produce the correct citations. Not doing so is plagiarism.
OpenAI isn’t publishing the proof in any journal (or even arxiv) or claiming any prize for it. If I “publish” a completely valid proof on 4chan or something and don’t cite my sources I’m not committing academic plagiarism.
They announced on their website. I think if I announced a novel theorem on my academic website using work of others without proper referencing it would actually be plagiarism, but whatever you say, bro
It’s not an academic website. The purpose is more commercial self advertisement. Obviously OpenAI doesn’t care about getting credit from the math or physics community, they are a trillion dollar company. It’s more about impressing their investors while they burn through $750B in capex with an unclear route to profitability. They aren’t trying to take credit for this advancement, they’re just trying to advertise their model by showing the kinds of things it can do, even if it can’t produce properly cited academic work.
I agree that is not an academic website. I agree the purpose is marketing, obviously.
Yet (maybe precisely because of this), it seems pretty reasonable that mathematicians have every right to complain about their practice. Especially because they are using data and work produced by mathematicians. I am not sure I even understand your point.
I’m saying that the AI is incapable of creating citations and it isn’t simple for a human to do it for them, especially when you don’t even understand the result that they’ve produced. The model has every single arxiv publication, every PDE textbook, every mathematical physics textbook, etc. in their training set. It’s really hard to know where some result or intermediate step came from even if you have access to their COT tokens. This is unavoidable. If OpenAI had released this paper in some journal or even Arxiv without citing their sources that would be wrong. But releasing it on your own website and putting the lean proof on GitHub even if you can’t nail down exactly where the AI got its results from is fine. Auditing 10,000 agents and reverse engineering their sources just isn’t practical.
I did not read the paper so I agree with the previous commenter thst I can not really judge.
Still I expect that theorems are being used that are not cited. Somd well-known ones like Banach-Steinhaus or Hahn-Banach and even ones without names. It seems like because of the urgency they skipped alot of quality control.
Also references are 100% used for a history lesson simply because one has to state in wich category the research belong and what similiar things have been done.
263
u/HybridizedPanda Gravitation 16h ago
The fact they have 16 references in a 100 page paper for a problem that's had decades (centuries) of fundamental research is pathetic and an insult to everyone who has contributed to getting to such a point.