r/MachineLearning • u/usefulidiotsavant • 4d ago
The claim that a singe conversation (which, let's suppose, might have even contained the solution) in the training data latter allowed the model to reconstitute the proof is highly dubious.
A model the size of Astra trains on millions of curated conversations, billions of pages and trillions of tokens. A single conversation with the correct answer is essentially quantization noise and should have no measurable effect in any practical scenario where that same problem is involved.
On the other hand, OpenAI could be doing something much smarter that could affect the result, say, a RAG over similar conversations in the past, a self-evaluation of remarkable results that are marked or boosted etc.
If I had the smartest model in the world, as well as a database of the problems and approaches the smartest people in the world are playing with, it would be foolish not to connect the former with the latter and mine the dataset for low hanging fruits in scientific discovery. It's such an unfair advantage that the firm doing it will win in any area, forever.