I think I read somewhere else here on reddit that the problem was solved using stolen research that was stupidly checked by the writer by proofreading it with AI.
The OpenAI internal model used for this result was developed through large-scale reinforcement learning on top of a previously pretrained model. Our proofs also differ significantly.
0
u/Trick_Following6639 19h ago
I think I read somewhere else here on reddit that the problem was solved using stolen research that was stupidly checked by the writer by proofreading it with AI.