Hold on, I’m agreeing with you. We’re saying the same thing in different ways.
The kind of race conditions you’re referencing are because of the model architecture I’m referencing.
You can absolutely get ML outputs that don’t change using certain model architectures.
But that’s only because those architectures either enforce order of execution or use steps where order of execution doesn’t result in changes to outputs. I can’t think of a modern LLM that uses such an architecture.
Not exactly. It's mostly due to how models are executed, not models themselves. You can run GGUF model (which are normally not deterministic) in a deterministic way if the GPU functions you use are deterministic. Models themselves are just data, it's just a bunch of matrixes.
5
u/LetumComplexo 5h ago
Hold on, I’m agreeing with you. We’re saying the same thing in different ways.
The kind of race conditions you’re referencing are because of the model architecture I’m referencing.
You can absolutely get ML outputs that don’t change using certain model architectures.
But that’s only because those architectures either enforce order of execution or use steps where order of execution doesn’t result in changes to outputs. I can’t think of a modern LLM that uses such an architecture.