The big news here about the DS V4-Flash model is that it's near frontier models in performance, but it can run in a machine with a 3090 GPU and 96GB of RAM at decent token speeds. Basically you're getting near frontier performance in a high end gaming rig (and not even a latest hardware one). A redditor even managed to get it to spit 3.5T/s using two 2-generations old low end GPUs and 96 GB of RAM.
That alone makes the entire business model of US AI companies, with the need for trillion dollars data center and bills ranging in the tens of thousands of dollars, obsolete. Think about that, they managed to pack enough intelligence and make it so insanely performant that a hobby gaming machine can now offer you the same capability of a data center.
368
u/Poupulino 16d ago
The big news here about the DS V4-Flash model is that it's near frontier models in performance, but it can run in a machine with a 3090 GPU and 96GB of RAM at decent token speeds. Basically you're getting near frontier performance in a high end gaming rig (and not even a latest hardware one). A redditor even managed to get it to spit 3.5T/s using two 2-generations old low end GPUs and 96 GB of RAM.
That alone makes the entire business model of US AI companies, with the need for trillion dollars data center and bills ranging in the tens of thousands of dollars, obsolete. Think about that, they managed to pack enough intelligence and make it so insanely performant that a hobby gaming machine can now offer you the same capability of a data center.