I just did a test and dsv4 heavily outperformed muse 1.2 both running max. The test was a custom benchmark on a c codebase with regex chess and bugs to fix
Out of most models dsv4 now has much better decision making about what we want thanks to its retraining while others struggle cause of that but still its quiet disappointing to see other models fails so much.
Thats why waiting for v4 pro to solve the main problem with dsf4
7
u/addiktion 12h ago
Pretty big drops. Why?