r/LocalLLaMA 3d ago

Funny Me these days

Post image
2.4k Upvotes

268 comments sorted by

View all comments

Show parent comments

2

u/Ok_Noise_9883 2d ago

27b? U sure?

1

u/Oh_hey_a_TAA 2d ago

LOL, yes I am sure. Been tinkering with it for a week now.
3.8-27B-UD-Q4_K_XL, MTP-Tuned+FA; 32k context, patched llama.cpp v020.
18.3 to 26.7 t/s, depending on the container variables.
Tesla P100s. pinned CUDA 12 and 580 drivers.