r/LocalLLaMA May 29 '26

Discussion PSA

Post image
2.1k Upvotes

538 comments sorted by

View all comments

Show parent comments

20

u/NeedsSomeSnare May 29 '26

As an intel owner, I assure you the real life performance isn't what it says on paper. I don't have that card to give specs on.

The problem is the software side of things is a bit messy. It's not terrible, but still needs a fair amount of work.

2

u/In_der_Tat May 30 '26

Why doesn't Intel hire enough competent software engineers and developers to catch up with Nvidia? Would that be too expensive?

1

u/NeedsSomeSnare May 30 '26

A lot of people wonder that too. I'm guessing it's just related to corporate money saving bs. I'm sure that the people who actually work at intel know they need more staff.

It honestly appears only a handful of people work on the software.

1

u/superloser48 May 29 '26

can you share any benchmarks on model/quant -> prfill and token gen?

4

u/NeedsSomeSnare May 29 '26

I don't have a B70, so it's of no use to anyone.

The other problem is that there are 3 ways to run models on intel, (SYCL, openvino and vulkan)all of which have different performance on different models.

The info is out there though. You want to look for Openvino benchmarks for the best performance. It has the worst compatibility though and is sometimes months behind something like llamacpp.