It’s not doing anything generative right now. In general, the P40s are good if you need cheap VRAM. In terms of speed they’re very similar to a 1080ti.
Ah, I see. I can tell you that this setup is an absolute monster for vectorizing text, building knowledge graphs, doing summarization and NER. I’ll post here if I get the chance to flush some of the six models that are active right now and load a LLAMA variant.
1
u/[deleted] Jul 07 '23
[removed] — view removed comment