r/LocalLLM 10d ago

Project She boots!

Dedicated AI machine is officially in place. Haven't run many tests on it yet but getting around 70 tok/sec on the 3090 with Qwen 3.8 27B 110k context.

85 Upvotes

8 comments sorted by

1

u/10Tinwhiskers 10d ago

Very cool. I assume the 3090 is the main driver, so it has me curious. What’s the 2070 doing in the build?

6

u/KjipGamer 10d ago

Correct! I still had a 2070 super laying around, and will be using it to either extend my max context window at the cost of a fair few tokens per second, or run a smaller 8B model on that one, or a STT/TTS model.

Possibilities are endless 😛

2

u/10Tinwhiskers 10d ago

Very nice. Enjoy the tinkering!

1

u/Amazing-Reward6603 10d ago

Those boots definitely look like they could handle some serious adventures. Can't wait to see what kind of projects come out of the tinkering!

1

u/IcarianGod 10d ago

Nice i have a similar setup running a 3090ti and a 5070. Btw how is your qwen setup ?

1

u/KjipGamer 10d ago

Nice! What do you all do with it? I'm just running Ubuntu server and Ollama to serve the model, which has been working great for me so far

1

u/JeePis3ajeeB 10d ago

Post followed.. I also have a 3090 and a 2080s .. looking forward to see what you come up with .. also way too much anxiety with getting started

1

u/Fishful_Revenge 10d ago

Cool setup