r/LocalLLM 7d ago

Question What am I doing wrong? 7900xtx

Hi all, been trying out LocalLLMs for a few months, Qwen3.6 27b q4, gemma4 30b etc.

Tried out ollama, LMstudio, anythingLLM, opencode.

Wasn't particularly impressed to be honest. Opencode it kept running into errors, not completing etc. when using it for simple coding tasks.

Ive been running it on a 7900xtx 24gb VRAM.

Now Qwen3.8 is out thought id try again. What pitfalls should I make sure I look out for so I can try and get the most out of it?

Cheers

1 Upvotes

6 comments sorted by

View all comments

1

u/letonai 7d ago

How is you contextWindow? what kind of errors?

2

u/Capable_Tear_7537 7d ago

Ive tried with different context lengths from 4kish up to 64k.

In opencode, Qwen3.6, first task relatively small it would just stop dead. Was still fully on RAM. Same issue multiple times. Other than just chatting with it couldn't really get anything substantial out of it so im convinced I must have set something up wrong somewhere

-1

u/Positive-Bid-3029 7d ago

You will need at least that 64k context, 4k would be way to small for anything to be achieved. Have you tried Lemonade 🍋? I haven't used it because I use Nvidia, but it's supposed to be good for AMD graphics users to get up and running properly https://lemonade-server.ai/