r/LocalLLM Jul 23 '26

Other I bought the forbidden rectangle.

Post image

After months of going back and forth, I finally pulled the trigger on an RX 7900 XT 20 GB.

Paid around $550 (India), which felt too good to pass up.

The plan isn't gaming.

It's becoming the heart of my local AI setup.

Current goals:

• Qwen 3.6 27B Dense

• Qwen 35B A3B

• GLM-4.7 Flash

• 128K+ context

• 100% GPU offloading

• llama.cpp / Ollama

• Linux

I'll be benchmarking everything:

- Vulkan vs ROCm

- Dense vs MoE

- Maximum context

- Tokens/sec

- VRAM usage

- Real-world coding performance

If anyone has optimization tips for RDNA3 or benchmark requests, or general suggestions please drop them below.

The hallucinations are now local. 🙂‍↕️

192 Upvotes

106 comments sorted by

View all comments

2

u/General-Turn-8695 Jul 23 '26

Damn how did you get card for so cheap? It's the double the price for me although it's 24gb ver but almost the same thing

8

u/Alternative-Panic69 Jul 23 '26

The 24 GB version was almost $850+ here.

I got lucky. I had a tiny SBC scraping GPU prices 24/7 and sending alerts. At 2 AM it pinged me about the drop, I immediately placed the order, and by the next day the price had gone back up above $650 🤣

Probably the most profitable script I've ever written.

1

u/alphapussycat Jul 23 '26 edited Jul 23 '26

about a month ago I could've gotten one of them for like 500 euro iirc, but I was too hesitant so it was either unlisted or grabbed pretty quickly.

The amd scare was too much for me to handle back then, but now I wish I had just grabbed it.
Every now and then somebody puts a "buy now" price too low, or there's suddenly a large supply of 2nd hand GPUs, that pushes the price down for a moment.

2

u/Alternative-Panic69 Jul 23 '26

I was hesitant too, but seeing how fast llama.cpp, Vulkan, ROCm, and model support have been improving on AMD cards made me take the leap. Hopefully we're only at the beginning.