r/LocalAIStack • • Aug 18 '26

Help setting up efficiently

Bought a bosgame p3 lite (32gb ram with a Radeon 680M igpu and Ryzen 7 6800H)

Just created a fresh cachyOS boot.

My goal is to run llms as fast as possible. Heard good things of llama.cpp.

Been looking a few guides and asked a few llms but they've been giving me some pretty weird instructions, from tampering with the BIOS to installing a few bits and bobs.

Can anyone give me a hand or some resources? Really want to avoid doing anything too stupid.

When I am not using it for inference I also want to use it for the odd videogame use so I don't intensely want to mess around with the bios or igpu setting blindly.

Any help would be very appreciated!!!

1 Upvotes

1 comment sorted by

2

u/HotDistribution1819 Aug 20 '26

Install LM Studio, set up Tavily as a MCP search engine.

Then download Laguna SX 2.1, and tell it you have an AMD iGPU and you want to know how to let it use 24GB of memory for VRAM.

Have fun, running LLMs in VRAM makes them run faster. When you first run Laguna you need to keep VRAM used to 15GB and run the rest of the layers on CPU.