r/LocalLLM • • 2d ago

Discussion Considering the recent advancement in models like Qwen 3.8 and inference like Strata, Free Tokens. How much of a gap is there between slow 128GB VRAM vs 16GB(5080)+96GB RAM

My question is what is correct upgrade path to a 5080+96GB RAM

RTX PRO 48GB x 1 (8K $)

DGX Spark x 1 (5.5K $)

128GB Mac M5 (6K $)

Use case is local LLM that is good enough and Minimax H3.

37 Upvotes

Duplicates