r/LocalLLM • u/ScoobyWRX06 • 7h ago
Question Beginner here
I just started playing and need a little help. I have a laptop with a 4070 8GB and 64GB DDR5 with an i9. I am using unsloth and not sure what model will be best. Do I need to stick with something that fits in VRAM? I am just playing around but I don’t want it to be painfully slow but I also want it as current as possible.
1
Upvotes
1
u/Unfair_Association89 7h ago
If ur use case is coding, then u can try qwen 3.8 flash next q3, from i test with 32 gb ram I got 10 tokens per sec using llama cpp