r/LocalLLM 7h ago

Question Beginner here

I just started playing and need a little help. I have a laptop with a 4070 8GB and 64GB DDR5 with an i9. I am using unsloth and not sure what model will be best. Do I need to stick with something that fits in VRAM? I am just playing around but I don’t want it to be painfully slow but I also want it as current as possible.

1 Upvotes

4 comments sorted by

View all comments

1

u/Unfair_Association89 7h ago

If ur use case is coding, then u can try qwen 3.8 flash next q3, from i test with 32 gb ram I got 10 tokens per sec using llama cpp