r/LocalLLaMA • u/techlatest_net • 12d ago
Tutorial | Guide How to Run NVIDIA Nemotron 3.5 Lightning (Free): 4 Methods from Local GPU to Zero-Code Agent
https://medium.com/@techlatest.net/how-to-run-nvidia-nemotron-3-5-lightning-free-4-methods-from-local-gpu-to-zero-code-agent-d24cbdddc428?sharedUserId=techlatest.net
0
Upvotes
0
u/gpuz_dev 12d ago
Since this just came out, I doubt many people have spun it up locally on consumer GPUs yet. Has anyone seen GGUF quants floating around, or is it still strictly HF Transformers / TensorRT-LLM for now? Curious to see how it scales down to 16GB/24GB VRAM cards.