How to Run Nemotron 3.5 Lightning Locally With llama.cpp

Nemotron 3.5 Lightning activates 3 billion parameters per token. Its Q4 GGUF is still nearly 24 GiB. If you want to run Nemotron 3.5 Lightning locally, that distinction determines whether you get a useful server or a very sophisticated out-of-memory…
