llama.cpp-dgx
llama.cpp fork optimized for NVIDIA DGX Spark / GB10 (Blackwell, SM 12.1) — TurboQuant weights + KV, NVFP4, DFlash MTP
// repository documentation
Was this content helpful?
(0 ratings)
