llama.cpp-dgx

(★ 31)

llama.cpp fork optimized for NVIDIA DGX Spark / GB10 (Blackwell, SM 12.1) — TurboQuant weights + KV, NVFP4, DFlash MTP

llama.cpp-dgx 최신버젼 다운로드

최종 버전 다운로드 (.zip)
// repository documentation