cpubrrr

(★ 56)

Frontier-class LLM inference on a laptop CPU — gpt-oss:20b at ~110 tok/s on Apple M4 Max, 7.5x llama.cpp, no GPU. From-scratch NEON/SME kernels in Rust.

cpubrrr 최신버젼 다운로드

최종 버전 다운로드 (.zip)
// repository documentation