C-ray — X-ray your containers. See beneath the runtime.
Explore Similar Repositories
lazy-moe:The GPU-free LLM inference engine. Combines lazy expert loading + TurboQuant KV compression to run models that shouldn't fit on your hardware. Built from scratch, fully local, zero cloud.