ray_vllm_inference

(★ 79)

A simple service that integrates vLLM with Ray Serve for fast and scalable LLM serving.

ray_vllm_inference 최신버젼 다운로드

최종 버전 다운로드 (.zip)
// repository documentation