llama-box
LM inference server implementation based on *.cpp.
파일 탐색기
최종 버전 다운로드 (.zip)- ci.yml
- prune.yml
- sync.yml
- ggml-cpu.patch
- batch.patch
- clip.patch
- common.patch
- context.patch
- dynamic_link.patch
- embedding.patch
- ggml-cpu.patch
- ggml-cuda.patch
- ggml-hip.patch
- ggml-metal.patch
- ggml-rpc.patch
- grammer.patch
- log.patch
- max_devices.patch
- model.patch
- model_py.patch
- mrope.patch
- ngram_cache.patch
- progress_callback.patch
- pure_cpu.patch
- sampling.patch
- seed.patch
- template.patch
- tool_calling.patch
- vendor_httplib.patch
- vocab.patch
- dynamic_link.patch
- log.patch
- progress_callback.patch
- util.patch
- clean.sh
- pull.sh
- sync.sh
- build-patch-reset.cmake
- build-patch.cmake
- build-windows-arm64.cmake
- configure-patch.cmake
- gen-version-cpp.cmake
- batch_chat.sh
- chat.sh
- chat_tool_current_date_time.sh
- chat_tool_get_temperature.sh
- chat_tool_get_weather.sh
- chat_tool_square_of_number.sh
- chat_tool_square_root_of_number.sh
- chat_tool_where_am_i.sh
- image_edit.sh
- image_generate.sh
- image_view.scpt
- .clang-format
- CMakeLists.txt
- engine.cpp
- engine_param.hpp
- httpserver.hpp
- rpcserver.hpp
- version.cpp.in
- z_multimodal.hpp
- z_stablediffusion.hpp
- z_utils.hpp
- .gitattributes
- .gitignore
- .gitmodules
- CMakeLists.txt
- concurrentqueue
- LICENSE
- llama.cpp
- readerwriterqueue
- README.md
- stable-diffusion.cpp
// repository documentation
Was this content helpful?
(0 ratings)
