hf_text-generation-inference/server
OlivierDehaene 9946165ee0
chore: add pre-commit (#1569)
2024-02-16 11:58:58 +01:00
..
custom_kernels chore: add pre-commit (#1569) 2024-02-16 11:58:58 +01:00
exllama_kernels chore: add pre-commit (#1569) 2024-02-16 11:58:58 +01:00
exllamav2_kernels chore: add pre-commit (#1569) 2024-02-16 11:58:58 +01:00
tests feat(server): add frequency penalty (#1541) 2024-02-08 18:41:25 +01:00
text_generation_server chore: add pre-commit (#1569) 2024-02-16 11:58:58 +01:00
.gitignore
Makefile
Makefile-awq chore: add pre-commit (#1569) 2024-02-16 11:58:58 +01:00
Makefile-eetq
Makefile-flash-att chore: add pre-commit (#1569) 2024-02-16 11:58:58 +01:00
Makefile-flash-att-v2
Makefile-selective-scan chore: add pre-commit (#1569) 2024-02-16 11:58:58 +01:00
Makefile-vllm
README.md chore: add pre-commit (#1569) 2024-02-16 11:58:58 +01:00
poetry.lock Update to peft 0.8.2 (#1537) 2024-02-08 12:44:04 +01:00
pyproject.toml Outlines guided generation (#1539) 2024-02-15 10:28:10 +01:00
requirements_common.txt
requirements_cuda.txt
requirements_rocm.txt

README.md

Text Generation Inference Python gRPC Server

A Python gRPC server for Text Generation Inference

Install

make install

Run

make run-dev