hf_text-generation-inference/server/text_generation_server/utils/gptq
OlivierDehaene 9946165ee0
chore: add pre-commit (#1569)
2024-02-16 11:58:58 +01:00
..
custom_autotune.py fix: fix quant linear autotune 2023-12-14 16:45:47 +01:00
exllama.py fix: fix gpt-q with groupsize = -1 (#1358) 2023-12-18 16:07:05 +01:00
exllamav2.py GPTQ support on ROCm (#1489) 2024-01-26 16:27:44 +01:00
quant_linear.py chore: add pre-commit (#1569) 2024-02-16 11:58:58 +01:00
quantize.py feat: format code (#1070) 2023-09-27 12:22:09 +02:00