hf_text-generation-inference/benchmark
Daniël de Kok 7735b385dc Prefix caching WIP 2024-08-09 14:52:59 +00:00
..
src Prefix caching WIP 2024-08-09 14:52:59 +00:00
Cargo.toml Rebase TRT-llm (#2331) 2024-07-31 10:33:10 +02:00
README.md chore: add pre-commit (#1569) 2024-02-16 11:58:58 +01:00

README.md

Text Generation Inference benchmarking tool

benchmark

A lightweight benchmarking tool based inspired by oha and powered by tui.

Install

make install-benchmark

Run

First, start text-generation-inference:

text-generation-launcher --model-id bigscience/bloom-560m

Then run the benchmarking tool:

text-generation-benchmark --tokenizer-name bigscience/bloom-560m