hf_text-generation-inference/benchmark
OlivierDehaene 50b495f3d8
feat: add more latency metrics in forward (#1346)
2023-12-14 15:59:38 +01:00
..
src feat: add more latency metrics in forward (#1346) 2023-12-14 15:59:38 +01:00
Cargo.toml Preping 1.1.0 (#1066) 2023-09-27 10:40:18 +02:00
README.md feat(benchmark): tui based benchmarking tool (#149) 2023-03-30 15:26:27 +02:00

README.md

Text Generation Inference benchmarking tool

benchmark

A lightweight benchmarking tool based inspired by oha and powered by tui.

Install

make install-benchmark

Run

First, start text-generation-inference:

text-generation-launcher --model-id bigscience/bloom-560m

Then run the benchmarking tool:

text-generation-benchmark --tokenizer-name bigscience/bloom-560m