LM/Performance: Difference between revisions
< LM
No edit summary |
|||
| Line 5: | Line 5: | ||
<syntaxhighlight lang="bash"> | <syntaxhighlight lang="bash"> | ||
llama-benchy --base-url https://models.nothanks.com/v1 --pp 512 --tg 128 --model my_model --api-key sk-12345678 | llama-benchy --base-url https://models.nothanks.com/v1 --pp 512 --tg 128 --model my_model --api-key sk-12345678 | ||
llama-benchy --base-url http://models.nothanks.com/v1 \ | |||
--model my_model --api-key sk-12345678 \ | |||
--pp 512 --tg 128 --concurrency 8 \ | |||
2>/dev/null | grep "^|" -B1 -A3 | |||
</syntaxhighlight> | </syntaxhighlight> | ||
Revision as of 09:12, 30 September 2026
Quick Test: llama-benchy
Use
llama-benchy --base-url https://models.nothanks.com/v1 --pp 512 --tg 128 --model my_model --api-key sk-12345678
llama-benchy --base-url http://models.nothanks.com/v1 \
--model my_model --api-key sk-12345678 \
--pp 512 --tg 128 --concurrency 8 \
2>/dev/null | grep "^|" -B1 -A3
Install
pipx install llama-benchy
Full Test: betterbench
Use
betterbench run --endpoint https://models.nothanks.com/v1 --model my_model --api-key sk-12345678
Install
git clone https://github.com/GGZ14/BetterBench.git
uv sync
.venv/bin/activate