llm-benchmark (ollama-benchmark)

LLM Benchmark for Throughput via Ollama (Local LLMs)

Installation Steps

pip install llm-benchmark

Usage for general users directly

llm_benchmark run

Installation and Usage in Video format

It's tested on Python 3.9 and above.

ollama installation with the following models installed

7B model can be run on machines with 8GB of RAM

13B model can be run on machines with 16GB of RAM

Usage explaination

On Windows, Linux, and macOS, it will detect memory RAM size to first download required LLM models.

When memory RAM size is greater than or equal to 4GB, but less than 7GB, it will check if gemma:2b exist. The program implicitly pull the model.

ollama pull gemma:2b

When memory RAM size is greater than 7GB, but less than 15GB, it will check if these models exist. The program implicitly pull these models

ollama pull gemma:2b
ollama pull gemma:7b
ollama pull mistral:7b
ollama pull llama2:7b
ollama pull llava:7b

When memory RAM siz is greater than 15GB, it will check if these models exist. The program implicitly pull these models

ollama pull gemma:2b
ollama pull gemma:7b
ollama pull mistral:7b
ollama pull llama2:7b
ollama pull llama2:13b
ollama pull llava:7b
ollama pull llava:13b

Python Poetry manually(advanced) installation

https://python-poetry.org/docs/#installing-manually

For developers to develop new features on Windows Powershell or on Ubuntu Linux or macOS

python3 -m venv .venv
. ./.venv/bin/activate
pip install -U pip setuptools
pip install poetry

Usage in Python virtual environment

poetry shell
poetry install
llm_benchmark hello jason

Example #1 send systeminfo and benchmark results to a remote server

llm_benchmark run

Example #2 Do not send systeminfo and benchmark results to a remote server

llm_benchmark run --no-sendinfo

Example #3 Benchmark run on explicitly given the path to the ollama executable (When you built your own developer version of ollama)

llm_benchmark run --ollamabin=~/code/ollama/ollama

Reference

Ollama

Name		Name	Last commit message	Last commit date
Latest commit History 55 Commits
.github/workflows		.github/workflows
llm_benchmark		llm_benchmark
tests		tests
.gitignore		.gitignore
LICENSE		LICENSE
README.md		README.md
llm-benchmark.gif		llm-benchmark.gif
pyproject.toml		pyproject.toml
requirements.txt		requirements.txt
setup.py		setup.py

Provide feedback

Saved searches

Use saved searches to filter your results more quickly

.github/workflows

.github/workflows

llm_benchmark

llm_benchmark

tests

tests

.gitignore

.gitignore

LICENSE

LICENSE

README.md

README.md

llm-benchmark.gif

llm-benchmark.gif

pyproject.toml

pyproject.toml

requirements.txt

requirements.txt

setup.py

setup.py

Repository files navigation

llm-benchmark (ollama-benchmark)

Installation Steps

Usage for general users directly

Installation and Usage in Video format

ollama installation with the following models installed

Usage explaination

Python Poetry manually(advanced) installation

For developers to develop new features on Windows Powershell or on Ubuntu Linux or macOS

Usage in Python virtual environment

Example #1 send systeminfo and benchmark results to a remote server

Example #2 Do not send systeminfo and benchmark results to a remote server

Example #3 Benchmark run on explicitly given the path to the ollama executable (When you built your own developer version of ollama)

Reference

About

Releases 20

Languages

License

aidatatools/ollama-benchmark

Folders and files

Latest commit

History

Repository files navigation

llm-benchmark (ollama-benchmark)

Installation Steps

Usage for general users directly

Installation and Usage in Video format

ollama installation with the following models installed

Usage explaination

Python Poetry manually(advanced) installation

For developers to develop new features on Windows Powershell or on Ubuntu Linux or macOS

Usage in Python virtual environment

Example #1 send systeminfo and benchmark results to a remote server

Example #2 Do not send systeminfo and benchmark results to a remote server

Example #3 Benchmark run on explicitly given the path to the ollama executable (When you built your own developer version of ollama)

Reference

About

Topics

Resources

License

Stars

Watchers

Forks

Languages