OllamaBench is a full-stack AI benchmarking platform built for evaluating locally-deployed Large Language Models (LLMs) through Ollama. It provides automated benchmark suites, multi-model comparison, and hardware-aware performance analysis.
The platform enables developers to run comprehensive quality and performance evaluations entirely offline, with interactive dashboards for analyzing results and exportable benchmark reports for sharing findings.
Built with a modern stack — React with TypeScript for the frontend, FastAPI with Python for the backend — OllamaBench bridges the gap between simple model testing and production-grade evaluation infrastructure.