← Back to Portfolio
Project 07 · Developer Tools · AI

OllamaBench

Automated performance and quality evaluation for locally deployed Large Language Models. Multi-model comparison with hardware-aware analysis.

OllamaBench Dashboard
Overview

OllamaBench is a full-stack AI benchmarking platform built for evaluating locally-deployed Large Language Models (LLMs) through Ollama. It provides automated benchmark suites, multi-model comparison, and hardware-aware performance analysis.

The platform enables developers to run comprehensive quality and performance evaluations entirely offline, with interactive dashboards for analyzing results and exportable benchmark reports for sharing findings.

Built with a modern stack — React with TypeScript for the frontend, FastAPI with Python for the backend — OllamaBench bridges the gap between simple model testing and production-grade evaluation infrastructure.

Architecture
React + TypeScript Frontend ↓ FastAPI REST Backend ↓ Ollama Local Model Server ↓ Benchmark Suite Engine ↓ Hardware Metrics Collector ↓ Results Aggregation & Analysis ↓ Interactive Dashboard & Reports
Key Features
Platform Interface
Tech Stack
Frontend
React TypeScript Vite Tailwind CSS
Backend
Python FastAPI Benchmark Runner
Database & Storage
SQLite JSON Metric Storage
AI & Infrastructure
Ollama Local LLMs Hardware Monitor