Explore Local AI Models

Interactive scatter plot and sortable table of every supported local LLM. Toggle the X and Y axes between VRAM, quality, tok/s, disk size, or context length to find the best model for your hardware.

Available axes

VRAM required (GB)How much GPU memory the model needs at the selected quant
Disk size (GB)Storage required for the GGUF file
Model size (params B)Total parameter count in billions
Est. tok/sBandwidth-based throughput estimate on selected GPU
Quality index0–100 aggregate quality score from Artificial Analysis
Context lengthMaximum context window in tokens

Presets

Best valueHighlights the Pareto frontier — highest quality per GB of VRAM
FastestSorts and colours models by estimated tokens/sec on your GPU
Fits my hardwareHides models that won't run on your selected accelerator

Click any point or table row to add the model to the Compare tool.