All tracked LLMs

97 open-weight models with benchmarks and hardware requirements.

This is the complete catalogue of open-weight large language models tracked by CanItRun. For each model you can see the exact VRAM needed at every quantization level (FP16 down to Q2_K), the estimated inference speed in tokens per second on any GPU we track, and benchmark scores from the Open LLM Leaderboard v2 and LMSYS Chatbot Arena.

Models range from sub-1B edge models to 1T+ parameter frontier models. Use the homepage calculator to filter by your specific hardware, or browse below to compare models by parameter count. Each model page shows the full compatibility breakdown across all {60+} GPUs we track.

Browse by size

License
Releaser
Sort

97 models

Guides by use case

Each of these goes deeper than the filters above: VRAM breakdowns, GPU pairings, and tradeoffs specific to one category of model.