Smeltcore - Which AI models actually run on your GPU
by•
Honest benchmarks for running open-weights AI models on your own GPU. 84 models × 27 consumer GPUs — from RTX 3060 to Apple Silicon — each pair gets an honest verdict: runs, tight fit, or won't fit, with real speed and VRAM numbers. Plus 800+ step-by-step install recipes (llama.cpp, Ollama, ComfyUI, MLX) and a GPU Advisor. No invented numbers, no signup, free under CC BY-SA.

Replies
finally a benchmark site that doesn't oversell. checked my 4070 against a few models and the VRAM numbers actually matched what ollama shows locally. the install recipes are a nice touch too.
@zaferzker Thank you for your feedback!
Finally a benchmark site that doesn't oversell performance. Checked the RTX 3060 vs 7B models section and the VRAM numbers matched what I actually get running llama.cpp locally. That kind of honesty is refreshing.
the "tight fit" verdict is genuinely useful, like knowing exactly when you'll be sweating at 99% VRAM is way more honest than those blanket "it runs" claims. and no signup for the full thing under CC BY-SA is a really solid move.
A repo of actual install recipes and real VRAM numbers is genuinely useful, way better than guessing from random Reddit threads. One idea if you're open to it: add a simple toggle or filter to compare models head-to-head on the same GPU, like side-by-side tokens/sec for Llama 3 8B vs Qwen 2.5 7B on a 4070. Would make picking between similarly sized models a lot less of a guessing game.
Pazi
Love the idea behind Smeltcore. Big congratulations on shipping and launching! 🚀