Can my PC run this LLM?

Type your GPU or Mac chip — see the models that actually run, their quant, the memory they need and an estimated tok/s band.

Popular picks

How it works

  1. Memory: weights + KV cache + compute buffer against what your GPU (or unified memory) can actually use.
  2. Speed: a memory-bandwidth model calibrated against 52 public benchmarks.
  3. Every number carries a band and a confidence label — measured, calibrated or theoretical.

Every number on this site ships with an error band and a confidence label: measured, calibrated (±12% for NVIDIA and AMD GPUs running a dense model fully on the GPU, ±20% otherwise), or theoretical (±30%).

Popular GPUs and Macs

Popular models