Known-issues sheets for local language models Rev 0.1.0 · reviewed every 7 days

Known-issues database · 41 models on file

What LLM your machine can run — and which one is a good fit for a specific job.

Pick your machine, say what you need the model for, and the output ranks what fits.

Step 01 — Your machine

Start from a machine like yours, or set the numbers below.

Machines with the same memory hold the same models, so what fits is the same answer for all of them. How fast they run it is not: a Tesla M40 and an RTX 4090 both hold a 24GB model, and the 4090 reads memory three and a half times quicker. Pick a machine and the recommendations below estimate both.

Your graphics card and your CPU both work here, and this asks about both. A model larger than the card does not stop: the card holds what it can and your CPU runs the rest out of system RAM, which is why the RAM figure and its speed change the answer even when you have a card. With no card at all, system RAM is the whole of it.

Step 02 — Workload

Output — Recommendation

Awaiting input. Choose a workload above.

All models