CHOOSE BY THE QUESTION

Model rankings & shortlists

Find candidates for your next deployment. Compare the specifications we can verify, then narrow the test to your workload.

Specifications first. Calculated rankings and editorial shortlists are labeled separately. Performance benchmarks are not available yet.

Our methodology

Explore the lists 6

Sources and ordering explained on every page

The measured rankings come next

These need reproducible GPU runs. No scores or winners have been assigned.

AWAITING TESTS

24 / 48 / 80 GB GPU fit

Working configurations at a stated precision, context, and concurrency.

AWAITING TESTS

Inference speed

Successful throughput and tail latency on the same hardware and workload.

AWAITING TESTS

Deployment cost

Measured capacity paired with dated instance pricing and utilization assumptions.

How to use these lists

Choose a question, inspect the inclusion rules, then open a model profile to check the original card. The lists cover a curated directory, not the whole open-model ecosystem. Position in an alphabetical shortlist does not indicate quality.

To compare individual candidates side by side, use the model comparison tool. For sizing context and runtime overhead, start with our GPU memory field note.