Models & benchmarks

Compare AI models without losing the evidence

Follow frontier, open and local models, see what benchmark results actually measure, and connect capability to access, hardware and cost.

Model landscape

Frontier and commercial models

Track the major model families, providers, access routes, context windows, published prices and practical strengths without treating marketing claims as benchmarks.

text modelsreasoningcodingmultimodal AIlong contextagentsconsumer accessAPI access
Run it yourself

Open, local and self-hosted AI

Understand open-weight licensing, local deployment, memory needs and the hardware trade-offs that determine what can realistically run outside a hosted service.

open weightslicensesquantizationVRAMRAMGPUsNPUsAI PCs
Measure capability

Benchmarks and evaluations

Compare models using separate evidence for different abilities instead of collapsing every task into one misleading score.

reasoningmathcodingagent taskslong contextdocumentsred teamingevaluation limits
Choose intelligently

Capability, access and value

A useful model comparison includes what it can do, what it costs, how it is accessed, whether it can be deployed locally and how strong the supporting evidence is.

cost per tokenfree accesssubscriptionslatencyprivacydeployment controlmodel fit