OpenRating is an independent, open-source AI rating platform. Join the community

OOpenRating

OpenRating Research

Research & Insights

Technical deep dives into benchmark integrity, LLM serving infrastructure forensics, and empirical pricing analysis.

OpenRating is an independent model evaluation platform. Every article here starts from primary measurements — benchmark runs, inference traces, and pricing data we collect ourselves — so the numbers behind the leaderboard are explainable.

What we cover

The research behind the rankings

OpenRating's leaderboard is built on measurements we run and verify ourselves. The blog publishes the methods, teardowns, and raw analysis behind those numbers — including the cases where popular benchmarks fail.

Benchmark integrity

How leaderboards can be gamed — and the forensic methods we use to keep OpenRating measurements honest.

AI hardware & serving

Silicon architectures, inference economics, and what watts-per-token means for the cost of every model you run.

Pricing & cost analysis

Empirical breakdowns of token pricing, cache economics, and the true per-task cost of frontier models.

Model forensics

Fingerprinting, proxy detection, and verification techniques for knowing which model is really answering.

Benchmark Newsletter

New teardowns & benchmark research in your inbox

Monthly empirical updates when labs drop new models. No marketing spam.

Join 4,200+ ML engineers and researchers. Unsubscribe at any time.