Independent AI leaderboard ranking 300+ LLMs by a composite score across reasoning, coding, agent capability, speed, and price, updated continuously.
LLM Stats is an independent benchmarking platform that ranks 300+ large language models including GPT, Claude, Gemini, Llama, and DeepSeek. Rankings are updated continuously from public benchmarks and live API metrics. The composite LLM Stats Score aggregates GPQA Diamond, SWE-Bench Verified, coding-arena performance, output throughput, and per-token pricing. The platform also runs contamination-proof independent evaluations not present in model training data. LLM Stats is a product of LLM Stats.
For people
For agents
Nothing listed yet.