Enterprise DNA Enterprise DNA
P Apps and SaaS Productivity low

LLM Stats

by Various

The LLM Leaderboard — independent ranking of GPT, Claude, Gemini, Llama, DeepSeek and 300+ AI models by intelligence, speed and price. Composite LLM Stats Score updated continuou

LLM Stats screenshot

Apps

LLM Stats

Added 1 June 2026

Overview

LLM Stats is an independent leaderboard ranking over 300 AI models by intelligence, speed, and price. It provides a composite LLM Stats Score that updates continuously using public benchmarks and live API metrics.

Best for

Best for
Developers and researchers who need to objectively compare model options for a project.

Use cases

  • Compare model performance across intelligence, speed, and price dimensions.
  • Select the most cost-effective model for a given application.
  • Track ranking changes and score trends over time for popular models.

Notes

LLM Stats is an independent leaderboard ranking over 300 AI models by intelligence, speed, and price. It provides a composite LLM Stats Score that updates continuously using public benchmarks and live API metrics.

Use cases

  • Compare model performance across intelligence, speed, and price dimensions.
  • Select the most cost-effective model for a given application.
  • Track ranking changes and score trends over time for popular models.

Pros

  • Covers a wide range of models including GPT, Claude, Gemini, Llama, and DeepSeek.
  • Composite score simplifies comparison across multiple benchmarks.
  • Data is continuously updated from live API metrics and public benchmarks.

Cons

  • Benchmarks may not fully represent real-world task performance.
  • Live API metrics can be affected by provider reliability or access restrictions.
  • Does not include fine-tuned or custom model variants.

Indexed from awesome-generative-ai and enriched against its public facts.

Pros

  • Covers a wide range of models including GPT, Claude, Gemini, Llama, and DeepSeek.
  • Composite score simplifies comparison across multiple benchmarks.
  • Data is continuously updated from live API metrics and public benchmarks.

Cons

  • Benchmarks may not fully represent real-world task performance.
  • Live API metrics can be affected by provider reliability or access restrictions.
  • Does not include fine-tuned or custom model variants.
Free 27-page guide

Get the free Developer’s Field Guide

A 27-page field guide to the AI coding workflow with Claude. Claude Code, MCP servers, the prompt patterns that work, and what to delegate. Free.

Enter your work email. We send it straight over, plus a few short notes worth knowing. Unsubscribe any time.

No spam. Unsubscribe any time.