Compare leading AI models through one auditable decision layer for performance, coding, value, speed, and efficiency.
Built for engineers, researchers, and teams choosing models without relying on vendor claims.
General keeps the Artificial Analysis order. Capability views use LLMStats names, scores, and ranks as published.
Performance Leaderboard: Published intelligence scores are used unchanged. No cost, speed, or completeness penalty influences this ranking.
General intelligence, pricing and API performance data from Artificial Analysis.
Experimental model-specific reactions from recent Reddit, Hacker News, and X when its official API is configured. Developer issue templates and news headlines are excluded from community quotes. Sentiment never changes leaderboard rankings.
Follow each company’s model releases chronologically and see how intelligence scores changed from one generation to the next.
Select a model company to view its release timeline.
Releases are placed by their published date and grouped by product family. Reasoning-effort variants are collapsed into one generation.
Ask a question in plain language. The Advisor answers only from the current LLMDEX dataset and shows whether Gemini or deterministic local analysis produced the result.
Download processed datasets for the dashboard, Power BI, and external analysis. Source failures are visible and never replaced with fabricated rows.
LLMDEX is designed as an auditable, source-transparent LLM analytics platform. Every ranking decision is documented, every data point is traceable, and no value is ever fabricated. Below is how the system works.
Artificial Analysis powers General, price, and API performance. LLMStats powers source-native capability views. LLMDEX preserves both observations separately and links them only at family level.
The Artificial Analysis Intelligence Index is used unchanged. LLMStats capability views preserve every benchmark column shown by the source. Missing values stay unavailable and never become zero.
The family score is 50% AA matched-universe percentile and 50% LLMStats General percentile. Raw source scores are never averaged; missing or unresolved families remain unscored.
Exact IDs, URLs, approved aliases, then provider + family + version are eligible. Approved variants share the same family LLMDEX score; source-native scores and model configurations remain separate.
Composite: 50% Performance + 30% Cost Efficiency + 20% Speed. Missing components redistribute weight proportionally. Uses adjusted performance for fairness.
Efficiency = Performance / Blended Cost, normalized as percentile ranks to prevent chart distortion. Only models with adjusted performance ≥ 25 qualify. Models with a listed zero API price may rank highest on API-price efficiency. Self-hosting, hardware, engineering, and operating costs are not included.
General intelligence, model pricing, generation throughput, latency, and API-performance observations are sourced from Artificial Analysis .
General and capability-specific benchmark rankings are sourced from LLMStats .
LLMDEX independently processes, links, and presents these observations. LLMDEX is not affiliated with, endorsed by, or sponsored by either source. Model coverage, metric definitions, providers, and update times may differ between sources.
LLMDEX does not independently execute every upstream benchmark. Third-party observations remain subject to their respective source terms.