Official AI Models & LLM Benchmark Hub (2026 Live Verified)

AI Models & LLM Intelligence Leaderboard

Comprehensive, zero-drift AI benchmarks aggregated across LMSYS Arena.ai (human Elo), Artificial Analysis (quality/speed ratio & price), SWE-bench Verified, GPQA Diamond, LiveBench, and MTEB.

#1 Frontier Reasoning
Claude Opus 5 (High)
Anthropic • Ultra-MoE
1504 Arena Elo • 78.6% SWE
#1 Test-Time Compute
GPT-5.6 Sol (xHigh)
OpenAI • Adaptive MoE
1498 Arena Elo • 97.6 Intel
#1 Quality / Speed Ratio
Gemini 3.7 Flash
Google • 2M Context
158 pts Ratio • 165 tps ($0.61/M)
#1 Open Weights MIT
DeepSeek-V4 Pro
DeepSeek • 750B MoE
1474 Arena Elo • $0.35/M
License: All Open Proprietary
# Model & Architecture Arena Elo (Human) SWE / GPQA Speed & Ratio Price ($/1M) Context Actions
1
BAAI BGE-en-ICL
BAAI 7B Dense 8GB open-source
N/A
76.5 MTEB
1.3k docs/s 94 pts
TTFT: 35ms
Free / Local
Blended 3:1
33k
Tokens
2
Voyage 3 Large
Voyage AI 1024-dim proprietary
N/A
75.8 MTEB
2.4k docs/s 93 pts
TTFT: 28ms
$0.12
Blended 3:1
32k
Tokens
3
BAAI BGE-Reranker-v2-m3
BAAI 567M Cross 4GB open-source
N/A
N/A MTEB
850 docs/s 92 pts
TTFT: 45ms
Free / Local
Blended 3:1
8k
Tokens
4
Cohere Rerank 3.5
Cohere Cross-Encoder proprietary
N/A
N/A MTEB
1.1k docs/s 92 pts
TTFT: 40ms
$1.00
Blended 3:1
4k
Tokens
5
text-embedding-3-large
OpenAI 3072-dim proprietary
N/A
71.2 MTEB
3.2k docs/s 88 pts
TTFT: 22ms
$0.13
Blended 3:1
8k
Tokens
6
Snowflake Arctic Embed M v2
Snowflake 305M Dense 2GB open-source
N/A
72.4 MTEB
4.8k docs/s 89 pts
TTFT: 15ms
Free / Local
Blended 3:1
8k
Tokens

Query this Leaderboard via REST API & Agents

Instant, zero-auth JSON endpoint with edge caching for programmatic model selection.

GET https://yakaai.com/api/leaderboard