RANKED • CURATED
LLMs Leaderboard
Tools with an independently verified benchmark score rank first, by that real-world score. Everything else is ranked by curated priority: quality, reliability, and unique capabilities.
RANK BY CATEGORY
All Tools
346tools
LLMs
119tools
IDEs & Coding Tools
60tools
Text → Image
54tools
Multimodal Reasoning
46tools
Image → Video
45tools
Text → Video
42tools
Image → Image
39tools
Image → 3D
28tools
Text → 3D
27tools
Text → Audio
26tools
AI Assistants
17tools
Video → Video
12tools
Multi-Service Platforms
10tools
Infrastructure
7tools
Agentic Browsers
6tools
REAL BENCHMARK SCORES
Source: Artificial Analysis Intelligence Index, as of 2026-08-04. Shown only for models with independently verified scores. Not every tool in this category has published, comparable data.
RESULTS
| Rank | Tool | Modality | Pricing |
|---|---|---|---|
| ① |
Claude Opus 5
Anthropic's frontier model, currently first on the Artificial Analysis Intelligence Index
|
LLMs, IDEs & Coding Tools, Multimodal Reasoning | Freemium |
| ② |
Claude Fable 5
Anthropic's Mythos-class creative model
|
LLMs, Multimodal Reasoning | Paid |
| ③ |
GPT-5.6 Sol
OpenAI's top-tier model for complex professional work
|
LLMs, IDEs & Coding Tools | Paid |
| 4 |
Kimi K3
Moonshot AI's 2.8-trillion-parameter open-weight flagship
|
LLMs, IDEs & Coding Tools | Freemium |
| 5 |
Claude Opus 4.8
Anthropic's powerful enterprise model from May 2026
|
LLMs, IDEs & Coding Tools | Enterprise |
| 6 |
GPT-5.5 Instant
OpenAI's fast default ChatGPT model from May 2026
|
LLMs | Freemium |
| 7 |
Grok 4.5
xAI's flagship coding model, trained in partnership with Cursor
|
LLMs, IDEs & Coding Tools | Paid |
| 8 |
Qwen 3.8-Max
Alibaba's 2.4-trillion-parameter flagship, currently in preview
|
LLMs, IDEs & Coding Tools | Paid |
| 9 |
Claude Sonnet 5
Anthropic's cheaper, near-Opus everyday model
|
LLMs, IDEs & Coding Tools | Freemium |
| 10 |
GLM-5.2
753B open-weight MoE coding model with a 1M-token context, MIT licensed
|
LLMs, IDEs & Coding Tools | Freemium |
| 11 |
Muse Spark 1.1
Meta's closed-weight agentic model, and its first paid model API
|
LLMs, Multimodal Reasoning | Freemium |
| 12 |
Gemini 3.5 Flash
Google's fast, capable multimodal model from I/O 2026
|
LLMs, Multimodal Reasoning | Freemium |
| 13 |
Gemini 3.6 Flash
Google's faster, sharper agentic-coding upgrade to 3.5 Flash
|
LLMs, IDEs & Coding Tools | Freemium |
| 14 |
MiniMax M3
MiniMax's 1M-context agentic frontier model
|
LLMs, AI Assistants | Freemium |
| 15 |
DeepSeek V4-Pro
DeepSeek's open-weight model with permanent pricing
|
LLMs, IDEs & Coding Tools, Multimodal Reasoning | Freemium |
| 16 |
Kimi K2.7-Code
Moonshot's specialized coding model
|
LLMs, IDEs & Coding Tools | Freemium |
| 17 |
Inkling
Thinking Machines' 975B Apache-2.0 model that takes text, images and audio natively
|
LLMs, Multimodal Reasoning | Free |
| 18 |
Claude Opus 4.6
The ceiling of enterprise autonomy with 1M context
|
LLMs, IDEs & Coding Tools | Enterprise |
| 19 |
NVIDIA Nemotron 3 Ultra
NVIDIA's 550B open-weights reasoning model, built for inference speed
|
LLMs | Free |
| 20 |
Claude 4.6 Sonnet
The industry standard for coding and nuanced instruction following
|
LLMs, IDEs & Coding Tools | Paid |
| 21 |
StepFun Step 3.7 Flash
StepFun's 198B MoE vision-language model
|
LLMs, Multimodal Reasoning | Freemium |
| 22 |
NVIDIA Nemotron 3 Nano Omni
One multimodal model for text, vision, audio, and video reasoning
|
LLMs, Multimodal Reasoning | Paid |
| 23 |
NotebookLM
Google's AI Research Assistant: The Ultimate Study Tool
|
LLMs, Text → Audio, AI Assistants | Free |
| 24 |
Grok
xAI's real-time AI assistant
|
LLMs | Paid |
| 25 |
DeepSeek
The Efficiency Revolution: Frontier Intelligence at 1/100th the Cost
|
LLMs, Multimodal Reasoning | Freemium |
| 26 |
Llama
Meta's open-source large language model
|
LLMs | Free |
| 27 |
Mistral AI
European open-source and commercial LLM
|
LLMs | Freemium |
| 28 |
Cohere
Enterprise-focused LLM platform
|
LLMs | Enterprise |
| 29 |
Qwen
Alibaba's multilingual open-source LLM
|
LLMs | Freemium |
| 30 |
Microsoft Phi
Microsoft's efficient small language models
|
LLMs | Free |
| 31 |
Gemma
Google's open-source lightweight LLM
|
LLMs | Free |
| 32 |
Kimi k1.5
The 'Next DeepSeek' Movement: o1-Level Reasoning at 1/100th the Cost
|
LLMs, Multimodal Reasoning | Freemium |
| 33 |
Qwen 2.5-VL
The Open Vision-Reasoner: SOTA Multimodal Performance
|
Multimodal Reasoning, LLMs | Free |
| 34 |
DBRX
Databricks' high-performance open-source LLM
|
LLMs | Enterprise |
| 35 |
Llama 3.2 Vision
Meta's Open Multimodal Standard
|
Multimodal Reasoning, LLMs | Free |
| 36 |
Pixtral Large
The Open Vision Frontier: 124B Multimodal Power
|
Multimodal Reasoning, LLMs | Freemium |
| 37 |
InternVL 2.5
The Open-Source Vision Giant: 78B Multimodal Leader
|
Multimodal Reasoning, LLMs | Free |
| 38 |
Cursor Composer 2.5
Cursor's agentic coding model for multi-file software engineering
|
IDEs & Coding Tools, LLMs | Paid |
| 39 |
OpenCode
Open-source, model-agnostic terminal coding agent
|
IDEs & Coding Tools, LLMs | Free |
| 40 |
Grok Build
xAI's agentic coding CLI for autonomous software engineering
|
IDEs & Coding Tools, LLMs | Freemium |
| 41 |
Gemini Omni
Google's unified multimodal generation model
|
LLMs, Text → Image, Text → Video, Text → Audio, Multimodal Reasoning | Freemium |
| 42 |
Gemini Spark
Google's personal AI agent for proactive assistance
|
AI Assistants, LLMs | Freemium |
| 43 |
Mistral Vibe
Mistral's unified work and coding agent
|
LLMs, IDEs & Coding Tools, AI Assistants | Freemium |
| 44 |
GPT-5.6 Terra
OpenAI's balanced GPT-5.6 model for intelligence and cost
|
LLMs, IDEs & Coding Tools | Freemium |
| 45 |
GPT-5.6 Luna
OpenAI's cost-optimized GPT-5.6 model for high-volume workloads
|
LLMs | Freemium |
| 46 |
Kimi K2.7 Code Highspeed
Faster inference variant of Kimi's coding specialist
|
LLMs, IDEs & Coding Tools | Freemium |
| 47 |
Claude Mythos 5
Limited-availability Mythos-class model without Fable 5 safety classifiers
|
LLMs, Multimodal Reasoning | Enterprise |
| 48 |
NVIDIA Nemotron 3 Nano
Compact 30B open-weight model with configurable reasoning for agents
|
LLMs | Free |
| 49 |
NVIDIA Nemotron 3 Super
120B open-weight hybrid MoE for efficient multi-agent reasoning
|
LLMs | Free |
| 50 |
GLM-5-Turbo
Optimized GLM-5 variant for fast sequential task execution
|
LLMs, IDEs & Coding Tools | Paid |
| 51 |
GLM-5V-Turbo
Multimodal coding and visual-reasoning agent model
|
LLMs, Multimodal Reasoning | Paid |
| 52 |
Mistral Medium 3.5
Mistral's mid-tier workhorse for reasoning, coding, and instruction
|
LLMs, IDEs & Coding Tools | Freemium |
| 53 |
MiniMax M2.7
Recursive self-improvement language model for real-world engineering
|
LLMs, IDEs & Coding Tools | Freemium |
| 54 |
MiniMax M2.7 Highspeed
Same M2.7 performance with significantly faster inference
|
LLMs, IDEs & Coding Tools | Freemium |
| 55 |
Mistral Small 4
Unified open-source small model for chat, reasoning, vision, and coding
|
LLMs, IDEs & Coding Tools, Multimodal Reasoning | Freemium |
| 56 |
DeepSeek V4-Flash
High-volume DeepSeek inference with a 1M-token context window
|
LLMs, IDEs & Coding Tools, Multimodal Reasoning | Freemium |
| 57 |
Hunyuan Hy3 Preview
Tencent's latest open-source MoE flagship with tool use
|
LLMs, IDEs & Coding Tools | Freemium |
| 58 |
Kimi K2.6
Moonshot's open-weight multimodal successor with long-context coding stability
|
LLMs, Multimodal Reasoning, IDEs & Coding Tools | Freemium |
| 59 |
Claude Opus 4.7
Frontier Opus model with higher-resolution vision and xhigh effort
|
LLMs, IDEs & Coding Tools, Multimodal Reasoning | Enterprise |
| 60 |
Mistral Large 3
Mistral's flagship open-weight multimodal frontier model
|
LLMs, Multimodal Reasoning | Freemium |
| 61 |
Ministral 3
Mistral's edge family of small, dense open-source models
|
LLMs | Freemium |
| 62 |
GLM-5.1
MIT-licensed MoE flagship for 8-hour autonomous coding sessions
|
LLMs, IDEs & Coding Tools | Paid |
| 63 |
GLM-4.7-Flash
Free universal GLM model with a 200K context window
|
LLMs | Free |
| 64 |
GLM-4V-Flash
Free vision model for image understanding and document snapshots
|
LLMs, Multimodal Reasoning | Free |
| 65 |
Grok 4.3
xAI's long-context flagship with a 1M-token window
|
LLMs, Multimodal Reasoning | Paid |
| 66 |
GLM-5
744B-parameter open-weight MoE flagship for agentic planning and execution
|
LLMs, IDEs & Coding Tools | Paid |
| 67 |
Qwen3-Coder-Next
80B parameter open-weight coding powerhouse
|
LLMs, IDEs & Coding Tools | Free |
| 68 |
GPT-5.3 Codex
The frontier model for complex reasoning and software architecture
|
LLMs, Multimodal Reasoning | Paid |
| 69 |
Gemini 3 Ultra
Native multimodal intelligence with a 10M context window
|
LLMs, Multimodal Reasoning | Paid |
| 70 |
Perplexity AI
The conversational search engine that replaced traditional search
|
LLMs | Freemium |
| 71 |
Consensus
AI search engine for peer-reviewed scientific research
|
LLMs | Freemium |
| 72 |
Grok 4.20
xAI's 2M-context beta model with multi-agent capabilities
|
LLMs, Multimodal Reasoning | Paid |
| 73 |
Kimi K2.5
Moonshot's open-weight multimodal generalist with agent swarms
|
LLMs, Multimodal Reasoning, IDEs & Coding Tools | Freemium |
| 74 |
Hunyuan TurboS
Tencent's fast, cost-efficient flagship Hunyuan model
|
LLMs | Freemium |
| 75 |
DeepSeek V3.2
The 128K-context MoE flagship that introduced sparse attention
|
LLMs, IDEs & Coding Tools, Multimodal Reasoning | Freemium |
| 76 |
GLM-4.7
Strong general-reasoning model with interleaved thinking
|
LLMs | Paid |
| 77 |
GLM-4.6V
Vision-language model for visual reasoning and UI replication
|
LLMs, Multimodal Reasoning | Paid |
| 78 |
Claude Opus 4.5
First Claude model with the effort parameter and context compaction
|
LLMs, IDEs & Coding Tools | Enterprise |
| 79 |
Hunyuan A13B Instruct
Tencent's efficient small-scale MoE instruct model
|
LLMs | Freemium |
| 80 |
Cursor Composer 1
Cursor's first-generation agentic coding model
|
IDEs & Coding Tools, LLMs | Paid |
| 81 |
Grok 4.1 Fast
xAI's high-volume, 2M-context workhorse model
|
LLMs | Paid |
| 82 |
Claude Haiku 4.5
Fast, low-cost Claude model with extended thinking and Computer Use
|
LLMs, AI Assistants | Paid |
| 83 |
GLM-OCR
Document parsing model for PDF and image OCR
|
LLMs, Multimodal Reasoning | Paid |
| 84 |
Claude Sonnet 4.5
Balanced Sonnet model with major coding and agentic improvements
|
LLMs, IDEs & Coding Tools | Paid |
| 85 |
Hunyuan 2.0 Think
The deep-thinking variant of Hunyuan 2.0
|
LLMs, Multimodal Reasoning | Freemium |
| 86 |
GLM-4.6
Mid-range coding and tool-calling model with 200K context
|
LLMs, IDEs & Coding Tools | Paid |
| 87 |
Grok Code Fast 1
xAI's fast, cheap coding specialist model
|
LLMs, IDEs & Coding Tools | Paid |
| 88 |
GLM-4.5-Air
Cost-efficient reasoning, coding, and agent model
|
LLMs, IDEs & Coding Tools | Paid |
| 89 |
Hunyuan 2.0 Instruct
Tencent's general-purpose instruction-tuned Hunyuan 2.0 model
|
LLMs | Freemium |
| 90 |
Qwen 3
Alibaba's open-source MoE flagship with thinking modes
|
LLMs, IDEs & Coding Tools, AI Assistants | Free |
| 91 |
Llama 4 Maverick
Meta's open-weight flagship with native multimodal reasoning
|
LLMs, Multimodal Reasoning | Free |
| 92 |
Llama 4 Scout
Long-context, efficient open multimodal model for edge and single-GPU use
|
LLMs, Multimodal Reasoning | Free |
| 93 |
Gemini 2.5 Pro
Google's high-performance reasoning model with advanced coding
|
LLMs, IDEs & Coding Tools, Multimodal Reasoning | Freemium |
| 94 |
Hunyuan T1
Tencent's Mamba-powered deep-thinking reasoning model
|
LLMs, Multimodal Reasoning | Freemium |
| 95 |
Gemma 3
Google's open multimodal model for research and developers
|
LLMs, Multimodal Reasoning | Free |
| 96 |
Perplexity Sonar Deep Research
Autonomous research agent that performs multi-source deep dives
|
LLMs, AI Assistants | Enterprise |
| 97 |
Qwen 2.5-Max
Alibaba's closed-API flagship before Qwen 3
|
LLMs, IDEs & Coding Tools | Paid |
| 98 |
Perplexity Sonar Reasoning Pro
Chain-of-thought reasoning model for multi-step logical analysis
|
LLMs | Paid |
| 99 |
DeepSeek R1
The open-weight reasoning model that sparked the efficiency revolution
|
LLMs, Multimodal Reasoning | Freemium |
| 100 |
Gemini 2.0 Flash
Google's low-latency agentic model with native tool use
|
LLMs, Multimodal Reasoning, AI Assistants | Freemium |
| 101 |
Llama 3.3
Efficient 70B open model matching 405B quality
|
LLMs | Free |
| 102 |
Llama-3.1-Nemotron Ultra
NVIDIA-aligned 253B Llama 3.1 for helpfulness and instruction following
|
LLMs | Free |
| 103 |
Llama-3.1-Nemotron Super
NVIDIA-aligned 49B Llama 3.1 for balanced performance
|
LLMs | Free |
| 104 |
Llama-3.1-Nemotron Nano
NVIDIA-aligned 8B Llama 3.1 for efficient inference
|
LLMs | Free |
| 105 |
Qwen 2.5-Coder
Alibaba's open coding-specialist model
|
LLMs, IDEs & Coding Tools | Free |
| 106 |
Hunyuan Large
Tencent's 389B-parameter open-source MoE language model
|
LLMs | Free |
| 107 |
Perplexity Sonar Pro
Advanced search model with deeper reasoning and richer citations
|
LLMs | Paid |
| 108 |
Qwen-Math
Open mathematical reasoning specialist
|
LLMs | Free |
| 109 |
Llama 3.1 405B
The first frontier-scale open-weight language model
|
LLMs | Free |
| 110 |
Gemini 1.5 Flash
Fast, cost-efficient multimodal model with a 1M context window
|
LLMs, Multimodal Reasoning | Freemium |
| 111 |
Gemini 1.5 Pro
Google's long-context multimodal flagship with up to 2M tokens
|
LLMs, Multimodal Reasoning | Freemium |
| 112 |
Qwen-VL-Max
Alibaba's strongest vision-language model
|
LLMs, Multimodal Reasoning | Freemium |
| 113 |
Perplexity Sonar
Lightweight, real-time search model for cited answers at low cost
|
LLMs | Paid |
| 114 |
Microsoft 365 Copilot
AI assistant embedded across Word, Excel, PowerPoint, Outlook, and Teams
|
LLMs, AI Assistants | Enterprise |
| 115 |
Microsoft Copilot
Microsoft's everyday AI assistant across web, PC, and mobile
|
LLMs, AI Assistants, Text → Image | Freemium |
| 116 |
Microsoft MAI Models (Build 2026)
Microsoft's unified AI model family from Build 2026
|
LLMs, Multimodal Reasoning, Text → Audio, Text → Image, IDEs & Coding Tools, AI Assistants | Enterprise |
| 117 |
Baidu ERNIE 4.5
Open-source MoE LLM with strong Chinese NLP and multimodal capabilities
|
LLMs | Freemium |
| 118 |
GLM-4.5
Advanced multilingual LLM with enhanced reasoning and long-context support
|
LLMs | Freemium |
| 119 |
Manus AI
Autonomous AI agent for complex multi-step workflows and research automation
|
LLMs | Enterprise |
No tools match your search/filter.