Foundation Models
Canonical reference pages for major LLMs and multimodal models. Each model links to benchmarks, guides, tools, research, and rankings.
14 models · Research feed · Benchmarks · LLMs guide
Featured Models
GPT-5
OpenAIOpenAI’s flagship general-purpose model for complex reasoning, coding, multimodal understanding, and agentic tool use.
Claude Opus
AnthropicAnthropic’s highest-capability Claude tier for deep reasoning, long-context analysis, coding, and careful instruction following.
Claude Sonnet
AnthropicAnthropic’s balanced Claude tier — strong quality at lower latency and cost than Opus, widely used for production agents and coding.
Gemini 2.5 Pro
GoogleGoogle’s Pro-class Gemini model for advanced reasoning and native multimodal workloads across text, images, audio, and long context.
DeepSeek V3
DeepSeekDeepSeek’s MoE general model — strong open-weight performance on coding and knowledge tasks with competitive API pricing.
DeepSeek R1
DeepSeekDeepSeek’s reasoning-focused model trained with reinforcement learning for multi-step math, science, and coding problem solving.
Qwen3
AlibabaAlibaba’s Qwen3 generation — strong multilingual open models spanning chat, reasoning modes, coding, and multimodal variants.
Llama 4
MetaMeta’s Llama 4 family — open-weight multimodal models designed for research and commercial use under Meta’s community license.
Proprietary
GPT-5
OpenAIOpenAI’s flagship general-purpose model for complex reasoning, coding, multimodal understanding, and agentic tool use.
Claude Opus
AnthropicAnthropic’s highest-capability Claude tier for deep reasoning, long-context analysis, coding, and careful instruction following.
Claude Sonnet
AnthropicAnthropic’s balanced Claude tier — strong quality at lower latency and cost than Opus, widely used for production agents and coding.
Gemini 2.5 Pro
GoogleGoogle’s Pro-class Gemini model for advanced reasoning and native multimodal workloads across text, images, audio, and long context.
Gemini Flash
GoogleGoogle’s fast, cost-efficient Gemini tier for high-throughput chat, classification, and multimodal apps where latency matters.
Mistral Large
MistralMistral’s flagship large model for enterprise reasoning, multilingual chat, and function calling via La Plateforme and cloud partners.
Grok
xAIxAI’s Grok model family — real-time oriented assistants with strong coding and reasoning, available via xAI API and consumer surfaces.
Command R+
CohereCohere’s Command R+ model optimized for retrieval-augmented generation, enterprise search, and multilingual business assistants.
Open Source
DeepSeek V3
DeepSeekDeepSeek’s MoE general model — strong open-weight performance on coding and knowledge tasks with competitive API pricing.
DeepSeek R1
DeepSeekDeepSeek’s reasoning-focused model trained with reinforcement learning for multi-step math, science, and coding problem solving.
Qwen3
AlibabaAlibaba’s Qwen3 generation — strong multilingual open models spanning chat, reasoning modes, coding, and multimodal variants.
Llama 4
MetaMeta’s Llama 4 family — open-weight multimodal models designed for research and commercial use under Meta’s community license.
Mixtral
MistralMistral’s sparse Mixture-of-Experts open models (e.g. Mixtral 8x7B / 8x22B) — efficient high-quality text generation for self-hosting.
Phi
MicrosoftMicrosoft’s Phi family of small language models — high capability per parameter for on-device, edge, and cost-sensitive deployments.
Reasoning
GPT-5
OpenAIOpenAI’s flagship general-purpose model for complex reasoning, coding, multimodal understanding, and agentic tool use.
Claude Opus
AnthropicAnthropic’s highest-capability Claude tier for deep reasoning, long-context analysis, coding, and careful instruction following.
Claude Sonnet
AnthropicAnthropic’s balanced Claude tier — strong quality at lower latency and cost than Opus, widely used for production agents and coding.
Gemini 2.5 Pro
GoogleGoogle’s Pro-class Gemini model for advanced reasoning and native multimodal workloads across text, images, audio, and long context.
DeepSeek V3
DeepSeekDeepSeek’s MoE general model — strong open-weight performance on coding and knowledge tasks with competitive API pricing.
DeepSeek R1
DeepSeekDeepSeek’s reasoning-focused model trained with reinforcement learning for multi-step math, science, and coding problem solving.
Qwen3
AlibabaAlibaba’s Qwen3 generation — strong multilingual open models spanning chat, reasoning modes, coding, and multimodal variants.
Llama 4
MetaMeta’s Llama 4 family — open-weight multimodal models designed for research and commercial use under Meta’s community license.
Mistral Large
MistralMistral’s flagship large model for enterprise reasoning, multilingual chat, and function calling via La Plateforme and cloud partners.
Grok
xAIxAI’s Grok model family — real-time oriented assistants with strong coding and reasoning, available via xAI API and consumer surfaces.
Command R+
CohereCohere’s Command R+ model optimized for retrieval-augmented generation, enterprise search, and multilingual business assistants.
Phi
MicrosoftMicrosoft’s Phi family of small language models — high capability per parameter for on-device, edge, and cost-sensitive deployments.
Coding
GPT-5
OpenAIOpenAI’s flagship general-purpose model for complex reasoning, coding, multimodal understanding, and agentic tool use.
Claude Opus
AnthropicAnthropic’s highest-capability Claude tier for deep reasoning, long-context analysis, coding, and careful instruction following.
Claude Sonnet
AnthropicAnthropic’s balanced Claude tier — strong quality at lower latency and cost than Opus, widely used for production agents and coding.
Gemini 2.5 Pro
GoogleGoogle’s Pro-class Gemini model for advanced reasoning and native multimodal workloads across text, images, audio, and long context.
DeepSeek V3
DeepSeekDeepSeek’s MoE general model — strong open-weight performance on coding and knowledge tasks with competitive API pricing.
DeepSeek R1
DeepSeekDeepSeek’s reasoning-focused model trained with reinforcement learning for multi-step math, science, and coding problem solving.
Qwen3
AlibabaAlibaba’s Qwen3 generation — strong multilingual open models spanning chat, reasoning modes, coding, and multimodal variants.
Llama 4
MetaMeta’s Llama 4 family — open-weight multimodal models designed for research and commercial use under Meta’s community license.
Mistral Large
MistralMistral’s flagship large model for enterprise reasoning, multilingual chat, and function calling via La Plateforme and cloud partners.
Mixtral
MistralMistral’s sparse Mixture-of-Experts open models (e.g. Mixtral 8x7B / 8x22B) — efficient high-quality text generation for self-hosting.
Grok
xAIxAI’s Grok model family — real-time oriented assistants with strong coding and reasoning, available via xAI API and consumer surfaces.
Phi
MicrosoftMicrosoft’s Phi family of small language models — high capability per parameter for on-device, edge, and cost-sensitive deployments.
Vision
GPT-5
OpenAIOpenAI’s flagship general-purpose model for complex reasoning, coding, multimodal understanding, and agentic tool use.
Claude Opus
AnthropicAnthropic’s highest-capability Claude tier for deep reasoning, long-context analysis, coding, and careful instruction following.
Claude Sonnet
AnthropicAnthropic’s balanced Claude tier — strong quality at lower latency and cost than Opus, widely used for production agents and coding.
Gemini 2.5 Pro
GoogleGoogle’s Pro-class Gemini model for advanced reasoning and native multimodal workloads across text, images, audio, and long context.
Gemini Flash
GoogleGoogle’s fast, cost-efficient Gemini tier for high-throughput chat, classification, and multimodal apps where latency matters.
Qwen3
AlibabaAlibaba’s Qwen3 generation — strong multilingual open models spanning chat, reasoning modes, coding, and multimodal variants.
Llama 4
MetaMeta’s Llama 4 family — open-weight multimodal models designed for research and commercial use under Meta’s community license.
Audio
Multimodal
GPT-5
OpenAIOpenAI’s flagship general-purpose model for complex reasoning, coding, multimodal understanding, and agentic tool use.
Claude Opus
AnthropicAnthropic’s highest-capability Claude tier for deep reasoning, long-context analysis, coding, and careful instruction following.
Claude Sonnet
AnthropicAnthropic’s balanced Claude tier — strong quality at lower latency and cost than Opus, widely used for production agents and coding.
Gemini 2.5 Pro
GoogleGoogle’s Pro-class Gemini model for advanced reasoning and native multimodal workloads across text, images, audio, and long context.
Gemini Flash
GoogleGoogle’s fast, cost-efficient Gemini tier for high-throughput chat, classification, and multimodal apps where latency matters.
Qwen3
AlibabaAlibaba’s Qwen3 generation — strong multilingual open models spanning chat, reasoning modes, coding, and multimodal variants.
Llama 4
MetaMeta’s Llama 4 family — open-weight multimodal models designed for research and commercial use under Meta’s community license.
Small Models
Gemini Flash
GoogleGoogle’s fast, cost-efficient Gemini tier for high-throughput chat, classification, and multimodal apps where latency matters.
Mixtral
MistralMistral’s sparse Mixture-of-Experts open models (e.g. Mixtral 8x7B / 8x22B) — efficient high-quality text generation for self-hosting.
Phi
MicrosoftMicrosoft’s Phi family of small language models — high capability per parameter for on-device, edge, and cost-sensitive deployments.