DataAIHub
DataAIHubNews · Research · Tools · Learning

Foundation Model Landscape

Seed foundation models grouped by provider — open-weight vs proprietary — from Entity Platform attributes.

All landscapes · Built from Entity Platform seeds — not a separate directory.

Alibaba

1 entity

Foundation models from Alibaba.

  • Qwen3

    Alibaba’s Qwen3 generation — strong multilingual open models spanning chat, reasoning modes, coding, and multimodal variants.

Anthropic

4 entities

Foundation models from Anthropic.

  • Claude Opus

    Anthropic’s Claude Opus 5 tier for complex agentic coding, enterprise work, long-context analysis, and careful instruction following. Claude Fable 5 sits above Opus for peak widely released capability.

  • Claude Sonnet

    Anthropic’s Claude Sonnet 5 tier — best combination of speed and intelligence for most production agents and coding, at lower cost than Opus.

  • Claude Fable

    Anthropic’s Claude Fable 5 — the most capable widely released Claude for long-horizon agents, deep reasoning, and demanding coding workflows. Mythos 5 is the limited-access peer for Project Glasswing.

  • Claude Haiku

    Anthropic’s fast, cost-efficient Claude tier for high-volume chat, classification, extraction, and sub-agent steps where latency and price matter more than peak reasoning.

DeepSeek

2 entities

Foundation models from DeepSeek.

  • DeepSeek R1

    DeepSeek’s reasoning-focused model trained with reinforcement learning for multi-step math, science, and coding problem solving.

  • DeepSeek V4

    DeepSeek’s V4 generation — deepseek-v4-pro (1.6T / 49B active) and deepseek-v4-flash (284B / 13B active) with 1M context, dual thinking modes, and strong agentic coding. Flash-0731 is the current Flash API revision.

Foundation models from Google DeepMind.

  • Gemini 2.5 Pro

    Google’s Pro-class Gemini model for advanced reasoning and native multimodal workloads across text, images, audio, and long context.

Meta

2 entities

Foundation models from Meta.

  • Llama 4

    Meta’s Llama 4 family — open-weight multimodal models designed for research and commercial use under Meta’s community license.

  • Muse Spark

    Meta Superintelligence Labs’ Muse Spark 1.1 — closed multimodal reasoning model for agentic tasks, coding, computer use, and 1M-context workflows via the Meta Model API (public preview).

Moonshot AI

1 entity

Foundation models from Moonshot AI.

  • Kimi K3

    Moonshot’s Kimi K3 — 2.8T MoE (104B active) open-weight multimodal agentic model with 1M context, native vision, and strong long-horizon coding. Weights on Hugging Face under the Kimi K3 License.

OpenAI

1 entity

Foundation models from OpenAI.

  • GPT-5.6

    OpenAI’s GPT-5.6 family (Sol flagship, Terra balanced, Luna cost-efficient) for complex reasoning, coding, multimodal understanding, and agentic tool use. The gpt-5.6 API alias routes to Sol.

Open weights

5 entities

Seed models with open-weight / open-source availability.

  • Llama 4

    Meta’s Llama 4 family — open-weight multimodal models designed for research and commercial use under Meta’s community license.

  • DeepSeek R1

    DeepSeek’s reasoning-focused model trained with reinforcement learning for multi-step math, science, and coding problem solving.

  • DeepSeek V4

    DeepSeek’s V4 generation — deepseek-v4-pro (1.6T / 49B active) and deepseek-v4-flash (284B / 13B active) with 1M context, dual thinking modes, and strong agentic coding. Flash-0731 is the current Flash API revision.

  • Kimi K3

    Moonshot’s Kimi K3 — 2.8T MoE (104B active) open-weight multimodal agentic model with 1M context, native vision, and strong long-horizon coding. Weights on Hugging Face under the Kimi K3 License.

  • Qwen3

    Alibaba’s Qwen3 generation — strong multilingual open models spanning chat, reasoning modes, coding, and multimodal variants.

API / proprietary

7 entities

Seed models primarily available via hosted APIs.

  • GPT-5.6

    OpenAI’s GPT-5.6 family (Sol flagship, Terra balanced, Luna cost-efficient) for complex reasoning, coding, multimodal understanding, and agentic tool use. The gpt-5.6 API alias routes to Sol.

  • Claude Opus

    Anthropic’s Claude Opus 5 tier for complex agentic coding, enterprise work, long-context analysis, and careful instruction following. Claude Fable 5 sits above Opus for peak widely released capability.

  • Claude Sonnet

    Anthropic’s Claude Sonnet 5 tier — best combination of speed and intelligence for most production agents and coding, at lower cost than Opus.

  • Claude Fable

    Anthropic’s Claude Fable 5 — the most capable widely released Claude for long-horizon agents, deep reasoning, and demanding coding workflows. Mythos 5 is the limited-access peer for Project Glasswing.

  • Claude Haiku

    Anthropic’s fast, cost-efficient Claude tier for high-volume chat, classification, extraction, and sub-agent steps where latency and price matter more than peak reasoning.

  • Gemini 2.5 Pro

    Google’s Pro-class Gemini model for advanced reasoning and native multimodal workloads across text, images, audio, and long context.

  • Muse Spark

    Meta Superintelligence Labs’ Muse Spark 1.1 — closed multimodal reasoning model for agentic tasks, coding, computer use, and 1M-context workflows via the Meta Model API (public preview).