Ecosystem Timeline

Key moments across seed companies, foundation models, and tools — curated entity events, linked back into the graph.

243 events from 12 companies, 13 models, and 10 tools.

2026

  1. Open sourceQwen3

    Qwen3.8-Flash-Next open weights

    Multimodal MoE (125B / 6B active + 51B n-gram embeddings) previewing the Qwen4 architecture; native 262K context (extensible to 1M). Cloud Qwen3.8-Flash is the production counterpart with 1M-default context and built-in tools.

  2. PlatformOpenAI

    GPT-5.6 Sol promotional API pricing

    Sol API drops to $4 input / $20 output per 1M tokens; promotional pricing through at least Nov 21, 2026. ChatGPT Pro/Plus/Business subscription prices unchanged.

  3. ModelDeepSeek

    V4-Flash-Vision-Exp multimodal API

    Experimental deepseek-v4-flash-vision-exp adds image+text at Flash rates; Files API enables free image reuse by file_id. API-only; open Flash weights remain text-first.

  4. PlatformGPT-5.6

    GPT-5.6 Sol promotional API pricing

    Sol API drops to $4 input / $20 output per 1M tokens (−20% / −33%); promotional pricing through at least Nov 21, 2026. Applies to API and eligible Codex/Work credits; ChatGPT subscription prices unchanged.

  5. ModelDeepSeek V4

    V4-Flash-Vision-Exp multimodal API

    Experimental deepseek-v4-flash-vision-exp adds image+text input at Flash rates (up to 384 tokens/image). Matches Flash on text; multimodal agents improve vs text Flash. API-only; no open vision weights yet. Files API for free image reuse shipped the same day.

  6. ProductGoogle DeepMind

    Antigravity in Gemini Enterprise + IDE extensions

    Antigravity ships inside eligible Gemini Enterprise subscriptions with admin/spend controls; VS Code, Visual Studio, JetBrains, and Zed extensions plus Antigravity 2.0 desktop and CLI.

  7. APIAnthropic

    Computer use GA + browser use toolset

    Computer use ships GA on the Claude API as computer_toolset_20260801 (no beta header; batch actions; zoom on by default). New browser_toolset_20260801 drives a host-provided browser viewport via accessibility tree and element refs. Available on Fable 5, Mythos 5, Opus 5, Sonnet 5, and Opus 4.8.

  8. APIAnthropic

    Files API and Agent Skills GA

    Files API (/v1/files) and Agent Skills (/v1/skills) leave beta on the Claude API—no beta headers required for GA request/response shapes. Earlier beta headers remain accepted for compatibility.

  9. APIClaude Opus

    Computer use GA + browser use

    Computer use is generally available as computer_toolset_20260801; browser_toolset_20260801 adds viewport/accessibility-tree control. Supported on Opus 5 (and peer Claude tiers).

  10. APIClaude Sonnet

    Computer use GA + browser use on Sonnet 5

    Sonnet 5 supports GA computer_toolset_20260801 and browser_toolset_20260801 on the Claude API for UI/browser automation agents.

  11. APIClaude Fable

    Computer use GA + browser use on Fable 5

    Fable 5 supports GA computer_toolset_20260801 and browser_toolset_20260801 on the Claude API.

  12. ProductCursor

    Cloud Agents Subscriptions and /goal

    Cloud agents can subscribe to PRs, Slack threads, or schedules; hold long-lived /goal objectives; run subagents on isolated VMs; and accept non-interrupting steering mid-run.

  13. ProductOpenAI

    ChatGPT for Teens

    Age-appropriate ChatGPT experience for ages 13–17 with stronger default safety protections, Study Mode nudges, and parental controls; auto-applied when age is stated or estimated under 18.

  14. ProductAnthropic

    Claude Code /design + Remote Control GA

    Week 34: /design research preview drafts editable UI artboards in CLI/Desktop (Pro/Max/Team/Enterprise, v2.1.233+). Remote Control leaves research preview so phone/claude.ai can start sessions on a local machine. Concise output style ships in v2.1.237.

  15. PlatformCursor

    Origin code hosting (early beta)

    Git forge inside Cursor for repos, pull requests, code browsing, and GitHub sync. Rolling out on paid plans; agent-native hosting features still to come.

  16. APIDeepSeek

    V4 peak/off-peak API pricing live

    Peak/off-peak rates replace the prior flat V4 API prices from 16:00 UTC. Peak hours 01:00–04:00 and 06:00–10:00 UTC; off-peak is half of peak.

  17. APIDeepSeek V4

    V4 peak/off-peak API pricing live

    Peak/off-peak rates replace the prior flat V4 API prices from 16:00 UTC. Peak hours 01:00–04:00 and 06:00–10:00 UTC; off-peak is half of peak.

  18. PlatformAnthropic

    Claude text watermarking (SynthID-Text)

    Anthropic documents SynthID-Text watermarks on Claude models launched on or after 2026-08-02 (EU AI Act transparency; applied globally). C2PA credentials on supported files; detection API forthcoming. Older models roll out over coming months.

  19. ProductAnthropic

    Claude Code auto mode default for Pro/Max/Team

    Auto mode becomes the default in Claude Code for Pro, Max, and Team plans; classifier overhead fees removed for those tiers. Enterprise remains opt-in.

  20. Open sourceQwen3

    Qwen3.8-27B open weights

    Dense 27B vision-language model on Hugging Face under Apache-2.0; native 262K context with image and video understanding.

  21. AcquisitionCursor

    Acquired by SpaceX

    Anysphere/Cursor becomes a wholly owned SpaceX subsidiary, completing the acquisition process that started with the SpaceXAI partnership in April.

  22. ProductGitHub Copilot

    Grok 4.6 in GitHub Copilot

    Grok 4.6 rolls out across VS Code, Visual Studio, CLI, cloud agent, Copilot app, JetBrains, Xcode, and Eclipse. Business/Enterprise admins must enable the policy (off by default).

  23. PlatformOpenAI

    GPT-5.6 Sol Ultrafast preview

    Limited API preview of Ultrafast mode for GPT-5.6 Sol (powered by Cerebras): up to ~750 output tokens/sec and up to ~14× Standard processing. Pricing and GA not published; waitlist expansion.

  24. ModelGoogle DeepMind

    Gemini 3.7 Flash GA

    Current Flash workhorse for coding and agents; introductory $0.75/$3.75 per 1M tokens through 2026-12-31.

  25. ModelDeepSeek

    DeepSeek V4-Pro generally available

    V4-Pro-0813 GA on app/web/API with agent upgrades; peak/off-peak API prices from 2026-08-16 16:00 UTC.

  26. Open sourceDeepSeek

    DeepSeek Harness developer preview

    MIT-licensed agent harness (dsh) with a plugin architecture (Cordis). GitHub repo created 2026-08-13; preview APIs may break.

  27. PlatformGPT-5.6

    Ultrafast mode preview for Sol

    Limited API preview of Ultrafast for GPT-5.6 Sol (Cerebras-backed): up to ~750 output tokens/sec and up to ~14× Standard. Pricing and GA not published.

  28. ModelGemini 3.1 Pro

    Gemini 3.7 Flash GA

    Current Flash workhorse (gemini-3.7-flash) for coding and agents; intro $0.75/$3.75 per 1M tokens through 2026-12-31.

  29. ModelDeepSeek V4

    V4-Pro-0813 generally available

    Pro leaves preview: agent upgrades, thinking effort low/high/max, native Responses API; peak/off-peak prices from 2026-08-16 16:00 UTC.

  30. ProductCursor

    Cloud Agent Builds

    Cursor prepares warm copies of cloud-agent environments in the background (no extra cost). Default-on for all environments on 2026-08-17.

  31. MilestoneCursor

    AIUC-1 agent security certification

    Independent AIUC-1 certification for agent security, safety, and reliability after a Schellman audit of controls and adversarial product tests.

  32. ProductGitHub Copilot

    Gemini 3.7 Flash in GitHub Copilot

    Google’s Gemini 3.7 Flash workhorse becomes a Copilot model option the same day as the model’s GA.

  33. Open sourceQwen3

    Qwen3.8-2.4T-A95B open weights

    Text-first 2.4T MoE weights on Hugging Face; cloud Qwen3.8-Max keeps vision, 1M-default context, and built-in tools.

  34. PlatformCursor

    Grok 4.6 in Cursor

    Grok 4.6 ships in Cursor with 2× included usage for the first week; tuned for long-running agents and visual work.

  35. PlatformGitHub Copilot

    Agent Plugins 1.0 generally available

    Portable plugins (skills + MCP servers) are GA in VS Code, Copilot CLI, the GitHub Copilot SDK, and the Copilot app on all Copilot plans.

  36. APIAnthropic

    Compliance API expands to Cowork and Claude Code

    Anthropic added Compliance API coverage for Cowork and Claude Code local session transcripts in beta for Enterprise organizations.

  37. Open sourceNVIDIA

    Nemotron 3.5 Lightning

    Open 30B MoE (3B active) for low-latency always-on agents, with NeMo Switchyard routing; weights on Hugging Face and build.nvidia.com.

  38. PlatformMistral AI

    Regional Endpoints GA + Priority Tier

    Customers can run inference in Europe or the US; Priority Tier public preview adds SLA-backed capacity. Platform also hosts third-party open models starting with Z.ai GLM-5.2.

  39. PlatformOpenAI

    Daybreak expands with GPT-5.6-Cyber

    OpenAI expanded Daybreak security access tiers and introduced GPT-5.6-Cyber for authorized defensive security research.

  40. APIAnthropic

    Sonnet 5 $2/$10 API pricing made standard

    Claude Sonnet 5 introductory rates ($2/$10 per 1M input/output tokens) become the standard price; the planned Sept 1, 2026 increase to $3/$15 is cancelled.

  41. Open sourceMeta

    Muse Glimmer 30B open weights

    Apache-2.0 30B on-device agent model from Meta Superintelligence Labs; distinct from closed Muse Spark and from Llama 4.

  42. APIClaude Sonnet

    Sonnet 5 $2/$10 pricing made standard

    Anthropic confirms Claude Sonnet 5 introductory API rates ($2 input / $10 output per 1M tokens) are now standard; the scheduled Sept 1, 2026 increase to $3/$15 will not occur.

  43. PlatformKimi K3

    Kimi K3 in GitHub Copilot

    Kimi K3 rolls out as a Copilot model option on Pro, Pro+, Max, Business, and Enterprise plans.

  44. Open sourceMuse Spark

    Muse Glimmer 30B open weights

    Apache-2.0 30B on-device agent model (sibling to closed Muse Spark); Hugging Face meta-models/Muse-Glimmer-30B.

  45. Open sourceMuse Glimmer

    Muse Glimmer 30B open weights

    Apache-2.0 ~30B dense multimodal agent model for on-device / single-GPU use; Hugging Face meta-models/Muse-Glimmer-30B. Distinct from closed Muse Spark and from Llama 4.

  46. PlatformMuse Glimmer

    Weights + quantizations on Hugging Face

    BF16 full weights, 4-bit variants for 24/32 GB GPUs, perception encoder, and DFlash drafter head published under Apache-2.0.

  47. PartnershipMuse Glimmer

    Edge and serving partner path

    Meta points developers to llama.cpp, MLX, ExecuTorch, Ollama, LM Studio, vLLM, SGLang, and hosts like Together / Fireworks / OpenRouter as integrations land.

  48. ReleasevLLM

    vLLM 0.27.0

    Day-0 Kimi K3 stack plus Qwen3.5/3.8-class model support; PyTorch 2.13 upgrade.

  49. ProductGitHub Copilot

    Kimi K3 in GitHub Copilot

    Moonshot’s Kimi K3 rolls out to Copilot Pro, Pro+, Max, Business, and Enterprise plans.

  50. DeprecationGemini 3.1 Pro

    gemini-2.5-pro shutdown scheduled

    Google lists gemini-2.5-pro shutdown for Oct 16, 2026; migrate to gemini-3.1-pro-preview.

  51. ProductGPT-5.6

    ChatGPT Sol update + Luna for Free/Go

    Plus/Pro get an updated GPT-5.6 Sol with a reasoning slider; Free/Go move to Luna defaults with unlimited text chats rolling out. Distinct from July Codex/Work API builds.

  52. PlatformPinecone

    Pinecone Nexus generally available

    Customer-cloud knowledge engine with KnowQL for agent-ready governed knowledge on top of Pinecone Database.

  53. ProductMeta

    Muse Spark 1.2 + Muse Code

    Coding-focused Muse Spark 1.2 and Muse Code terminal agent beta; Meta Model API access expands globally.

  54. ModelMuse Spark

    Muse Spark 1.2

    Coding-focused upgrade with more coding compute and environment diversity; available via Meta Model API with expanded global access.

  55. ProductMuse Spark

    Muse Code beta

    Terminal coding agent powered by Muse Spark 1.2 with persistent background subagents and a replay-exact event-log runtime.

  56. ModelMuse Glimmer

    Muse Spark 1.2 sibling context

    Closed Muse Spark 1.2 / Muse Code beta ship days earlier as the paid Meta Model API path; Glimmer is the open local counterpart.

  57. ReleaseQdrant

    Qdrant 1.19 Turbo4 + memory tiers

    Turbo4 4-bit TurboQuant storage datatype, unified pinned/cached/cold memory tiers, per-tenant IDF for sparse search, and keyword prefix filters.

  58. ReleaseWeaviate

    Weaviate 1.39.0

    Namespaces, 4-bit rotational quantization (RQ4), hybrid MMR, dedicated Search REST API, gRPC-web endpoint, and alter-schema improvements.

  59. ModelQwen3

    Qwen3.8-Max API

    2.4T / 95B-active Max-class model on QwenCloud (qwen3.8-max) for coding and long-horizon work.

  60. ProductCursor

    Google Workspace MCP plugins

    Marketplace plugins connect agents to Gmail, Google Drive, and Calendar without leaving the IDE.

  61. ModelDeepSeek V4

    V4-Flash-0731 agent upgrade

    Official Flash release with stronger agentic post-training; MIT weights on Hugging Face; same deepseek-v4-flash API ID.

  62. PlatformGPT-5.6

    Terra/Luna API price reductions

    Luna −80% and Terra −20% API pricing; Sol unchanged. Fast mode replaces Priority Processing in the API.

  63. ReleaseMilvus

    Milvus 3.0.0 lake-native GA

    External Collections index lake files in place (Parquet, Lance, Iceberg, Vortex); richer server-side retrieval (sparse SINDI, faceted search, schema evolution). Apache-2.0.

  64. ProductGitHub Copilot

    Grok 4.5 in GitHub Copilot

    xAI’s Grok 4.5 becomes available as a Copilot model option across VS Code, CLI, and cloud agents.

  65. Open sourceMoonshot AI

    Kimi K3 open weights

    Full checkpoint published on Hugging Face under the Kimi K3 License.

  66. Open sourceKimi K3

    Open weights on Hugging Face

    Full 2.8T checkpoint published under the Kimi K3 License (not MIT/Apache).

  67. ModelClaude Opus

    Claude Opus 5 GA

    Opus 5 ships as Anthropic’s everyday high-capability tier for agentic coding and 1M context at the same API rates as Opus 4.8.

  68. DeprecationDeepSeek V4

    Legacy chat/reasoner aliases retire

    deepseek-chat and deepseek-reasoner stop after Jul 24 2026 UTC; traffic moves to V4 Flash.

  69. PlatformCursor

    Cursor Router (Auto mode)

    Auto mode routes requests by task complexity with Cost, Balance, and Intelligence optimization modes.

  70. ModelGoogle DeepMind

    Gemini 3.6 Flash GA

    Flash workhorse ships generally available; 3.5 Pro remains in partner testing.

  71. ModelGemini 3.1 Pro

    Gemini 3.6 Flash GA

    Flash workhorse for agentic/coding/volume; Google confirms 3.5 Pro still testing with partners.

  72. ModelMoonshot AI

    Kimi K3 launch

    2.8T MoE multimodal agentic model ships on Kimi product and API.

  73. APIMoonshot AI

    platform.kimi.ai

    OpenAI/Anthropic-compatible API for Kimi K3 and related models.

  74. APIKimi K3

    Kimi K3 API / product launch

    Moonshot launches Kimi K3 via platform.kimi.ai and consumer Kimi surfaces.

  75. ResearchKimi K3

    KDA + Stable LatentMoE

    Kimi Delta Attention and 16-of-896 expert MoE target long-context efficiency.

  76. ModelKimi K3

    Native multimodal + 1M context

    Text, image, and video in one open model with a 1,048,576-token window.

  77. ModelMeta

    Muse Spark 1.1 + Meta Model API

    Paid multimodal agentic model via Meta’s first public Model API preview.

  78. ModelMuse Spark

    Muse Spark 1.1

    Major gains in agentic coding, tool/computer use, and multimodal understanding.

  79. APIMuse Spark

    Meta Model API public preview

    Meta’s first paid model API (OpenAI-compatible) opens Muse Spark to developers.

  80. ProductMuse Spark

    Thinking mode in Meta AI

    Consumer Meta AI / meta.ai surfaces expose Muse Spark Thinking mode.

  81. ModelGPT-5.6

    GPT-5.6 family (Sol / Terra / Luna)

    Capability tiers make routing explicit: Sol for peak, Terra balanced, Luna cost-efficient.

  82. APIGPT-5.6

    gpt-5.6 API routing

    API aliases route to Sol by default with Terra/Luna for cost control.

  83. ModelClaude Sonnet

    Claude Sonnet 5

    Sonnet 5 ships as the speed/intelligence balance for most production Claude traffic.

  84. APIClaude Sonnet

    Introductory pricing window

    Aggressive intro rates push Sonnet as the default production Claude.

  85. ModelClaude Fable

    Claude Fable 5 launch

    Most capable widely released Claude for long-horizon agents and deep reasoning.

  86. ProductClaude Fable

    Mythos 5 (Glasswing)

    Limited-access peer of Fable for approved Project Glasswing customers.

  87. APIClaude Fable

    claude-fable-5 API

    Generally available on Claude API and major cloud partners at premium rates.

  88. ProductClaude Fable

    Adaptive thinking + effort

    Effort parameter replaces manual thinking budgets for depth control.

  89. ProductClaude Fable

    1M context default

    Million-token context and large max outputs for long agent threads.

  90. ProductGPT-5.6

    MCP and agent tooling deepen

    Tool use, structured output, and MCP-style workflows become first-class for agents.

  91. ProductClaude Opus

    1M-token context class

    Document-heavy analysis and long agent threads become practical.

  92. ProductClaude Sonnet

    Long-context Sonnet

    1M-token-class context lands on the production tier, not only Opus.

  93. ModelGPT-5.6

    Multimodal + long context

    Text, image, and audio understanding with million-token-class context for research agents.

  94. ModelDeepSeek

    DeepSeek V4 Preview

    V4-Pro and V4-Flash bring 1M context and thinking modes to the DeepSeek API.

  95. ModelDeepSeek V4

    DeepSeek V4 Preview

    V4-Pro (1.6T/49B) and V4-Flash (284B/13B) ship with 1M context and thinking modes.

  96. APIDeepSeek V4

    deepseek-v4-pro / deepseek-v4-flash

    Official API IDs replace assuming historic chat/reasoner names map to V3/R1 weights.

  97. ModelMuse Spark

    Muse Spark first release

    Meta Superintelligence Labs ships its first closed multimodal Muse model.

  98. ModelGemini 3.1 Pro

    Gemini 3.1 Pro Preview

    Current Pro-class pin (gemini-3.1-pro-preview) for hard multimodal reasoning while 3.5 Pro remains in partner testing.

2025

  1. PlatformClaude Opus

    MCP support expands

    Model Context Protocol tooling broadens Claude’s agent integration surface.

  2. ModelClaude Haiku

    Claude Haiku 4.5 era

    Fast Claude tier with stronger reasoning at economical rates.

  3. ModelOpenAI

    GPT-5 family

    Next major foundation-model generation across chat and API.

  4. ModelAnthropic

    Claude 4 updates

    Next-generation Opus and Sonnet strengthen coding, agents, and long-document workflows.

  5. ModelMeta

    Llama 4

    Next open multimodal generation for self-hosted stacks.

  6. ModelLlama 4

    Llama 4 family

    Open-weight multimodal Scout / Maverick-class models for research and commercial use.

  7. ResearchLlama 4

    MoE variants

    Mixture-of-Experts variants expand efficiency for open multimodal serving.

  8. ModelQwen3

    Qwen3 generation

    Multilingual open models spanning chat, reasoning modes, coding, and multimodal variants.

  9. ProductQwen3

    Thinking / reasoning modes

    Explicit reasoning modes make Qwen competitive on hard multi-step tasks.

  10. ModelGoogle DeepMind

    Gemini 2.5 Pro

    Stronger reasoning and long-context performance for Vertex AI and AI Studio.

  11. ModelGemini 3.1 Pro

    Gemini 2.5 Pro class (legacy)

    Earlier Pro-class Gemini for advanced reasoning and native multimodal workloads.

  12. BenchmarkDeepSeek R1

    Reasoning benchmark wave

    GPQA / LiveBench-style comparisons put R1 on the open-reasoning map.

  13. MilestoneDeepSeek

    Industry cost/performance shock

    R1 coverage forces hyperscalers and labs to revisit pricing and open-weight strategy.

  14. ModelDeepSeek

    DeepSeek-R1

    Reasoning model resets expectations for open-weight chain-of-thought performance.

  15. ResearchDeepSeek R1

    DeepSeek-R1 technical report

    RL-trained reasoning model with public traces reshapes open-model expectations.

  16. Open sourceDeepSeek R1

    Open weights release

    Weights and distillations enable self-hosted reasoning baselines worldwide.

  17. APIDeepSeek R1

    DeepSeek API availability

    Hosted API makes R1 accessible without local GPU fleets.

  18. PlatformClaude Haiku

    Production high-volume default

    Becomes the default Claude choice when latency and unit economics dominate.

2024

  1. ModelDeepSeek

    DeepSeek-V3

    MoE open-weight model delivers frontier-competitive quality at lower cost.

  2. ModelDeepSeek R1

    V3 generalist context

    Strong MoE generalist precedes R1 and pairs with it in many stacks.

  3. ProductClaude Opus

    Computer-use style workflows

    UI and tool-driven agent patterns push Claude into operational automation.

  4. ProductCursor

    Enterprise adoption

    Privacy modes, team controls, and org rollouts make Cursor a standard engineering tool.

  5. ProductGitHub Copilot

    Agent mode expansion

    Multi-step agent workflows push Copilot closer to AI-native IDE competitors.

  6. ModelOpenAI

    o1 reasoning models

    Models optimized for multi-step reasoning before answering.

  7. ResearchGPT-5.6

    Reasoning-model era (o1 lineage)

    OpenAI popularizes deliberate reasoning traces that later feed GPT-5 thinking modes.

  8. PlatformSnowflake

    Cortex AI expansion

    Deepens SQL-native and app patterns for LLMs on enterprise data.

  9. ModelQwen3

    Qwen2.5 momentum

    Strong coding and size-ladder releases build the Qwen open ecosystem.

  10. ProductLangGraph

    Studio / debugging experience

    Visual debugging and graph inspection improve day-to-day agent development.

  11. PlatformvLLM

    Production hardening

    Metrics, continuous batching, and model coverage deepen for real traffic.

  12. PlatformLlamaIndex

    Production RAG patterns

    Indexing strategies, reranking, and eval loops become standard guidance.

  13. ModelClaude Sonnet

    Claude 3.5 Sonnet era

    Sonnet becomes the breakout Claude for coding and everyday agent work.

  14. PlatformDatabricks

    Mosaic AI expansion

    Deepens train/serve/evaluate for enterprise RAG and agent workloads.

  15. ProductClaude Haiku

    Tool use on Haiku

    Reliable tool calling makes Haiku practical for high-volume agent workers.

  16. ProductGemini 3.1 Pro

    Long-context Gemini push

    Million-token-class context becomes a Gemini product differentiator.

  17. PlatformQwen3

    vLLM serving maturity

    Production serving recipes for Qwen solidify across open inference stacks.

  18. PlatformLangGraph

    LangGraph Platform momentum

    Managed deployment and ops story expands beyond the open-source library.

  19. ProductPinecone

    Inference / embedding services

    Vector DB expands toward integrated embedding and inference workflows.

  20. PlatformCursor

    Multi-model backend era

    Teams routinely switch among GPT, Claude, and other models inside the same IDE.

  21. ProductLangGraph

    Human-in-the-loop patterns

    Interrupt/resume flows make approval steps practical in production agents.

  22. PlatformQdrant

    Distributed production deployments

    Clustering and scale-out patterns mature for larger self-hosted estates.

  23. ModelSnowflake

    Arctic models

    Signals Snowflake as both a data platform and a model contributor.

  24. ModelMeta

    Llama 3

    Major quality jump for open-weight chat and coding models.

  25. ModelLlama 4

    Llama 3 quality jump

    Major open-weight quality leap that sets up the Llama 4 generation.

  26. PlatformLangChain

    LangSmith production focus

    Observability, evaluation, and deployment tighten the path from prototype to production.

  27. ProductvLLM

    LoRA / multi-adapter serving

    Serving many fine-tuned adapters efficiently becomes a production requirement.

  28. ModelDatabricks

    DBRX open model

    Signals Databricks as both a platform and a model contributor in open LLM space.

  29. PlatformNVIDIA

    Blackwell architecture

    Next-generation accelerated computing platform for training and inference at AI-factory scale.

  30. PlatformNVIDIA

    NVIDIA NIM

    Microservices for deploying optimized inference in enterprise AI factories.

  31. ModelAnthropic

    Claude 3 family

    Opus, Sonnet, and Haiku establish a clear quality–latency–cost ladder for production.

  32. ModelClaude Opus

    Claude 3 Opus launches

    Establishes Opus as Anthropic’s peak-capability brand in the Claude family.

  33. ModelClaude Sonnet

    Claude 3 Sonnet launches

    Defines the mid-tier Claude brand for balanced cost and capability.

  34. ModelClaude Haiku

    Claude 3 Haiku launches

    Defines Haiku as Anthropic’s speed and cost brand in the Claude family.

  35. ModelClaude Haiku

    Vision on Haiku

    Image understanding lands on the fast tier for multimodal pipelines.

  36. PlatformLangGraph

    Checkpoints & persistence mature

    Durable state, retries, and long-running workflows become first-class.

  37. PlatformLlamaIndex

    Workflows / agents expansion

    Framework grows from pure RAG into broader workflow and agent patterns.

  38. ModelMistral AI

    Mistral Large

    Flagship commercial model strengthens La Plateforme for enterprise APIs.

  39. ProductMistral AI

    Le Chat

    Consumer/business chat surface for Mistral models.

  40. ModelGemini 3.1 Pro

    Gemini 1.5 era

    Native multimodal + long context establishes Gemini’s modern identity.

  41. PlatformWeaviate

    Generative / RAG features

    Closer integration with generative search patterns for end-to-end RAG apps.

  42. PlatformGitHub Copilot

    Copilot Enterprise era

    Org knowledge, admin controls, and enterprise packaging deepen GitHub-native AI.

  43. Open sourceLangGraph

    LangGraph open release

    Graph-based runtime for durable, stateful LangChain agents.

  44. ProductLangChain

    LangGraph introduced

    Stateful graph runtime complements LangChain’s higher-level abstractions.

  45. APIDeepSeek

    DeepSeek API platform

    OpenAI-compatible API makes DeepSeek easy to try in existing clients.

  46. PlatformQwen3

    Hugging Face distribution

    Hub becomes the default discovery path for Qwen weights and demos.

  47. PlatformPinecone

    Serverless / usage-based era

    Packaging shifts toward more elastic, usage-based vector infrastructure.

  48. ProductCursor

    Agentic coding features expand

    Multi-file edits and agent-style workflows deepen IDE autonomy.

  49. PlatformvLLM

    Ecosystem default for open serving

    Becomes a standard choice next to TGI and TensorRT-LLM in many stacks.

2023

  1. ModelMicrosoft

    Phi small models

    Shows competitive quality from small, efficient models for edge and cost-sensitive use.

  2. ModelMistral AI

    Mixtral 8x7B

    Sparse MoE open model becomes a default efficient alternative to dense LLMs.

  3. ModelGoogle DeepMind

    Gemini 1.0

    Natively multimodal foundation model family launches across consumer and cloud surfaces.

  4. ModelGemini 3.1 Pro

    Gemini family launches

    Google’s unified multimodal foundation-model brand debuts.

  5. PlatformOpenAI

    Assistants API

    Platform primitives for tool use, retrieval, and agent-style apps.

  6. PlatformSnowflake

    Snowflake Cortex

    Brings LLM functions and AI services into the governed Data Cloud.

  7. ProductQdrant

    Hybrid search capabilities

    Sparse + dense retrieval patterns expand beyond pure vector similarity.

  8. ModelMistral AI

    Mistral 7B

    Small open model punches above its size and accelerates European open LLM adoption.

  9. ProductCursor

    Multi-file editing / Composer

    Repo-aware edits across files shift Cursor from autocomplete to agentic coding.

  10. APIvLLM

    OpenAI-compatible serving

    Drop-in API compatibility accelerates migration from hosted APIs to self-host.

  11. ProductLangChain

    Expression Language / LCEL era

    Composable runnables become the preferred way to build chains and pipelines.

  12. ProductLlamaIndex

    Evaluation & observability focus

    RAG quality tooling expands beyond connectors and indexes.

  13. Open sourceMeta

    Llama 2 open release

    Openly available weights for commercial and research use.

  14. Open sourceLlama 4

    Llama 2 commercial license

    Open weights become broadly usable for commercial products under Meta’s terms.

  15. ModelAnthropic

    Claude 2

    Longer context and stronger conversational performance for enterprise assistants.

  16. FoundedDeepSeek

    DeepSeek founded

    Hangzhou lab focuses on efficient, high-capability open models.

  17. AcquisitionDatabricks

    MosaicML acquisition

    Brings training and generative AI tooling into Mosaic AI on the Lakehouse.

  18. ProductPinecone

    Hybrid / metadata retrieval focus

    Sparse-dense and metadata filtering become table stakes for enterprise RAG.

  19. Open sourcevLLM

    PagedAttention / vLLM rise

    Open serving engine makes high-throughput LLM inference widely accessible.

  20. ProductMilvus

    GPU indexing momentum

    GPU-accelerated indexing deepens Milvus’s large-scale search positioning.

  21. MilestoneNVIDIA

    AI infrastructure surge

    GPUs become the default substrate for training and inference at global scale.

  22. ProductWeaviate

    Hybrid search prominence

    Keyword + vector retrieval becomes a primary reason teams evaluate Weaviate.

  23. LeadershipGoogle DeepMind

    Google DeepMind formed

    Google Brain and DeepMind combine—research, models, and product under one AI organization.

  24. FoundedMistral AI

    Mistral AI founded

    Paris-based lab founded by former DeepMind / Meta researchers.

  25. LeadershipGemini 3.1 Pro

    Google Brain × DeepMind

    Research and model product organizations consolidate under one AI org.

  26. ProductLlamaIndex

    Data connector ecosystem

    Ingestion from docs, APIs, and SaaS sources becomes a core product strength.

  27. ProductGitHub Copilot

    Copilot Chat

    Chat moves Copilot beyond autocomplete into conversational coding help.

  28. ModelOpenAI

    GPT-4 introduced

    Multimodal foundation model with stronger reasoning and coding.

  29. PlatformLangChain

    Integration explosion

    Model, vector-store, and tool connectors make LangChain the glue layer for LLM stacks.

  30. ProductQdrant

    Payload filtering strength

    Rich metadata filters become a defining reason teams pick Qdrant for RAG.

  31. ProductCursor

    Cursor IDE gains traction

    AI-native fork of VS Code becomes a default coding environment for many teams.

  32. ModelMeta

    LLaMA to researchers

    Research weights that sparked the open-model wave.

  33. ResearchLlama 4

    LLaMA research weights

    Research release that sparks the modern open-model wave.

  34. PartnershipMicrosoft

    Expanded OpenAI investment

    Deepens Azure OpenAI as the enterprise channel for ChatGPT-class models.

  35. PlatformHugging Face

    Hub as open-model hub

    Models, datasets, and Spaces concentrate open AI collaboration and distribution.

  36. FoundedMoonshot AI

    Moonshot AI founded

    Beijing lab builds the Kimi assistant and frontier MoE models.

2022

  1. ProductOpenAI

    ChatGPT released

    Consumer chat product that brought LLMs into mainstream use.

  2. Open sourceLlamaIndex

    LlamaIndex (GPT Index) adoption

    Becomes a leading open framework for connecting LLMs to private data.

  3. PlatformHugging Face

    Inference Endpoints

    Managed production inference for Hub models without self-hosting ops.

  4. Open sourceLangChain

    LangChain gains adoption

    Becomes the default open framework for composing LLM apps.

  5. ProductGitHub Copilot

    GitHub Copilot GA

    Generally available as a paid coding assistant for individuals and teams.

  6. PlatformPinecone

    Pod-based scale era

    Production customers standardize on managed indexes for semantic search and RAG.

  7. PlatformQdrant

    Qdrant Cloud

    Managed option alongside self-hosted deployments.

  8. ModelGoogle DeepMind

    PaLM

    Pathways Language Model showcases large-scale Google LLM research before Gemini.

  9. PlatformNVIDIA

    Hopper architecture (H100)

    Defines the GPU generation that trains and serves most frontier models of the mid-2020s.

  10. PlatformWeaviate

    Weaviate Cloud

    Managed offering expands beyond self-hosted deployments.

  11. PlatformMilvus

    Zilliz Cloud managed offering

    Managed Milvus becomes the enterprise path without self-host ops.

2021

  1. PlatformHugging Face

    Spaces

    Turns the Hub into a place to demo and share interactive ML apps.

  2. ProductMicrosoft

    GitHub Copilot

    Coding assistant brings LLM copilots into everyday developer workflows.

  3. ProductGitHub Copilot

    GitHub Copilot technical preview

    Inline AI completions bring LLM coding assistants into mainstream IDEs.

  4. ProductWeaviate

    Modules / vectorizers

    Pluggable modules make embedding and enrichment part of the DB workflow.

  5. Open sourceMilvus

    CNCF sandbox / cloud-native path

    Milvus joins the cloud-native ecosystem path for large-scale vector infra.

  6. ProductPinecone

    Pinecone product traction

    Managed vector search becomes mainstream for production RAG.

  7. FoundedAnthropic

    Anthropic founded

    Safety-focused lab founded by former OpenAI researchers; Constitutional AI becomes a defining idea.

  8. Open sourceQdrant

    Qdrant open-source growth

    Rust vector engine gains adoption for filtered similarity search.

  9. PlatformMilvus

    Milvus 2.x distributed architecture

    Cloud-native redesign enables billion-scale indexes and separation of storage/compute.

2020

  1. FundingSnowflake

    Snowflake IPO

    Public listing accelerates enterprise Data Cloud adoption.

  2. APIOpenAI

    OpenAI API launched

    Developers gain API access to GPT models for production apps.

  3. Open sourceHugging Face

    Datasets library

    Standardizes how the community loads, shares, and versions training data.

  4. PlatformDatabricks

    Lakehouse positioning

    Unifies data warehousing and data lakes as the substrate for analytics and AI.

2019

  1. Open sourceMilvus

    Milvus open-sourced

    Zilliz releases Milvus as an open vector database for similarity search.

  2. PartnershipOpenAI

    Microsoft strategic partnership

    Microsoft invests $1B and becomes a primary cloud partner.

  3. PartnershipMicrosoft

    OpenAI partnership begins

    Multi-year investment and Azure exclusivity set up GPT distribution.

  4. Open sourceWeaviate

    Weaviate open-source vector DB

    Early open vector database with a modular architecture.

2018

  1. Open sourceHugging Face

    Transformers library

    Becomes the default open toolkit for pretrained NLP models.

  2. Open sourceDatabricks

    MLflow open-sourced

    Becomes the default open toolkit for experiment tracking and model registry.

2017

  1. ResearchGoogle DeepMind

    Transformer architecture published

    “Attention Is All You Need” (Google Research) becomes the foundation of modern LLMs.

  2. PlatformNVIDIA

    TensorRT era

    Inference optimization stack becomes central to production deep learning serving.

2016

  1. Open sourceMeta

    PyTorch released

    Becomes the dominant research framework and later the industry standard for training modern neural nets.

  2. ResearchGoogle DeepMind

    AlphaGo vs Lee Sedol

    Landmark demonstration of deep reinforcement learning on a grandmaster-level game.

  3. FoundedHugging Face

    Hugging Face founded

    Started as a chat app company; pivoted into open ML tooling.

2015

  1. FoundedOpenAI

    OpenAI established

    Founded as a nonprofit AI research lab in San Francisco.

2013

  1. FoundedDatabricks

    Databricks founded

    Created by the creators of Apache Spark to commercialize big data and ML platforms.

2012

  1. FoundedSnowflake

    Snowflake founded

    Cloud data platform company founded to separate storage and compute.

2010

  1. FoundedGoogle DeepMind

    DeepMind founded

    London research lab focused on general-purpose AI systems.

  2. PlatformMicrosoft

    Azure cloud era

    Azure becomes the cloud substrate for later AI services and OpenAI hosting.

2006

  1. PlatformNVIDIA

    CUDA introduced

    GPU programming platform that later underpins modern AI training and inference.

1975

  1. FoundedMicrosoft

    Microsoft founded

    Started as a microcomputer software company in Albuquerque.