Ecosystem Timeline
Key moments across seed companies, foundation models, and tools — curated entity events, linked back into the graph.
243 events from 12 companies, 13 models, and 10 tools.
2026
- Open sourceQwen3
Qwen3.8-Flash-Next open weights
Multimodal MoE (125B / 6B active + 51B n-gram embeddings) previewing the Qwen4 architecture; native 262K context (extensible to 1M). Cloud Qwen3.8-Flash is the production counterpart with 1M-default context and built-in tools.
- PlatformOpenAI
GPT-5.6 Sol promotional API pricing
Sol API drops to $4 input / $20 output per 1M tokens; promotional pricing through at least Nov 21, 2026. ChatGPT Pro/Plus/Business subscription prices unchanged.
- ModelDeepSeek
V4-Flash-Vision-Exp multimodal API
Experimental deepseek-v4-flash-vision-exp adds image+text at Flash rates; Files API enables free image reuse by file_id. API-only; open Flash weights remain text-first.
- PlatformGPT-5.6
GPT-5.6 Sol promotional API pricing
Sol API drops to $4 input / $20 output per 1M tokens (−20% / −33%); promotional pricing through at least Nov 21, 2026. Applies to API and eligible Codex/Work credits; ChatGPT subscription prices unchanged.
- ModelDeepSeek V4
V4-Flash-Vision-Exp multimodal API
Experimental deepseek-v4-flash-vision-exp adds image+text input at Flash rates (up to 384 tokens/image). Matches Flash on text; multimodal agents improve vs text Flash. API-only; no open vision weights yet. Files API for free image reuse shipped the same day.
- ProductGoogle DeepMind
Antigravity in Gemini Enterprise + IDE extensions
Antigravity ships inside eligible Gemini Enterprise subscriptions with admin/spend controls; VS Code, Visual Studio, JetBrains, and Zed extensions plus Antigravity 2.0 desktop and CLI.
- APIAnthropic
Computer use GA + browser use toolset
Computer use ships GA on the Claude API as computer_toolset_20260801 (no beta header; batch actions; zoom on by default). New browser_toolset_20260801 drives a host-provided browser viewport via accessibility tree and element refs. Available on Fable 5, Mythos 5, Opus 5, Sonnet 5, and Opus 4.8.
- APIAnthropic
Files API (/v1/files) and Agent Skills (/v1/skills) leave beta on the Claude API—no beta headers required for GA request/response shapes. Earlier beta headers remain accepted for compatibility.
- APIClaude Opus
Computer use is generally available as computer_toolset_20260801; browser_toolset_20260801 adds viewport/accessibility-tree control. Supported on Opus 5 (and peer Claude tiers).
Computer use GA + browser use on Sonnet 5
Sonnet 5 supports GA computer_toolset_20260801 and browser_toolset_20260801 on the Claude API for UI/browser automation agents.
- APIClaude Fable
Computer use GA + browser use on Fable 5
Fable 5 supports GA computer_toolset_20260801 and browser_toolset_20260801 on the Claude API.
- ProductCursor
Cloud Agents Subscriptions and /goal
Cloud agents can subscribe to PRs, Slack threads, or schedules; hold long-lived /goal objectives; run subagents on isolated VMs; and accept non-interrupting steering mid-run.
- ProductOpenAI
Age-appropriate ChatGPT experience for ages 13–17 with stronger default safety protections, Study Mode nudges, and parental controls; auto-applied when age is stated or estimated under 18.
- ProductAnthropic
Claude Code /design + Remote Control GA
Week 34: /design research preview drafts editable UI artboards in CLI/Desktop (Pro/Max/Team/Enterprise, v2.1.233+). Remote Control leaves research preview so phone/claude.ai can start sessions on a local machine. Concise output style ships in v2.1.237.
- PlatformCursor
Origin code hosting (early beta)
Git forge inside Cursor for repos, pull requests, code browsing, and GitHub sync. Rolling out on paid plans; agent-native hosting features still to come.
- APIDeepSeek
V4 peak/off-peak API pricing live
Peak/off-peak rates replace the prior flat V4 API prices from 16:00 UTC. Peak hours 01:00–04:00 and 06:00–10:00 UTC; off-peak is half of peak.
- APIDeepSeek V4
V4 peak/off-peak API pricing live
Peak/off-peak rates replace the prior flat V4 API prices from 16:00 UTC. Peak hours 01:00–04:00 and 06:00–10:00 UTC; off-peak is half of peak.
- PlatformAnthropic
Claude text watermarking (SynthID-Text)
Anthropic documents SynthID-Text watermarks on Claude models launched on or after 2026-08-02 (EU AI Act transparency; applied globally). C2PA credentials on supported files; detection API forthcoming. Older models roll out over coming months.
- ProductAnthropic
Claude Code auto mode default for Pro/Max/Team
Auto mode becomes the default in Claude Code for Pro, Max, and Team plans; classifier overhead fees removed for those tiers. Enterprise remains opt-in.
- Open sourceQwen3
Dense 27B vision-language model on Hugging Face under Apache-2.0; native 262K context with image and video understanding.
- AcquisitionCursor
Anysphere/Cursor becomes a wholly owned SpaceX subsidiary, completing the acquisition process that started with the SpaceXAI partnership in April.
- ProductGitHub Copilot
Grok 4.6 rolls out across VS Code, Visual Studio, CLI, cloud agent, Copilot app, JetBrains, Xcode, and Eclipse. Business/Enterprise admins must enable the policy (off by default).
- PlatformOpenAI
Limited API preview of Ultrafast mode for GPT-5.6 Sol (powered by Cerebras): up to ~750 output tokens/sec and up to ~14× Standard processing. Pricing and GA not published; waitlist expansion.
- ModelGoogle DeepMind
Current Flash workhorse for coding and agents; introductory $0.75/$3.75 per 1M tokens through 2026-12-31.
- ModelDeepSeek
DeepSeek V4-Pro generally available
V4-Pro-0813 GA on app/web/API with agent upgrades; peak/off-peak API prices from 2026-08-16 16:00 UTC.
- Open sourceDeepSeek
DeepSeek Harness developer preview
MIT-licensed agent harness (dsh) with a plugin architecture (Cordis). GitHub repo created 2026-08-13; preview APIs may break.
- PlatformGPT-5.6
Ultrafast mode preview for Sol
Limited API preview of Ultrafast for GPT-5.6 Sol (Cerebras-backed): up to ~750 output tokens/sec and up to ~14× Standard. Pricing and GA not published.
- ModelGemini 3.1 Pro
Current Flash workhorse (gemini-3.7-flash) for coding and agents; intro $0.75/$3.75 per 1M tokens through 2026-12-31.
- ModelDeepSeek V4
V4-Pro-0813 generally available
Pro leaves preview: agent upgrades, thinking effort low/high/max, native Responses API; peak/off-peak prices from 2026-08-16 16:00 UTC.
- ProductCursor
Cursor prepares warm copies of cloud-agent environments in the background (no extra cost). Default-on for all environments on 2026-08-17.
- MilestoneCursor
AIUC-1 agent security certification
Independent AIUC-1 certification for agent security, safety, and reliability after a Schellman audit of controls and adversarial product tests.
- ProductGitHub Copilot
Gemini 3.7 Flash in GitHub Copilot
Google’s Gemini 3.7 Flash workhorse becomes a Copilot model option the same day as the model’s GA.
- Open sourceQwen3
Qwen3.8-2.4T-A95B open weights
Text-first 2.4T MoE weights on Hugging Face; cloud Qwen3.8-Max keeps vision, 1M-default context, and built-in tools.
- PlatformCursor
Grok 4.6 ships in Cursor with 2× included usage for the first week; tuned for long-running agents and visual work.
- PlatformGitHub Copilot
Agent Plugins 1.0 generally available
Portable plugins (skills + MCP servers) are GA in VS Code, Copilot CLI, the GitHub Copilot SDK, and the Copilot app on all Copilot plans.
- APIAnthropic
Compliance API expands to Cowork and Claude Code
Anthropic added Compliance API coverage for Cowork and Claude Code local session transcripts in beta for Enterprise organizations.
- Open sourceNVIDIA
Open 30B MoE (3B active) for low-latency always-on agents, with NeMo Switchyard routing; weights on Hugging Face and build.nvidia.com.
- PlatformMistral AI
Regional Endpoints GA + Priority Tier
Customers can run inference in Europe or the US; Priority Tier public preview adds SLA-backed capacity. Platform also hosts third-party open models starting with Z.ai GLM-5.2.
- PlatformOpenAI
Daybreak expands with GPT-5.6-Cyber
OpenAI expanded Daybreak security access tiers and introduced GPT-5.6-Cyber for authorized defensive security research.
- APIAnthropic
Sonnet 5 $2/$10 API pricing made standard
Claude Sonnet 5 introductory rates ($2/$10 per 1M input/output tokens) become the standard price; the planned Sept 1, 2026 increase to $3/$15 is cancelled.
- Open sourceMeta
Apache-2.0 30B on-device agent model from Meta Superintelligence Labs; distinct from closed Muse Spark and from Llama 4.
Sonnet 5 $2/$10 pricing made standard
Anthropic confirms Claude Sonnet 5 introductory API rates ($2 input / $10 output per 1M tokens) are now standard; the scheduled Sept 1, 2026 increase to $3/$15 will not occur.
- PlatformKimi K3
Kimi K3 rolls out as a Copilot model option on Pro, Pro+, Max, Business, and Enterprise plans.
- Open sourceMuse Spark
Apache-2.0 30B on-device agent model (sibling to closed Muse Spark); Hugging Face meta-models/Muse-Glimmer-30B.
- Open sourceMuse Glimmer
Apache-2.0 ~30B dense multimodal agent model for on-device / single-GPU use; Hugging Face meta-models/Muse-Glimmer-30B. Distinct from closed Muse Spark and from Llama 4.
- PlatformMuse Glimmer
Weights + quantizations on Hugging Face
BF16 full weights, 4-bit variants for 24/32 GB GPUs, perception encoder, and DFlash drafter head published under Apache-2.0.
- PartnershipMuse Glimmer
Meta points developers to llama.cpp, MLX, ExecuTorch, Ollama, LM Studio, vLLM, SGLang, and hosts like Together / Fireworks / OpenRouter as integrations land.
- ReleasevLLM
Day-0 Kimi K3 stack plus Qwen3.5/3.8-class model support; PyTorch 2.13 upgrade.
- ProductGitHub Copilot
Moonshot’s Kimi K3 rolls out to Copilot Pro, Pro+, Max, Business, and Enterprise plans.
- DeprecationGemini 3.1 Pro
gemini-2.5-pro shutdown scheduled
Google lists gemini-2.5-pro shutdown for Oct 16, 2026; migrate to gemini-3.1-pro-preview.
- ProductGPT-5.6
ChatGPT Sol update + Luna for Free/Go
Plus/Pro get an updated GPT-5.6 Sol with a reasoning slider; Free/Go move to Luna defaults with unlimited text chats rolling out. Distinct from July Codex/Work API builds.
- PlatformPinecone
Pinecone Nexus generally available
Customer-cloud knowledge engine with KnowQL for agent-ready governed knowledge on top of Pinecone Database.
- ProductMeta
Coding-focused Muse Spark 1.2 and Muse Code terminal agent beta; Meta Model API access expands globally.
- ModelMuse Spark
Coding-focused upgrade with more coding compute and environment diversity; available via Meta Model API with expanded global access.
- ProductMuse Spark
Terminal coding agent powered by Muse Spark 1.2 with persistent background subagents and a replay-exact event-log runtime.
- ModelMuse Glimmer
Muse Spark 1.2 sibling context
Closed Muse Spark 1.2 / Muse Code beta ship days earlier as the paid Meta Model API path; Glimmer is the open local counterpart.
- ReleaseQdrant
Qdrant 1.19 Turbo4 + memory tiers
Turbo4 4-bit TurboQuant storage datatype, unified pinned/cached/cold memory tiers, per-tenant IDF for sparse search, and keyword prefix filters.
- ReleaseWeaviate
Namespaces, 4-bit rotational quantization (RQ4), hybrid MMR, dedicated Search REST API, gRPC-web endpoint, and alter-schema improvements.
- ModelQwen3
2.4T / 95B-active Max-class model on QwenCloud (qwen3.8-max) for coding and long-horizon work.
- ProductCursor
Marketplace plugins connect agents to Gmail, Google Drive, and Calendar without leaving the IDE.
- ModelDeepSeek V4
Official Flash release with stronger agentic post-training; MIT weights on Hugging Face; same deepseek-v4-flash API ID.
- PlatformGPT-5.6
Terra/Luna API price reductions
Luna −80% and Terra −20% API pricing; Sol unchanged. Fast mode replaces Priority Processing in the API.
- ReleaseMilvus
External Collections index lake files in place (Parquet, Lance, Iceberg, Vortex); richer server-side retrieval (sparse SINDI, faceted search, schema evolution). Apache-2.0.
- ProductGitHub Copilot
xAI’s Grok 4.5 becomes available as a Copilot model option across VS Code, CLI, and cloud agents.
- Open sourceMoonshot AI
Full checkpoint published on Hugging Face under the Kimi K3 License.
- Open sourceKimi K3
Full 2.8T checkpoint published under the Kimi K3 License (not MIT/Apache).
- ModelClaude Opus
Opus 5 ships as Anthropic’s everyday high-capability tier for agentic coding and 1M context at the same API rates as Opus 4.8.
- DeprecationDeepSeek V4
Legacy chat/reasoner aliases retire
deepseek-chat and deepseek-reasoner stop after Jul 24 2026 UTC; traffic moves to V4 Flash.
- PlatformCursor
Auto mode routes requests by task complexity with Cost, Balance, and Intelligence optimization modes.
- ModelGoogle DeepMind
Flash workhorse ships generally available; 3.5 Pro remains in partner testing.
- ModelGemini 3.1 Pro
Flash workhorse for agentic/coding/volume; Google confirms 3.5 Pro still testing with partners.
- ModelMoonshot AI
2.8T MoE multimodal agentic model ships on Kimi product and API.
- APIMoonshot AI
OpenAI/Anthropic-compatible API for Kimi K3 and related models.
- APIKimi K3
Moonshot launches Kimi K3 via platform.kimi.ai and consumer Kimi surfaces.
- ResearchKimi K3
Kimi Delta Attention and 16-of-896 expert MoE target long-context efficiency.
- ModelKimi K3
Native multimodal + 1M context
Text, image, and video in one open model with a 1,048,576-token window.
- ModelMeta
Muse Spark 1.1 + Meta Model API
Paid multimodal agentic model via Meta’s first public Model API preview.
- ModelMuse Spark
Major gains in agentic coding, tool/computer use, and multimodal understanding.
- APIMuse Spark
Meta’s first paid model API (OpenAI-compatible) opens Muse Spark to developers.
- ProductMuse Spark
Consumer Meta AI / meta.ai surfaces expose Muse Spark Thinking mode.
- ModelGPT-5.6
GPT-5.6 family (Sol / Terra / Luna)
Capability tiers make routing explicit: Sol for peak, Terra balanced, Luna cost-efficient.
- APIGPT-5.6
API aliases route to Sol by default with Terra/Luna for cost control.
- ModelClaude Sonnet
Sonnet 5 ships as the speed/intelligence balance for most production Claude traffic.
Aggressive intro rates push Sonnet as the default production Claude.
- ModelClaude Fable
Most capable widely released Claude for long-horizon agents and deep reasoning.
- ProductClaude Fable
Limited-access peer of Fable for approved Project Glasswing customers.
- APIClaude Fable
Generally available on Claude API and major cloud partners at premium rates.
- ProductClaude Fable
Effort parameter replaces manual thinking budgets for depth control.
- ProductClaude Fable
Million-token context and large max outputs for long agent threads.
- ProductGPT-5.6
Tool use, structured output, and MCP-style workflows become first-class for agents.
- ProductClaude Opus
Document-heavy analysis and long agent threads become practical.
- ProductClaude Sonnet
1M-token-class context lands on the production tier, not only Opus.
- ModelGPT-5.6
Text, image, and audio understanding with million-token-class context for research agents.
- ModelDeepSeek
V4-Pro and V4-Flash bring 1M context and thinking modes to the DeepSeek API.
- ModelDeepSeek V4
V4-Pro (1.6T/49B) and V4-Flash (284B/13B) ship with 1M context and thinking modes.
- APIDeepSeek V4
deepseek-v4-pro / deepseek-v4-flash
Official API IDs replace assuming historic chat/reasoner names map to V3/R1 weights.
- ModelMuse Spark
Meta Superintelligence Labs ships its first closed multimodal Muse model.
- ModelGemini 3.1 Pro
Current Pro-class pin (gemini-3.1-pro-preview) for hard multimodal reasoning while 3.5 Pro remains in partner testing.
2025
- PlatformClaude Opus
Model Context Protocol tooling broadens Claude’s agent integration surface.
- ModelClaude Haiku
Fast Claude tier with stronger reasoning at economical rates.
- ModelOpenAI
Next major foundation-model generation across chat and API.
- ModelAnthropic
Next-generation Opus and Sonnet strengthen coding, agents, and long-document workflows.
- ModelMeta
Next open multimodal generation for self-hosted stacks.
- ModelLlama 4
Open-weight multimodal Scout / Maverick-class models for research and commercial use.
- ResearchLlama 4
Mixture-of-Experts variants expand efficiency for open multimodal serving.
- ModelQwen3
Multilingual open models spanning chat, reasoning modes, coding, and multimodal variants.
- ProductQwen3
Explicit reasoning modes make Qwen competitive on hard multi-step tasks.
- ModelGoogle DeepMind
Stronger reasoning and long-context performance for Vertex AI and AI Studio.
- ModelGemini 3.1 Pro
Earlier Pro-class Gemini for advanced reasoning and native multimodal workloads.
- BenchmarkDeepSeek R1
GPQA / LiveBench-style comparisons put R1 on the open-reasoning map.
- MilestoneDeepSeek
Industry cost/performance shock
R1 coverage forces hyperscalers and labs to revisit pricing and open-weight strategy.
- ModelDeepSeek
Reasoning model resets expectations for open-weight chain-of-thought performance.
- ResearchDeepSeek R1
RL-trained reasoning model with public traces reshapes open-model expectations.
- Open sourceDeepSeek R1
Weights and distillations enable self-hosted reasoning baselines worldwide.
- APIDeepSeek R1
Hosted API makes R1 accessible without local GPU fleets.
- PlatformClaude Haiku
Production high-volume default
Becomes the default Claude choice when latency and unit economics dominate.
2024
- ModelDeepSeek
MoE open-weight model delivers frontier-competitive quality at lower cost.
- ModelDeepSeek R1
Strong MoE generalist precedes R1 and pairs with it in many stacks.
- ProductClaude Opus
UI and tool-driven agent patterns push Claude into operational automation.
- ProductCursor
Privacy modes, team controls, and org rollouts make Cursor a standard engineering tool.
- ProductGitHub Copilot
Multi-step agent workflows push Copilot closer to AI-native IDE competitors.
- ModelOpenAI
Models optimized for multi-step reasoning before answering.
- ResearchGPT-5.6
Reasoning-model era (o1 lineage)
OpenAI popularizes deliberate reasoning traces that later feed GPT-5 thinking modes.
- PlatformSnowflake
Deepens SQL-native and app patterns for LLMs on enterprise data.
- ModelQwen3
Strong coding and size-ladder releases build the Qwen open ecosystem.
- ProductLangGraph
Visual debugging and graph inspection improve day-to-day agent development.
- PlatformvLLM
Metrics, continuous batching, and model coverage deepen for real traffic.
- PlatformLlamaIndex
Indexing strategies, reranking, and eval loops become standard guidance.
- ModelClaude Sonnet
Sonnet becomes the breakout Claude for coding and everyday agent work.
- PlatformDatabricks
Deepens train/serve/evaluate for enterprise RAG and agent workloads.
- ProductClaude Haiku
Reliable tool calling makes Haiku practical for high-volume agent workers.
- ProductGemini 3.1 Pro
Million-token-class context becomes a Gemini product differentiator.
- PlatformQwen3
Production serving recipes for Qwen solidify across open inference stacks.
- PlatformLangGraph
Managed deployment and ops story expands beyond the open-source library.
- ProductPinecone
Inference / embedding services
Vector DB expands toward integrated embedding and inference workflows.
- PlatformCursor
Teams routinely switch among GPT, Claude, and other models inside the same IDE.
- ProductLangGraph
Interrupt/resume flows make approval steps practical in production agents.
- PlatformQdrant
Distributed production deployments
Clustering and scale-out patterns mature for larger self-hosted estates.
- ModelSnowflake
Signals Snowflake as both a data platform and a model contributor.
- ModelMeta
Major quality jump for open-weight chat and coding models.
- ModelLlama 4
Major open-weight quality leap that sets up the Llama 4 generation.
- PlatformLangChain
Observability, evaluation, and deployment tighten the path from prototype to production.
- ProductvLLM
Serving many fine-tuned adapters efficiently becomes a production requirement.
- ModelDatabricks
Signals Databricks as both a platform and a model contributor in open LLM space.
- PlatformNVIDIA
Next-generation accelerated computing platform for training and inference at AI-factory scale.
- PlatformNVIDIA
Microservices for deploying optimized inference in enterprise AI factories.
- ModelAnthropic
Opus, Sonnet, and Haiku establish a clear quality–latency–cost ladder for production.
- ModelClaude Opus
Establishes Opus as Anthropic’s peak-capability brand in the Claude family.
- ModelClaude Sonnet
Defines the mid-tier Claude brand for balanced cost and capability.
- ModelClaude Haiku
Defines Haiku as Anthropic’s speed and cost brand in the Claude family.
- ModelClaude Haiku
Image understanding lands on the fast tier for multimodal pipelines.
- PlatformLangGraph
Checkpoints & persistence mature
Durable state, retries, and long-running workflows become first-class.
- PlatformLlamaIndex
Framework grows from pure RAG into broader workflow and agent patterns.
- ModelMistral AI
Flagship commercial model strengthens La Plateforme for enterprise APIs.
- ProductMistral AI
Consumer/business chat surface for Mistral models.
- ModelGemini 3.1 Pro
Native multimodal + long context establishes Gemini’s modern identity.
- PlatformWeaviate
Closer integration with generative search patterns for end-to-end RAG apps.
- PlatformGitHub Copilot
Org knowledge, admin controls, and enterprise packaging deepen GitHub-native AI.
- Open sourceLangGraph
Graph-based runtime for durable, stateful LangChain agents.
- ProductLangChain
Stateful graph runtime complements LangChain’s higher-level abstractions.
- APIDeepSeek
OpenAI-compatible API makes DeepSeek easy to try in existing clients.
- PlatformQwen3
Hub becomes the default discovery path for Qwen weights and demos.
- PlatformPinecone
Packaging shifts toward more elastic, usage-based vector infrastructure.
- ProductCursor
Agentic coding features expand
Multi-file edits and agent-style workflows deepen IDE autonomy.
- PlatformvLLM
Ecosystem default for open serving
Becomes a standard choice next to TGI and TensorRT-LLM in many stacks.
2023
- ModelMicrosoft
Shows competitive quality from small, efficient models for edge and cost-sensitive use.
- ModelMistral AI
Sparse MoE open model becomes a default efficient alternative to dense LLMs.
- ModelGoogle DeepMind
Natively multimodal foundation model family launches across consumer and cloud surfaces.
- ModelGemini 3.1 Pro
Google’s unified multimodal foundation-model brand debuts.
- PlatformOpenAI
Platform primitives for tool use, retrieval, and agent-style apps.
- PlatformSnowflake
Brings LLM functions and AI services into the governed Data Cloud.
- ProductQdrant
Sparse + dense retrieval patterns expand beyond pure vector similarity.
- ModelMistral AI
Small open model punches above its size and accelerates European open LLM adoption.
- ProductCursor
Repo-aware edits across files shift Cursor from autocomplete to agentic coding.
- APIvLLM
Drop-in API compatibility accelerates migration from hosted APIs to self-host.
- ProductLangChain
Expression Language / LCEL era
Composable runnables become the preferred way to build chains and pipelines.
- ProductLlamaIndex
Evaluation & observability focus
RAG quality tooling expands beyond connectors and indexes.
- Open sourceMeta
Openly available weights for commercial and research use.
- Open sourceLlama 4
Open weights become broadly usable for commercial products under Meta’s terms.
- ModelAnthropic
Longer context and stronger conversational performance for enterprise assistants.
- FoundedDeepSeek
Hangzhou lab focuses on efficient, high-capability open models.
- AcquisitionDatabricks
Brings training and generative AI tooling into Mosaic AI on the Lakehouse.
- ProductPinecone
Hybrid / metadata retrieval focus
Sparse-dense and metadata filtering become table stakes for enterprise RAG.
- Open sourcevLLM
Open serving engine makes high-throughput LLM inference widely accessible.
- ProductMilvus
GPU-accelerated indexing deepens Milvus’s large-scale search positioning.
- MilestoneNVIDIA
GPUs become the default substrate for training and inference at global scale.
- ProductWeaviate
Keyword + vector retrieval becomes a primary reason teams evaluate Weaviate.
- LeadershipGoogle DeepMind
Google Brain and DeepMind combine—research, models, and product under one AI organization.
- FoundedMistral AI
Paris-based lab founded by former DeepMind / Meta researchers.
- LeadershipGemini 3.1 Pro
Research and model product organizations consolidate under one AI org.
- ProductLlamaIndex
Ingestion from docs, APIs, and SaaS sources becomes a core product strength.
- ProductGitHub Copilot
Chat moves Copilot beyond autocomplete into conversational coding help.
- ModelOpenAI
Multimodal foundation model with stronger reasoning and coding.
- PlatformLangChain
Model, vector-store, and tool connectors make LangChain the glue layer for LLM stacks.
- ProductQdrant
Rich metadata filters become a defining reason teams pick Qdrant for RAG.
- ProductCursor
AI-native fork of VS Code becomes a default coding environment for many teams.
- ModelMeta
Research weights that sparked the open-model wave.
- ResearchLlama 4
Research release that sparks the modern open-model wave.
- PartnershipMicrosoft
Deepens Azure OpenAI as the enterprise channel for ChatGPT-class models.
- PlatformHugging Face
Models, datasets, and Spaces concentrate open AI collaboration and distribution.
- FoundedMoonshot AI
Beijing lab builds the Kimi assistant and frontier MoE models.
2022
- ProductOpenAI
Consumer chat product that brought LLMs into mainstream use.
- Open sourceLlamaIndex
LlamaIndex (GPT Index) adoption
Becomes a leading open framework for connecting LLMs to private data.
- PlatformHugging Face
Managed production inference for Hub models without self-hosting ops.
- Open sourceLangChain
Becomes the default open framework for composing LLM apps.
- ProductGitHub Copilot
Generally available as a paid coding assistant for individuals and teams.
- PlatformPinecone
Production customers standardize on managed indexes for semantic search and RAG.
- PlatformQdrant
Managed option alongside self-hosted deployments.
- ModelGoogle DeepMind
Pathways Language Model showcases large-scale Google LLM research before Gemini.
- PlatformNVIDIA
Defines the GPU generation that trains and serves most frontier models of the mid-2020s.
- PlatformWeaviate
Managed offering expands beyond self-hosted deployments.
- PlatformMilvus
Managed Milvus becomes the enterprise path without self-host ops.
2021
- PlatformHugging Face
Turns the Hub into a place to demo and share interactive ML apps.
- ProductMicrosoft
Coding assistant brings LLM copilots into everyday developer workflows.
- ProductGitHub Copilot
GitHub Copilot technical preview
Inline AI completions bring LLM coding assistants into mainstream IDEs.
- ProductWeaviate
Pluggable modules make embedding and enrichment part of the DB workflow.
- Open sourceMilvus
CNCF sandbox / cloud-native path
Milvus joins the cloud-native ecosystem path for large-scale vector infra.
- ProductPinecone
Managed vector search becomes mainstream for production RAG.
- FoundedAnthropic
Safety-focused lab founded by former OpenAI researchers; Constitutional AI becomes a defining idea.
- Open sourceQdrant
Rust vector engine gains adoption for filtered similarity search.
- PlatformMilvus
Milvus 2.x distributed architecture
Cloud-native redesign enables billion-scale indexes and separation of storage/compute.
2020
- FundingSnowflake
Public listing accelerates enterprise Data Cloud adoption.
- APIOpenAI
Developers gain API access to GPT models for production apps.
- Open sourceHugging Face
Standardizes how the community loads, shares, and versions training data.
- PlatformDatabricks
Unifies data warehousing and data lakes as the substrate for analytics and AI.
2019
- Open sourceMilvus
Zilliz releases Milvus as an open vector database for similarity search.
- PartnershipOpenAI
Microsoft strategic partnership
Microsoft invests $1B and becomes a primary cloud partner.
- PartnershipMicrosoft
Multi-year investment and Azure exclusivity set up GPT distribution.
- Open sourceWeaviate
Weaviate open-source vector DB
Early open vector database with a modular architecture.
2018
- Open sourceHugging Face
Becomes the default open toolkit for pretrained NLP models.
- Open sourceDatabricks
Becomes the default open toolkit for experiment tracking and model registry.
2017
- ResearchGoogle DeepMind
Transformer architecture published
“Attention Is All You Need” (Google Research) becomes the foundation of modern LLMs.
- PlatformNVIDIA
Inference optimization stack becomes central to production deep learning serving.
2016
- Open sourceMeta
Becomes the dominant research framework and later the industry standard for training modern neural nets.
- ResearchGoogle DeepMind
Landmark demonstration of deep reinforcement learning on a grandmaster-level game.
- FoundedHugging Face
Started as a chat app company; pivoted into open ML tooling.
2015
- FoundedOpenAI
Founded as a nonprofit AI research lab in San Francisco.
2013
- FoundedDatabricks
Created by the creators of Apache Spark to commercialize big data and ML platforms.
2012
- FoundedSnowflake
Cloud data platform company founded to separate storage and compute.
2010
- FoundedGoogle DeepMind
London research lab focused on general-purpose AI systems.
- PlatformMicrosoft
Azure becomes the cloud substrate for later AI services and OpenAI hosting.
2006
- PlatformNVIDIA
GPU programming platform that later underpins modern AI training and inference.
1975
- FoundedMicrosoft
Started as a microcomputer software company in Albuquerque.