Updated

AI model release tracker

Major AI model releases from 2022 to today.

Provider
Year

201 releases shown

September 2026

11 releases
  1. OpenAI
    ModelsLatest update

    GPT-6 Sol

    OpenAI released GPT-6 Sol for coding and professional work, with API pricing of $2 per million input tokens and $10 per million output tokens. Access is rolling out in ChatGPT Work and Codex; it is not yet available in Chat.

    For buildersEvaluate a coding or research workflow on your own tasks, comparing quality and total cost with the model you currently use.

  2. OpenAI
    Models

    GPT-6 Luna

    OpenAI released GPT-6 Luna alongside Sol as a lower-cost model, with API pricing of $0.10 per million input tokens and $0.50 per million output tokens. Free and Go users can access Luna in the desktop app.

    For buildersTest a high-volume workflow such as classifying support requests or extracting structured fields from incoming documents.

  3. xAI
    Models

    Grok Voice Transcribe 2.0

    xAI released Grok Voice Transcribe 2.0, a speech-to-text model available through its Speech-to-Text API for batch and streaming transcription. Company-listed pricing is unchanged from version 1.0: $0.10 per hour of audio for batch and $0.20 per hour for streaming.

    For buildersSend a noisy support call or meeting recording through the Speech-to-Text API and compare speaker labels, timestamps, and formatted emails or phone numbers against your current transcriber.

  4. Google
    Models

    Gemini 3.8 Live

    Google introduced two live dialogue models, with a separate Extended Thinking option for more complex voice tasks.

    For buildersPrototype a voice onboarding assistant that answers questions and walks new users through a multi-step setup.

  5. TypeSafe AI
    Models

    Jev (early access)

    TypeSafe AI released Jev, its first public System One model, in early access. It evaluates supplied state and returns typed decisions with probabilities instead of generating free-form text.

    For buildersTry classifying support tickets or routing requests using predefined choices, with a review path for uncertain decisions.

  6. DeepSeek
    Models

    DeepSeek V4.1 Flash

    DeepSeek released a new Flash model with native visual understanding and improvements to agent workflows.

    For buildersTest screenshots and document images alongside the text-only tasks in your evaluation set.

  7. OpenAI
    Models

    GPT Image 2.5

    OpenAI introduced an image model with sharper details, more precise editing, and faster generation.

    For buildersUse repeated edits to evaluate whether a product image keeps its key details intact.

  8. OpenAI
    Models

    GPT-6 Astra

    OpenAI released GPT-6 Astra, a new generation of model for demanding reasoning and computer-based work.

    For buildersEvaluate a difficult multi-step task against explicit acceptance criteria before changing your default model.

  9. Google
    Models

    Gemini 3.8 Flash

    Google announced 3.8 Flash with improvements to coding, reasoning, and multi-step agent workflows, alongside a separate Cyber variant.

    For buildersEvaluate it on your app's real tasks: extracting data, routing support requests, or carrying out a sequence of tool calls.

August 2026

4 releases
  1. Google
    Models

    Gemini 3.5 Transcribe

    Google launched a speech-to-text model designed to turn raw audio into polished, formatted text through the Gemini API.

    For buildersAdd voice notes to a CRM or turn recorded customer interviews into searchable transcripts for your team.

  2. xAI
    Models

    Grok 4.6

    Grok 4.6 introduced improvements for longer-running agents and interactive visual work.

    For buildersTest an agent on both the implementation and the rendered result of an interface change.

July 2026

9 releases
  1. MiniMax
    Models

    MiniMax H3

    MiniMax launched a multimodal generation model producing video with native stereo audio.

    For buildersPrototype a scene with coordinated visual action and environmental sound.

  2. ByteDance
    Models

    Seedance 2.5

    ByteDance introduced longer video generation with expanded reference and editing capabilities.

    For buildersTry a continuous sequence with a clear beginning, middle, and end.

  3. Anthropic
    Models

    Claude Opus 5

    Opus 5 launched with improvements in coding and knowledge work, positioned below Fable in price.

    For buildersBenchmark the full workflow, including tool calls and retries, when comparing model costs.

  4. Alibaba
    Models

    Qwen-Image-3.0

    Qwen introduced the third generation of its foundational image-generation model family.

    For buildersCompare typography, instruction following, and edit consistency on the same creative brief.

  5. xAI
    Models

    Grok 4.5

    Grok 4.5 launched for coding, agentic tasks, and knowledge work, including API access.

    For buildersUse realistic engineering tasks to compare completion quality and total token use.

  6. Moonshot AI
    Models

    Kimi K3

    Kimi K3 launched with native vision and a one-million-token context window. Hosted access came first, with model weights scheduled afterward.

    For buildersTest a large-repository coding task with screenshots and written requirements together.

  7. OpenAI
    Models

    GPT-Live

    OpenAI introduced a new generation of voice models for more natural spoken interaction.

    For buildersTest interruptions and turn-taking as well as transcription accuracy in a voice assistant.

  8. ByteDance
    Models

    Seedream 5.0 Pro

    ByteDance introduced an image generation model emphasizing reasoning and design-oriented creation.

    For buildersEvaluate a complete visual layout against a written design brief.

June 2026

10 releases
  1. Anthropic
    Models

    Claude Sonnet 5

    Sonnet 5 launched with improved planning, tool use, coding, and professional work at the Sonnet price tier.

    For buildersMeasure whether a smaller model can complete the same workflow reliably before moving to a higher tier.

  2. Google
    Models

    Nano Banana 2 Lite

    Google introduced a faster, lower-cost Gemini image model for generation and editing.

    For buildersUse a lower-cost image model to explore many directions before refining a final asset.

  3. Google
    Models

    DiffusionGemma

    Google introduced a Gemma-based model using diffusion techniques to accelerate text generation.

    For buildersMeasure both output quality and interactive latency when comparing generation architectures.

  4. Google
    Models

    Gemma 4 12B

    A 12B multimodal model joined Gemma 4, positioned for laptop deployment.

    For buildersA laptop-based assistant can process local documents without sending every input to a hosted model.

  5. MiniMax
    Models

    MiniMax M3

    M3 introduced native multimodality and support for a one-million-token context window alongside coding improvements.

    For buildersEvaluate visual inputs and a large codebase in the same workflow.

May 2026

3 releases
  1. Mistral
    Models

    Mistral Medium 3.5

    Mistral introduced a 128B open-weight model for longer coding and productivity workflows.

    For buildersTest whether an agent maintains context and follows the brief across a full implementation task.

April 2026

10 releases
  1. OpenAI
    Models

    GPT-5.5

    OpenAI released GPT-5.5 for coding, research, data analysis, and work across computer tools. API availability followed on April 24.

    For buildersTurn a messy brief and spreadsheet into a small internal tool, then check the calculations and behavior against the original data.

  2. Moonshot AI
    Models

    Kimi K2.6

    Moonshot released an updated open model focused on coding and sustained agent workflows.

    For buildersEvaluate changes spanning multiple files rather than judging only single code snippets.

  3. Anthropic
    Models

    Claude Opus 4.7

    Anthropic introduced an Opus update with stronger coding, vision, and complex multi-step task performance.

    For buildersFor visual app work, test whether the model follows both functional requirements and design references.

  4. OpenAI
    Models

    GPT-Rosalind (research preview)

    A specialized reasoning model for biology and drug-discovery research launched through a restricted research preview.

    For buildersSpecialized models require domain-specific evaluation and an appropriate access program.

  5. Google
    Models

    Gemini 3.1 Flash TTS

    A new text-to-speech model expanded control over expressive generated speech.

    For buildersPrototype spoken instructions and compare clarity across different speaking styles.

  6. Google
    Models

    Gemma 4

    Gemma 4 introduced open models focused on reasoning and agent workflows across different hardware sizes.

    For buildersTest an offline assistant on the device class your users actually own.

March 2026

8 releases
  1. Google
    Models

    Gemini 3.1 Flash Live

    Google introduced an audio model for faster live conversations with longer conversational context.

    For buildersTest a spoken support flow across interruptions and follow-up questions.

  2. MiniMax
    Models

    MiniMax M2.7

    MiniMax introduced a model focused on software engineering, professional tasks, and coordinated agent work.

    For buildersTest an engineering task that requires planning, implementation, and review.

  3. OpenAI
    Models

    GPT-5.4 mini and nano

    OpenAI added smaller GPT-5.4 models aimed at fast, cost-sensitive workloads.

    For buildersEvaluate separate models for routine classification and more demanding coding requests.

  4. Mistral
    Models

    Mistral Small 4

    Mistral Small 4 combined reasoning, multimodal understanding, and agentic coding in one model.

    For buildersA unified model can simplify an app that mixes screenshots, text, and code.

  5. OpenAI
    Models

    GPT-5.4 and GPT-5.4 Pro

    GPT-5.4 brought reasoning, coding, and professional tool workflows together, with a Pro variant for harder tasks.

    For buildersCompare the model on an end-to-end workflow spanning code, documents, and data.

February 2026

12 releases
  1. Google
    Models

    Nano Banana 2

    Google introduced an image model combining capabilities from Nano Banana Pro with Flash-class speed.

    For buildersCompare image quality and turnaround time for batches of marketing variations.

  2. Google
    Models

    Gemini 3.1 Pro

    Google introduced Gemini 3.1 Pro for more complex reasoning and multimodal tasks.

    For buildersEvaluate the model on realistic combinations of documents, images, and instructions.

  3. Alibaba
    Models

    Qwen3.5

    Qwen3.5 introduced an open-weight native vision-language model with reasoning and agent capabilities.

    For buildersEvaluate multilingual and visual inputs together when your users work across languages.

  4. MiniMax
    Models

    MiniMax M2.5

    MiniMax released a model focused on coding, tool use, and practical productivity tasks.

    For buildersMeasure success on complete workflows with a clear finished output.

  5. ByteDance
    Models

    Seedance 2.0

    ByteDance released a multimodal video model accepting text, image, audio, and video references.

    For buildersTry guiding a video with a reference clip and a separate audio track.

  6. Anthropic
    Models

    Claude Opus 4.6

    Opus 4.6 improved coding and longer agent tasks, with a one-million-token context window in beta.

    For buildersLarge-codebase work needs checks for both retrieval accuracy and correctness of the final change.

  7. OpenAI
    Models

    GPT-5.3-Codex

    OpenAI introduced a coding model combining software-engineering ability with broader reasoning and computer-based work.

    For buildersGive a coding agent a reproducible issue and verify both its fix and its tests.

  8. Kuaishou
    Models

    Kling 3.0 model family

    Kuaishou launched Video 3.0, Video 3.0 Omni, Image 3.0, and Image 3.0 Omni with stronger consistency and narrative control.

    For buildersTest whether references preserve characters across a multi-shot sequence.

January 2026

1 release
  1. Moonshot AI
    Models

    Kimi K2.5

    Kimi K2.5 combined visual understanding, coding, and agent capabilities.

    For buildersTry translating a screenshot into a working interface and reviewing the result visually.

December 2025

13 releases
  1. MiniMax
    Models

    MiniMax M2.1

    M2.1 expanded multilingual programming and development capabilities.

    For buildersEvaluate code quality across the languages used in your application.

  2. OpenAI
    Models

    GPT-5.2-Codex

    GPT-5.2-Codex launched in Codex for paid ChatGPT users, with API access planned afterward.

    For buildersTest a migration or refactor that requires maintaining context across a long session.

  3. Google
    Models

    Gemini 3 Flash

    Google expanded Gemini 3 with a faster, lower-cost Flash model for reasoning and multimodal tasks.

    For buildersUse a fast model for interactive tasks where waiting time affects the user experience.

  4. OpenAI
    Models

    GPT Image 1.5

    OpenAI released an image model with faster generation and more precise edits that preserve existing details.

    For buildersEvaluate a series of edits for consistency, rather than judging only the first image.

  5. NVIDIA
    Models

    Nemotron 3 Nano

    NVIDIA released Nano as the first available model in its Nemotron 3 family, with larger models announced for later release.

    For buildersEvaluate an open model for a focused agent workload on your own infrastructure.

  6. OpenAI
    Models

    GPT-5.2

    The GPT-5.2 series launched for professional work, including documents, spreadsheets, coding, and longer tool workflows.

    For buildersTest complete deliverables, such as an editable report and its calculations, rather than isolated answers.

  7. Amazon
    Models

    Amazon Nova 2 Sonic

    Amazon released its second-generation real-time conversational speech model.

    For buildersPrototype a spoken assistant and measure its handling of interruptions and follow-ups.

  8. DeepSeek
    Models

    DeepSeek V3.2 and V3.2-Speciale

    DeepSeek released reasoning-first models for agent workflows, with a higher-reasoning Speciale variant on a temporary endpoint.

    For buildersDifferent reasoning tiers should be compared on successful task completion and cost.

  9. Runway
    Models

    Runway Gen-4.5

    Runway announced a video model with improved motion quality, prompt adherence, and visual fidelity.

    For buildersCompare complex motion prompts and multi-part scene descriptions.

November 2025

10 releases
  1. Black Forest Labs
    Models

    FLUX.2

    Black Forest Labs introduced an image model for generation and editing with multiple visual references.

    For buildersUse reference images to test product, character, and style consistency across a set of assets.

  2. Anthropic
    Models

    Claude Opus 4.5

    Opus 4.5 launched with stronger coding, agent, and computer-use capabilities at a lower Opus price.

    For buildersCompare completion quality and the number of retries, not just a model’s per-token price.

  3. OpenAI
    Models

    GPT-5.1-Codex-Max

    OpenAI introduced a coding model designed for longer engineering tasks and more efficient reasoning.

    For buildersCompare sustained debugging across several iterations.

  4. xAI
    Models

    Grok 4.1 Fast

    xAI released a fast model for tool-using agents, alongside its Agent Tools API.

    For buildersMeasure the latency of the whole tool workflow, not just the first generated token.

  5. Google
    Models

    Gemini 3 Pro (preview)

    Gemini 3 Pro launched with improvements in reasoning, multimodal understanding, and agentic coding.

    For buildersTest visual reasoning and code generation together when building interfaces from references.

  6. xAI
    Models

    Grok 4.1

    An updated Grok model focused on conversational quality, nuanced instructions, and collaborative interactions.

    For buildersEvaluate tone and instruction following using the conversations your users actually have.

  7. ElevenLabs
    Models

    Scribe v2 Realtime

    ElevenLabs introduced a streaming speech-to-text model for low-latency transcription.

    For buildersEvaluate live captions with background noise, interruptions, and specialized vocabulary.

  8. Moonshot AI
    Models

    Kimi K2 Thinking

    Moonshot introduced a thinking model designed for reasoning with repeated tool use.

    For buildersTest multi-step research tasks that require gathering and reconciling information.

October 2025

3 releases
  1. MiniMax
    Models

    MiniMax M2

    MiniMax released an open model focused on coding and agent tool use.

    For buildersCompare repository tasks that require both code changes and command execution.

  2. Google
    Models

    Veo 3.1

    Veo 3.1 improved realism, audio, and creative control in video generation.

    For buildersEvaluate whether a generated sequence follows the intended story and preserves its subjects.

September 2025

7 releases
  1. OpenAI
    Models

    Sora 2

    Sora 2 introduced video with synchronized dialogue and sound effects. The Sora product was later discontinued in April 2026.

    For buildersA historical example of integrating video and audio generation into one model.

  2. Anthropic
    Models

    Claude Sonnet 4.5

    Sonnet 4.5 launched with advances in coding, computer use, and complex agent workflows.

    For buildersEvaluate long-running agents on completion quality, recovery from errors, and total cost.

  3. DeepSeek
    Models

    DeepSeek V3.2-Exp

    An experimental model introduced DeepSeek Sparse Attention for more efficient long-context processing.

    For buildersLong-context efficiency matters for repeated analysis of large documents or codebases.

  4. Luma AI
    Models

    Luma Ray3

    Luma unveiled a reasoning-driven video model with creative controls and an HDR production pipeline.

    For buildersExplore generated shots that need controlled transitions and consistent visual direction.

  5. OpenAI
    Models

    GPT-5-Codex

    OpenAI released a GPT-5 model specialized for agentic software development.

    For buildersEvaluate a complete repository change with implementation and test execution.

  6. Moonshot AI
    Models

    Kimi K2 0905

    An updated Kimi K2 improved agentic coding and expanded its context window to 256K.

    For buildersUse a multi-file app change to assess whether longer context translates into better implementation.

August 2025

5 releases
  1. DeepSeek
    Models

    DeepSeek V3.1

    DeepSeek introduced a hybrid model combining thinking and non-thinking modes, with stronger agent capabilities.

    For buildersUse reasoning selectively for the steps that require planning or difficult decisions.

  2. OpenAI
    Models

    GPT-5

    GPT-5 introduced a system combining fast responses and deeper reasoning, with improvements across coding and multimodal tasks.

    For buildersEvaluate both quick responses and longer reasoning on the same realistic set of app tasks.

  3. Anthropic
    Models

    Claude Opus 4.1

    An Opus update improved agentic coding, reasoning, and detail tracking during research and analysis.

    For buildersMulti-file refactoring is a useful test of whether a model makes precise changes.

July 2025

3 releases
  1. Moonshot AI
    Models

    Kimi K2

    Moonshot released open-weight mixture-of-experts models focused on knowledge, coding, and agentic tool use.

    For buildersEvaluate how a model chooses and sequences tools, as well as the final answer.

  2. xAI
    Models

    Grok 4 and Grok 4 Heavy

    Grok 4 launched with reasoning, native tool use, and real-time search integration; Heavy provided a higher-compute option.

    For buildersGround current-information tasks in retrieved sources and verify the resulting citations.

June 2025

4 releases
  1. Alibaba
    Models

    Qwen VLo (preview)

    Alibaba previewed a model combining visual understanding with image generation and editing.

    For buildersExplore edits described through both a reference image and written instructions.

  2. Mistral
    Models

    Magistral

    Mistral introduced its first reasoning-model family, emphasizing multilingual reasoning.

    For buildersEvaluate complex reasoning in the languages your app supports.

  3. ElevenLabs
    Models

    Eleven v3 (alpha)

    An expressive text-to-speech model debuted in alpha with audio tags and multi-speaker dialogue.

    For buildersDirect delivery and emotion for a narrated lesson, then listen for clarity and consistency.

May 2025

7 releases
  1. DeepSeek
    Models

    DeepSeek-R1-0528

    An updated R1 checkpoint improved reasoning and added support for structured outputs and function calling.

    For buildersEvaluate a reasoning workflow that must return valid structured data.

  2. Mistral
    Models

    Devstral

    Mistral and All Hands AI released an open model for agentic software-engineering tasks.

    For buildersRepository navigation and tool use matter when a model must fix a real software issue.

  3. Google
    Models

    Imagen 4

    Google introduced its next image-generation model with improvements in visual detail and text rendering.

    For buildersCompare image models on the text, layout, and brand constraints in your actual creative brief.

  4. Google
    Models

    Veo 3

    Google introduced Veo 3 for video generation with native audio.

    For buildersPrototype short scenes with speech and environmental sound before committing to a full production.

  5. Mistral
    Models

    Mistral Medium 3

    Mistral introduced a model positioned to balance capability, efficiency, and enterprise deployment needs.

    For buildersInclude latency and hosting requirements in your model comparison.

April 2025

7 releases
  1. Alibaba
    Models

    Qwen3

    Qwen3 introduced open-weight dense and mixture-of-experts models with thinking and non-thinking modes.

    For buildersTest whether extra reasoning improves accuracy enough to justify added latency.

  2. OpenAI
    Models

    OpenAI o3 and o4-mini

    The new reasoning models combined extended thinking with tool use and visual analysis.

    For buildersA reasoning workflow can inspect an image, run calculations, and gather evidence before returning an answer.

  3. OpenAI
    Models

    GPT-4.1, mini, and nano

    Three API models launched with improved coding and instruction following, and context windows up to one million tokens.

    For buildersDifferent model sizes let an app route simple extraction and complex code work to different cost tiers.

  4. Meta
    Models

    Llama 4 Scout and Maverick

    Meta released two open-weight, natively multimodal mixture-of-experts models. Behemoth was described but not released.

    For buildersMultimodal open weights offer another route to custom document and image-understanding systems.

March 2025

6 releases
  1. Runway
    Models

    Runway Gen-4

    Runway introduced a video model focused on consistent characters, objects, and environments across shots.

    For buildersBuild a short storyboard using the same subject across several scenes.

  2. OpenAI
    Models

    GPT-4o image generation

    OpenAI introduced native image generation with improved prompt following, text rendering, and image-based editing.

    For buildersCreate a visual from a brief, then refine it using the previous output as context.

  3. Cohere
    Models

    Command A

    Cohere released Command A for enterprise language and agent workflows.

    For buildersEvaluate grounded answers and tool use against internal business documents.

  4. Google
    Models

    Gemma 3

    Google released a new open-model family with multilingual and multimodal capabilities.

    For buildersOn-device document and image understanding can reduce reliance on hosted APIs.

February 2025

4 releases
  1. OpenAI
    Models

    GPT-4.5 (research preview)

    OpenAI released a research preview emphasizing conversational quality, broad knowledge, and instruction following.

    For buildersThis release illustrates that conversational quality and explicit reasoning are different model strengths.

  2. ElevenLabs
    Models

    Scribe

    ElevenLabs released its first speech-to-text model, with multilingual transcription, timestamps, and speaker labels.

    For buildersCreate searchable meeting or interview transcripts with links back to the original recording.

January 2025

3 releases
  1. OpenAI
    Models

    OpenAI o3-mini

    A small reasoning model arrived with a focus on science, mathematics, and coding.

    For buildersReasoning effort can be adjusted to match the difficulty of a technical task.

  2. DeepSeek
    Models

    DeepSeek-R1

    DeepSeek released its reasoning model and six smaller distilled models with open weights.

    For buildersCompare reasoning quality across model sizes before choosing a deployment approach.

  3. Luma AI
    Models

    Luma Ray2

    Luma launched Ray2 on a new multimodal architecture for video generation.

    For buildersPrototype short scenes and check whether movement remains plausible throughout the clip.

December 2024

7 releases
  1. DeepSeek
    Models

    DeepSeek-V3

    DeepSeek released a large mixture-of-experts language model with open weights.

    For buildersA milestone for comparing hosted APIs with independently deployable language models.

  2. Google
    Models

    Veo 2 (preview)

    Google introduced Veo 2 with improved realism and cinematic control through a limited VideoFX rollout.

    For buildersCompare motion and camera behavior over a series of generated clips.

  3. Microsoft
    Models

    Phi-4

    Microsoft introduced a 14-billion-parameter language model emphasizing reasoning, initially through Azure AI Foundry.

    For buildersCompare a compact model on focused mathematical and structured reasoning tasks.

  4. Google
    Models

    Gemini 2.0 Flash Experimental

    Gemini 2.0 began with an experimental Flash release and expanded native multimodal and tool-use capabilities.

    For buildersA reference point for moving from chat responses to models that participate in tool workflows.

  5. OpenAI
    Models

    Sora Turbo

    A faster version of Sora became available for video generation after the earlier research preview.

    For buildersThis historical release marks the move from a research demonstration to a usable video model.

  6. OpenAI
    Models

    OpenAI o1

    The full o1 model arrived in ChatGPT alongside a higher-compute o1 pro mode.

    For buildersCompare reasoning depth against latency when designing a multi-step assistant.

October 2024

2 releases

September 2024

4 releases
  1. Meta
    Models

    Llama 3.2

    Meta released small text models for edge devices and larger models with vision capabilities.

    For buildersMatch model size and modality to the device and task instead of defaulting to the largest model.

  2. Alibaba
    Models

    Qwen2.5

    Alibaba’s Qwen team released a broad family of open-weight language models, with coding and mathematics variants.

    For buildersChoose the smallest model that meets the requirements of your language or code task.

  3. OpenAI
    Models

    OpenAI o1-mini

    OpenAI released a smaller reasoning model focused on tasks such as coding and mathematics.

    For buildersA lower-cost reasoning tier offered another option for technical tasks.

  4. OpenAI
    Models

    OpenAI o1-preview

    An early reasoning model became available in ChatGPT and to eligible API users, using more computation before answering.

    For buildersReasoning models introduced a new tradeoff between response speed and work on difficult problems.

August 2024

2 releases
  1. Black Forest Labs
    Models

    FLUX.1

    Black Forest Labs introduced its first FLUX image generation models.

    For buildersCompare prompt adherence and visual consistency across a small set of brand assets.

July 2024

4 releases
  1. Mistral
    Models

    Mistral Large 2

    The second Mistral Large generation introduced 128K context and stronger multilingual and coding capabilities.

    For buildersEvaluate retrieval across long inputs before relying on a large advertised context window.

  2. Meta
    Models

    Llama 3.1

    Llama 3.1 expanded context to 128K, added language support, and introduced the 405B model.

    For buildersCompare the infrastructure costs of larger open-weight models with a managed API.

  3. OpenAI
    Models

    GPT-4o mini

    A smaller, lower-cost model launched with text and vision support.

    For buildersSmall models can make classification, extraction, and repeated background tasks more economical.

June 2024

4 releases
  1. Google
    Models

    Gemma 2

    Gemma 2 launched in 9B and 27B sizes with improvements to model quality and inference efficiency.

    For buildersSelect a model size against your available memory and expected request volume.

  2. Anthropic
    Models

    Claude 3.5 Sonnet

    The first Claude 3.5 model introduced improvements in coding, visual reasoning, and speed.

    For buildersA milestone for generating interfaces and reviewing screenshots alongside code.

  3. Alibaba
    Models

    Qwen2

    Alibaba released five Qwen2 model sizes, including dense and mixture-of-experts models.

    For buildersUse the size range to compare hosting requirements for a multilingual assistant.

May 2024

5 releases
  1. Mistral
    Models

    Codestral

    Mistral introduced its first dedicated code model, trained across more than 80 programming languages.

    For buildersCode completion and repository-level agents are different tasks and should be evaluated separately.

  2. Google
    Models

    Gemini 1.5 Flash

    Google introduced a lighter Gemini model designed for speed and efficiency at scale.

    For buildersLatency-sensitive summarization and extraction often benefit from a smaller model tier.

  3. Google
    Models

    Veo (limited preview)

    Google introduced Veo for high-definition video generation with limited creator access and a waitlist.

    For buildersExplore how camera and cinematic language translate into generated video.

  4. OpenAI
    Models

    GPT-4o

    OpenAI introduced its omni model for text, vision, and audio. Text and image inputs launched first; other modalities followed.

    For buildersA useful reference point for the shift toward assistants that understand screenshots and spoken input.

April 2024

4 releases
  1. Microsoft
    Models

    Phi-3 mini

    Microsoft released Phi-3 mini, a compact language model designed for constrained compute environments.

    For buildersTest whether a small model can handle a focused task close to the user device.

  2. Meta
    Models

    Llama 3

    Meta introduced the next generation of Llama with improved language and instruction-following capabilities.

    For buildersUse a consistent evaluation set when comparing hosted and self-hosted models.

  3. Cohere
    Models

    Command R+

    Cohere introduced a larger RAG-oriented model for enterprise workloads.

    For buildersCompare answer grounding and tool-use reliability on representative business documents.

March 2024

4 releases
  1. xAI
    Models

    Grok-1.5 (announcement)

    xAI announced Grok-1.5 with improved reasoning and a 128K context window ahead of early-user rollout.

    For buildersCompare long-document tasks with the preceding Grok generation.

  2. Anthropic
    Models

    Claude 3 Haiku

    Haiku joined the Claude 3 family as its fast, lower-cost model.

    For buildersA useful historical example of routing straightforward requests to a smaller model.

  3. Cohere
    Models

    Command R

    Cohere introduced Command R for retrieval-augmented generation in production applications.

    For buildersBuild answers around retrieved documents and verify that citations support the claims.

  4. Anthropic
    Models

    Claude 3 Opus and Sonnet

    Opus and Sonnet launched as the first available Claude 3 models, with vision support and different capability tiers.

    For buildersCompare a larger and smaller model on the same document-understanding tasks.

February 2024

3 releases
  1. Mistral
    Models

    Mistral Large

    Mistral introduced its flagship model for multilingual reasoning, text transformation, and code generation.

    For buildersTest multilingual prompts using native-language examples from your actual users.

  2. Google
    Models

    Gemma 2B and 7B

    Google introduced Gemma, a family of lightweight open models related to its Gemini research.

    For buildersSmaller downloadable models can support experimentation on hardware you control.

  3. Google
    Models

    Gemini 1.5 Pro (preview)

    Google introduced a Gemini 1.5 Pro preview with a breakthrough in long-context multimodal understanding.

    For buildersLarge document and video collections became practical inputs for a single model request.

December 2023

3 releases
  1. Mistral
    Models

    Mixtral 8x7B

    Mistral announced an open-weight sparse mixture-of-experts model under Apache 2.0.

    For buildersMixture-of-experts models balance total model capacity against the computation used per token.

  2. Google
    Models

    Gemini 1.0 Pro and Nano

    Google introduced the Gemini family. Pro and Nano began rolling out first, while Ultra was announced for a later release.

    For buildersThe family established separate model sizes for cloud applications and on-device tasks.

November 2023

3 releases
  1. Anthropic
    Models

    Claude 2.1

    Claude 2.1 expanded context to 200K tokens and added a tool-use beta.

    For buildersLong context and external tools enable assistants grounded in a company’s documents.

  2. OpenAI
    Models

    GPT-4 Turbo (preview)

    GPT-4 Turbo launched in preview with a 128K context window and lower API prices.

    For buildersLonger context made it possible to supply much larger documents in a single request.

October 2023

1 release

September 2023

1 release
  1. Mistral
    Models

    Mistral 7B

    Mistral released a compact language model under the Apache 2.0 license.

    For buildersA small self-hosted model can be a useful baseline for a narrowly defined text task.

July 2023

3 releases
  1. Anthropic
    Models

    Claude 2

    Claude 2 launched with improved coding, mathematics, reasoning, and longer responses, alongside a public web beta.

    For buildersLong-document assistants became a practical use case for the Claude family.

May 2023

1 release
  1. Google
    Models

    PaLM 2

    Google introduced a language model family with improved multilingual, reasoning, and coding capabilities.

    For buildersA useful milestone in the evolution of multilingual assistants before Gemini.

March 2023

3 releases
  1. Adobe
    Models

    Adobe Firefly (beta)

    Adobe introduced the first Firefly model in beta for image generation and text effects.

    For buildersA milestone in bringing generative image models into established design workflows.

  2. Anthropic
    Models

    Claude and Claude Instant

    Anthropic introduced its first broadly offered Claude models, including a faster, lighter Instant version.

    For buildersAn early choice of model tiers for summarization, writing, and conversational applications.

  3. OpenAI
    Models

    GPT-4

    OpenAI introduced GPT-4, with improved reasoning and a multimodal architecture accepting text and images.

    For buildersUse this release to trace the move from text assistants toward visual reasoning.

February 2023

1 release
  1. Meta
    Models

    LLaMA (research release)

    Meta introduced a family of foundation language models for research, ranging from 7B to 65B parameters.

    For buildersA foundational milestone in the development of downloadable language models.

November 2022

2 releases
  1. Stability AI
    Models

    Stable Diffusion 2.0

    Stable Diffusion 2.0 introduced a new text encoder, higher-resolution generation, and depth-guided image transformations.

    For buildersExplore how preserving depth can help an image edit retain its original composition.

September 2022

1 release
  1. OpenAI
    Models

    Whisper

    OpenAI released open-source speech-recognition models for multilingual transcription and translation into English.

    For buildersTurn recorded interviews or lessons into searchable transcripts.

August 2022

1 release
  1. Stability AI
    Models

    Stable Diffusion

    Stability AI, LMU Munich researchers, and Runway co-released Stable Diffusion for text-to-image generation.

    For buildersOpen image-model weights enabled local generation and a large ecosystem of custom workflows.

May 2022

1 release
  1. Meta
    Models

    OPT-175B (research access)

    Meta shared a large pretrained language model for research, alongside smaller models and training documentation.

    For buildersA milestone in making large language model research more reproducible.

April 2022

1 release

About the tracker

A history of major AI model launches across language, coding, image, video, and speech, with official sources and practical notes for builders.

What does this AI release tracker cover?

Major models from 2022 onward, including OpenAI, Anthropic, Google, Meta, DeepSeek, Alibaba, Mistral, and image, video, and speech providers. Same-day model variants may share an entry. Product features, integrations, minor checkpoints, and community fine-tunes are excluded.

How are release dates and claims verified?

Dates follow the linked official announcement or documented release milestone. Restricted previews and announcements ahead of rollout are labeled. Access can vary by account, region, or plan, and historical models may no longer be available. Summaries describe the announcement; project ideas are our editorial suggestions, not hands-on benchmark results.

How should I choose a model for my project?

Start with your task, budget, and the tools your app needs. Test a candidate model on realistic inputs and compare its accuracy, latency, and cost. A newer release is a reason to evaluate, not an automatic reason to switch.

Compare AI and no-code tools →
Where can I learn to build with these tools?

Follow the related guides in the timeline, browse our tool directory, or choose a practical AI and no-code course to build a working project step by step.