Updated

AI models released in October 2026

7 model releases from 7 providers in October 2026, with announcement dates, official sources, and practical notes for builders.

Releases this month

Dates refer to the linked announcement or release milestone. Previews are labeled in the entries; historical availability and pricing may have changed.

Release details

  1. Aleph Alpha

    Kolibri

    Aleph Alpha released Kolibri, an English-German open-weight Mixture-of-Experts language model with 78B total parameters, 3B active parameters, and a context length of up to 1 million tokens. Full weights are available under Apache 2.0; the provider describes it as specialized for regulated, mission-critical work.

    For buildersEvaluate it locally on representative English and German documents, including long-context retrieval and abstention cases, before considering it for regulated workflows.

  2. Ai2

    AstaBrief 8B

    Ai2 released AstaBrief 8B, an open-weights model that turns a research question and retrieved literature into a cited scientific report in one pass. The model and training data are available for download; it is also offered as Fast mode in Ai2’s Asta platform.

    For buildersTry it on a small set of research questions with known source papers, then check citation support, omitted caveats, and whether the report is useful as a first draft.

  3. Cactus

    Cactus Whistle

    Cactus released Whistle, a 16.9 MB on-device speech-recognition model for CPU deployment. It transcribes up to 30 seconds of audio in seven languages and also returns word timestamps and speech embeddings without sending audio off-device.

    For buildersTest it on short, representative clips on the target device, measuring word errors, timestamps, latency, and performance on names before using transcripts to trigger actions.

  4. Decagon

    Chord

    Decagon introduced Chord, its first speech model, trained and post-trained for customer conversations and used by its Voice 3 system. The company says Chord shapes speech phrase by phrase and is trained on licensed data and consented voice talent; Voice 3 supports more than 70 languages.

    For buildersIf evaluating a customer-service voice stack, test real turn-taking, interruptions, confirmation details, pronunciation, and handoff behavior across representative calls.

  5. Cloudflare

    Cloudflare Clef and Clef-flash

    Cloudflare released Clef and Clef-flash, open-weight decision models hosted on Workers AI that return probabilities for typed choices instead of generating text. Clef adds image input and a 64K context window; the weights are available under Apache 2.0.

    For buildersEvaluate one bounded task such as support-ticket routing against your current classifier or model, measuring decision quality, confidence calibration, latency, and how often a human should review the result.

  6. Tavus

    Griffin-Lite (research preview)

    Tavus introduced Griffin, a real-time video-to-video model that combines perception, conversational behavior, and audiovisual generation. Only Griffin-Lite is currently available, as a research preview for select early testers; a wider release is still to come.

    For buildersFor a face-to-face agent prototype, evaluate whether live video and interruption handling improve conversation flow, and assess disclosure and consent requirements before user testing.

  7. Microsoft

    MAI-Transcribe-2-Streaming, MAI-Voice-2.1, and MAI-Voice-2.1-Flash

    Microsoft introduced MAI-Transcribe-2-Streaming for live speech transcription and two speech-generation models, MAI-Voice-2.1 and MAI-Voice-2.1-Flash. Transcribe supports 60 languages with automatic detection and partial transcripts in just over 100 ms; the voice models support multilingual speech generation, with Flash designed for low-latency workloads.

    For buildersPrototype a voice-agent loop using representative call audio, then compare transcription latency and short-utterance accuracy alongside response quality and voice consistency across the languages you need.