MiniMax M2.1
M2.1 expanded multilingual programming and development capabilities.
For buildersEvaluate code quality across the languages used in your application.
13 model releases from 9 providers in December 2025, with announcement dates, official sources, and practical notes for builders.
Dates refer to the linked announcement or release milestone. Previews are labeled in the entries; historical availability and pricing may have changed.
M2.1 expanded multilingual programming and development capabilities.
For buildersEvaluate code quality across the languages used in your application.
GLM-4.7 improved coding and agent task execution.
For buildersCompare complete development tasks.
GPT-5.2-Codex launched in Codex for paid ChatGPT users, with API access planned afterward.
For buildersTest a migration or refactor that requires maintaining context across a long session.
Google expanded Gemini 3 with a faster, lower-cost Flash model for reasoning and multimodal tasks.
For buildersUse a fast model for interactive tasks where waiting time affects the user experience.
OpenAI released an image model with faster generation and more precise edits that preserve existing details.
For buildersEvaluate a series of edits for consistency, rather than judging only the first image.
NVIDIA released Nano as the first available model in its Nemotron 3 family, with larger models announced for later release.
For buildersEvaluate an open model for a focused agent workload on your own infrastructure.
The GPT-5.2 series launched for professional work, including documents, spreadsheets, coding, and longer tool workflows.
For buildersTest complete deliverables, such as an editable report and its calculations, rather than isolated answers.
Mistral released the second Devstral coding generation in 123B and 24B sizes.
For buildersCheck each variant’s license and hardware requirements before choosing a deployment.
Amazon introduced Nova 2 Lite and a preview of Nova 2 Pro with expanded reasoning capabilities.
For buildersEvaluate everyday automation separately from complex reasoning workloads.
Amazon released its second-generation real-time conversational speech model.
For buildersPrototype a spoken assistant and measure its handling of interruptions and follow-ups.
Mistral released Large 3 alongside smaller 3B, 8B, and 14B models under Apache 2.0.
For buildersA family of model sizes supports deployment from smaller devices to larger server workloads.
DeepSeek released reasoning-first models for agent workflows, with a higher-reasoning Speciale variant on a temporary endpoint.
For buildersDifferent reasoning tiers should be compared on successful task completion and cost.
Runway announced a video model with improved motion quality, prompt adherence, and visual fidelity.
For buildersCompare complex motion prompts and multi-part scene descriptions.