FLUX.2
Black Forest Labs introduced an image model for generation and editing with multiple visual references.
For buildersUse reference images to test product, character, and style consistency across a set of assets.
10 model releases from 7 providers in November 2025, with announcement dates, official sources, and practical notes for builders.
Dates refer to the linked announcement or release milestone. Previews are labeled in the entries; historical availability and pricing may have changed.
Black Forest Labs introduced an image model for generation and editing with multiple visual references.
For buildersUse reference images to test product, character, and style consistency across a set of assets.
Opus 4.5 launched with stronger coding, agent, and computer-use capabilities at a lower Opus price.
For buildersCompare completion quality and the number of retries, not just a model’s per-token price.
Nano Banana Pro brought Gemini 3 reasoning to image generation, including better text rendering and grounded visual content.
For buildersTest diagrams, campaign assets, and localized layouts with readable text.
OpenAI introduced a coding model designed for longer engineering tasks and more efficient reasoning.
For buildersCompare sustained debugging across several iterations.
xAI released a fast model for tool-using agents, alongside its Agent Tools API.
For buildersMeasure the latency of the whole tool workflow, not just the first generated token.
Gemini 3 Pro launched with improvements in reasoning, multimodal understanding, and agentic coding.
For buildersTest visual reasoning and code generation together when building interfaces from references.
An updated Grok model focused on conversational quality, nuanced instructions, and collaborative interactions.
For buildersEvaluate tone and instruction following using the conversations your users actually have.
The GPT-5.1 update introduced more conversational responses and more adaptive use of reasoning time.
For buildersA task router can reserve longer thinking for the inputs that benefit from it.
ElevenLabs introduced a streaming speech-to-text model for low-latency transcription.
For buildersEvaluate live captions with background noise, interruptions, and specialized vocabulary.
Moonshot introduced a thinking model designed for reasoning with repeated tool use.
For buildersTest multi-step research tasks that require gathering and reconciling information.