Runway Gen-4
Runway introduced a video model focused on consistent characters, objects, and environments across shots.
For buildersBuild a short storyboard using the same subject across several scenes.
6 model releases from 4 providers in March 2025, with announcement dates, official sources, and practical notes for builders.
Dates refer to the linked announcement or release milestone. Previews are labeled in the entries; historical availability and pricing may have changed.
Runway introduced a video model focused on consistent characters, objects, and environments across shots.
For buildersBuild a short storyboard using the same subject across several scenes.
Google introduced the first Gemini 2.5 model as an experimental reasoning model with stronger coding capabilities.
For buildersA useful milestone for comparing multimodal reasoning with earlier Gemini generations.
OpenAI introduced native image generation with improved prompt following, text rendering, and image-based editing.
For buildersCreate a visual from a brief, then refine it using the previous output as context.
OpenAI released GPT-4o Transcribe, GPT-4o mini Transcribe, and GPT-4o mini TTS for speech recognition and generation.
For buildersCombine transcription and speech generation to prototype a voice-based support assistant.
Cohere released Command A for enterprise language and agent workflows.
For buildersEvaluate grounded answers and tool use against internal business documents.
Google released a new open-model family with multilingual and multimodal capabilities.
For buildersOn-device document and image understanding can reduce reliance on hosted APIs.