14 articles about video generation
Google's Gemini Omni 1.1 Flash brings studio-quality video generation, scene extension, first-and-last-frame interpolation, and crisp 4K upscaling to Flow, AI Studio, the Gemini app, and Gemini Enterprise.
Google's multimodal video generation model exits beta with 1080p output, 15-second clips, and 40% faster generation — now production-ready for developers.
BFL releases FLUX Upscale, a standalone video upscaling tool that regenerates footage from 480p up to native 4K while fixing generated-video artifacts like smudged faces and gridded textures.
Black Forest Labs ships FLUX Upscale, a new tool and API endpoint that regenerates any video at up to native 4K resolution — extending the FLUX 3 family beyond image generation into video pipelines.
ByteDance released Seedance 2.5, generating 30-second native 4K video with 50 reference inputs and synchronized audio. The model doubles clip length and adds region-level editing, making it the most capable AI video generator available.
MiniMax launches H3, a general-purpose multimodal generation model that unifies text, image, video, and audio in one context — 15-second 2K video with native stereo sound, with open weights promised within days.
Grok Imagine Video 1.5 beat Sora 2, Veo 3.1, and Kling in blind user benchmarks at 86% lower cost, with native audio and dialogue in one pass.
Luma AI launches the Ray3.2 model and API with frame-level creative control, plus Luma Scenes for storyboard-first video workflows — and opens an Open Physical AI Lab to attack generalization in physical AI.
Runway releases Aleph, its most advanced video generation model yet, featuring 4K resolution output, precise motion control, and real-time editing capabilities.
Pika releases version 3.0 with synchronized lip movements for AI-generated characters, multi-character dialogue scenes, and real-time video generation.
Stability AI open-sources Stable Video Diffusion 4K, bringing high-resolution video generation to the open-source community with unprecedented quality.
Google launched Gemini Omni Flash, the first model in the Omni family, at I/O 2026 — a multimodal video generation and editing system that accepts text, images, audio, and video as input.
Kuaishou's Kling AI released version 3.0 with native 4K video generation, multi-shot storyboarding, and native audio — positioning it as the most feature-complete AI video platform for professional creators.
Google launched Gemini 3.5 Flash with frontier agentic performance, introduced Gemini Spark as a persistent personal AI assistant, and unveiled Omni Flash for video generation at I/O 2026.