Loading article…
Google DeepMind’s Veo 3.1 adds native 9:16 vertical video, 4K upscaling, and 48kHz synchronized audio, aiming to lead in cinematic AI storytelling tools.
Google DeepMind has launched Veo 3.1, an updated AI video generation model that introduces native 9:16 vertical video support and high-fidelity 48kHz synchronized audio. The release targets professional creators and social media workflows by allowing users to generate full-screen portrait content directly, eliminating the need for manual cropping [1].
| At a glance | |
|---|---|
| Developer | Google DeepMind |
| Key Feature | Native 9:16 vertical video generation |
| Audio Capability | 48kHz synchronized dialogue |
| Output Resolution | Up to 4K |
The core update to Veo 3.1 centers on the "Ingredients to Video" feature, which now supports native vertical aspect ratios [1]. By inputting reference images, users can generate expressive, full-screen portrait videos that maintain character and object consistency across different scenes and camera angles [2]. This capability is designed to streamline the production of content for platforms like YouTube Shorts, Instagram, and TikTok [1].
Beyond format adjustments, the model introduces 4K upscaling, a significant upgrade from standard outputs intended to meet the texture and clarity requirements of professional and enterprise productions [1]. The integration of 48kHz audio—capable of generating real, synchronized dialogue rather than just ambient sound—positions the model as a tool for narrative-driven filmmaking [2]. To ensure transparency, Google has implemented SynthID watermarking across all AI-generated videos, allowing users to verify content origin via Gemini [1].
Veo 3.1 is currently being deployed across Google’s ecosystem, including the Gemini app, YouTube Create, and enterprise-facing tools like Vertex AI and Google Vids [1]. The model’s performance has been measured against industry benchmarks, including Meta’s MovieGenBench, where it outperformed other models in overall preference and prompt adherence as of October 2025 [3].
While competitors such as Kling 3.0 and Seedance 2.0 offer alternatives in raw motion control and multimodal input, Google is positioning Veo 3.1 as the primary choice for fotorrealistic narrative scenes requiring credible, synchronized speech [2]. The company has also partnered with director Darren Aronofsky’s venture, Primordial Soup, to test the model’s ability to integrate live-action footage with AI-generated sequences [3].
| Benchmark Comparison | Performance Metric |
|---|---|
| MovieGenBench (Oct 2025) | Best overall preference [3] |
| VBench I2V (Oct 2025) | Preferred for prompt intent [3] |
The shift toward native vertical generation and synchronized dialogue suggests that Google is prioritizing the needs of social media creators and narrative filmmakers simultaneously. Whether these tools successfully replace traditional editing workflows will depend on how effectively the model maintains character consistency over extended, complex sequences.
Coverage is mostly measured — 295 of 300 reports stay neutral.
Every Monday — the token unlocks, Fed dates & catalysts set to move crypto and markets this week. So you’re never blindsided.
Free · 3-min read · one-click unsubscribe
AI-assisted synthesis by the TrendWatcher Editorial Desk · sourced from 3 outlets · Sep 14, 2026 · How we report
Google is purchasing 396 megawatts of clean energy from Fervo Energy's Cape Station enhanced geothermal systems project in Utah. As of September 2026, this power is intended to serve as a foundational building block for a planned data center, with the option for Google to increase its offtake to nearly 1 gigawatt by June 2030.
Google is expected to begin receiving power from the 396-megawatt capacity contract starting in 2028. The Cape Station project itself is expected to begin producing its first power by the end of 2026.
The Google Pixel 11 Pro Fold is priced at $1,899. As of September 2026, the device features an IP68 rating for dust and water protection and an olive-green finish.