較早收集於 11h

Google Veo 影片模型上線 AI Gateway

Google Veo 影片模型上線 AI Gateway
PostLinkedIn
閱讀原文: Vercel News
#photorealistic#image-to-video#native-audio#4k-resolutionvercel-ai-gateway

💡Veo 3.1's 4K photoreal videos + audio now via Vercel—perfect for cinematic AI prototypes.

⚡ 30-Second TL;DR

有什麼變化

Veo 3.1 寫實文字轉影片含音效

為什麼重要

此更新將 Google 高保真 Veo 帶給 Vercel 開發者,實現含音效的高解析生產級影片應用。簡化結合如 Gemini 圖像生成與 Veo 動畫的工作流程。

下一步行動

Generate a Veo video with google/veo-3.1-generate-001 and generateAudio: true in AI SDK 6.

誰應關注:Developers & AI Engineers

關鍵要點

  • Veo 3.1 寫實文字轉影片含音效
  • 圖像轉影片支援自訂起始圖像
  • 最高 4K 解析度及快速生成模式
  • 模型:google/veo-3.1-generate-001 及快速變體
  • AI Gateway Playground 無程式碼測試

🧠 深度解析

背景與延伸:來自公開資料,非原文內容。引用 6 個來源。

🔑 增強重點摘要

  • Veo 3.1, released in public preview on October 15, 2025, introduces photorealistic text-to-video generation with native synchronized audio including dialogue, ambient sounds, and music, marking a major upgrade from Veo 2 which lacked audio[1][3][4].
  • Supports image-to-video workflows, including referencing up to three images, providing first and last frame images, and extending Veo-created videos, with image-to-video launched for Veo 3 Preview on July 31, 2025[1].
  • Achieves up to 4K resolution with upsampling support added January 13, 2026, alongside 1080p, portrait videos, and new aspect ratios like 9:16 for reference-to-video[1][2].
  • Includes fast generation variants like Veo 3.1 Fast and Veo 3 Fast Preview, enabling rapid iterations alongside standard models such as google/veo-3.1-generate-001[1].
  • Veo 3 excels in cinematic visuals, realistic lighting, smoother motion, and better prompt understanding compared to Veo 2, though limited by short clip durations (4-8 seconds), occasional inconsistencies, and higher costs[3][4].
📊 競品分析▸ Show
FeatureGoogle Veo 3/3.1OpenAI Sora 2Kuaishou Kling 2.6
Audio GenerationFully synchronized dialogue, ambient, music from text [3][4]No native audio mentioned [3]Natural sound effects [3]
Motion/VisualsCinematic, realistic lighting, smooth motion [4]Cinematic narrative depth [3]Precise motion control, consistency [3]
InputsText, image (up to 3 refs), frames [1]Text prompts [3]Motion control focus [3]
StrengthsAudio integration, photorealism [3][4]Narrative depth [3]Fast, low-cost iteration [3]
Pricing/BenchmarksExpensive, limited access [4]Higher credit costs [3]Lower credit costs [3]

🛠️ 技術深入

  • Veo 3.1 models (e.g., veo-3.1-generate-001, fast variants) support 4, 6, 8-second durations, generally available short-duration videos since September 8, 2025[1].
  • Reference-to-video features: up to three images, first/last frames, video extension; 9:16 aspect ratio, 4K/1080p upsampling added January 13, 2026[1][2].
  • Image-to-video capability launched July 31, 2025 for Veo 3 Preview[1].
  • Generates videos with synchronized audio (dialogue, ambient, music) directly from text/image prompts, eliminating separate post-production[3][4].
  • Improved prompt adherence for camera angles, moods, lighting; realistic motion blur, reflections[4].

🔮 前景展望AI analysis grounded in cited sources

Vercel's AI Gateway integration of Veo 3.1 simplifies access via no-code playground and AI SDK 6 for Pro/Enterprise users, reducing multi-model friction and enabling creators to combine Veo with competitors like Sora 2 and Kling 2.6 for specialized workflows, potentially accelerating AI video adoption in filmmaking and social media by streamlining audio-inclusive generation[3].

時間線

2025-04
Released Veo 2.0-generate-001 as generally available text- and image-to-video model
2025-07
Launched image-to-video for Veo 3 Preview and released Veo 3 Fast Preview
2025-09
Veo 3 short-duration videos (4-8s) generally available
2025-10
Released Veo 3.1 and 3.1 Fast in public preview with image referencing, frame control, video extension
2026-01
Added 4K resolutions, portrait support for Veo; Veo 3.1 updates for 9:16 aspect ratio and upsampling
📰

AI 週報

閱讀本週精選 AI 大事摘要 →

👉相關動態

AI 策展新聞聚合。所有內容版權歸原始發布者所有。
原始來源: Vercel News

這是摘要,不是原文。去看原站,或訂閱每週簡報。

每週 AI 簡報

每週一封,可隨時退訂。