最高分辨率
1080p
最长时长
8s
输入
文本、图片
起步价
12 积分
概览
关于 Veo 3.1
Veo 3.1 is Google DeepMind's flagship text-to-video model, generating up to 8-second 1080p clips with synchronized audio, realistic physics, and cinema-grade lighting. The 3.1 release brings tighter prompt adherence, sharper character consistency across frames, and dramatically reduced morphing artifacts that plagued earlier video models. Use it for narrative shots, product films, and dialogue scenes where audio matters.
优势
- Native audio generation including dialogue, foley, and ambient sound
- Best-in-class prompt adherence for complex compositions
- Cinematic lighting and shallow depth-of-field by default
- Stable character identity across full 8-second clips
- 1080p resolution at native generation, no upscale required
适合场景
- Narrative short films and ad spots
- Product hero videos with synchronized voiceover
- Cinematic establishing shots and B-roll
- Dialogue scenes with realistic lip-sync
价格
单次生成成本
5s clip
12积分
Standard 5 seconds
8s clip
20积分
Extended 8 seconds
积分成本仅供参考,可能随供应商价格调整。订阅积分每月刷新,充值积分永不过期。
并排对比
对比 Veo 3.1
问题
常见问题
What's new in Veo 3.1 versus Veo 3?
Veo 3.1 ships with tighter prompt adherence, improved character consistency across frames, and reduced morphing artifacts. Native audio quality also takes a noticeable step up.
Can Veo 3.1 generate dialogue?
Yes. Veo 3.1 supports dialogue in prompts and produces synchronized lip-sync. Specify the line in your prompt and Veo will generate matching mouth movement.
How long can Veo 3.1 clips be?
Veo 3.1 generates up to 8 seconds per clip at 1080p. For longer narratives, chain multiple clips with consistent character prompts.
Does Veo 3.1 support image-to-video?
Yes. Upload a starting frame and Veo 3.1 will animate it while preserving subject identity and scene composition.
准备使用 Veo 3.1 生成?
免费开始,无需信用卡。
免费试用 Veo 3.1