DeepInfra raises $107M Series B to scale the inference cloud — read the announcement

google logo

google/

veo-3.1

Partner

$0.4000

/ second

Veo 3.1 is the latest text-to-video model from Google that generates high-fidelity, cinematic videos with synchronized audio from a simple text prompt. It excels at creating realistic and imaginative scenes with a deep understanding of natural language and visual dynamics.

Public
google/veo-3.1 cover image
api

Input

Prompt

text prompt

Please upload an image file

You need to log in to use this model

Log In

Settings

Enhance Prompt

Whether to enhance the prompt with additional context

Generate Audio

Whether to generate audio for the video

Negative Prompt

Negative text prompt to avoid unwanted content. (Default: empty)

Aspect Ratio

Aspect ratio of the output video

Seed

Seed for reproducibility, must be a non-negative integer (Default: empty, 0 ≤ seed ≤ 4294967295)

Resolution

Resolution of the output video

Sample Count

Number of samples to generate, must be between 1 and 4 (Default: empty, 1 ≤ sample_count ≤ 4)

Person Generation

Whether to allow adult faces are generated in the video

Output