API Free WeekHundreds of APIs are free to call this week.Browse free APIs

Gemini Omni Flash 1.1 Image to Video

google/gemini-omni-flash/1.1/image-to-video

Gemini Omni Flash 1.1 by Google - animate still images into dynamic videos with AI. Transform photos into cinematic clips with natural motion, camera movement, and optional audio generation.

Base price
$0.80USD / run
Execution
async
Model type
video
Input fields
6

Try the model

Playground

Open playground
Input
The text prompt describing how the first image should be animated or interpolated into the optional end image.

PNG, JPEG, WebP, or GIF · 20 MiB maximum

URL of the first frame to animate.

PNG, JPEG, WebP, or GIF · 20 MiB maximum

Optional URL of the end frame. When provided, the model interpolates between the two images in their listed order.
The resolution of the generated video. Allowed values: 360p, 720p, 1080p, 4k.
310
The duration of the generated video, in seconds. Range: 3 to 10.
The aspect ratio of the generated image. Allowed values: 16:9, 9:16.
OutputReady

Your output will appear here

Complete the inputs, then click Run.

Specifications

Pricing

Base price
$0.80 / run
Billing formula
(params.duration ?? 8) * ((params.resolution ?? "720p") == "360p" ? 0.03 : (params.resolution ?? "720p") == "1080p" ? 0.15 : (params.resolution ?? "720p") == "4k" ? 0.30 : 0.10)

Context & modalities

Input
Schema-defined
Output
video

Capabilities

Chat
Not supported
Vision
Not supported
Reasoning
Not supported
Structured output
Not supported
Function calling
Not supported
Audio input
Not supported

Access

Provider
Google
Model ID
google/gemini-omni-flash/1.1/image-to-video
Execution
async
API
Unified Run API
Endpoint
/v1/run

API README

Gemini Omni Flash 1.1 Image To Video

google/gemini-omni-flash/1.1/image-to-video animates a supplied still image into a coherent video, using the key visual to anchor subject identity and opening composition while the prompt directs movement and sound. Gemini Omni Flash 1.1 improves visual quality and instruction following while adding extension, first-and-last-frame interpolation, and delivery from 360p through 4K. This combination makes the model a practical choice when the creative outcome depends on those qualities rather than on a generic media conversion.

For production work, this stable 1.1 image-to-video workflow supports three-to-ten-second generation and multiple resolution tiers, giving teams a clearer path from fast exploration to high-resolution finishing. The result is most reliable when the source material and creative brief clearly describe the intended subject, progression, visual or sonic character, and the qualities that must remain unchanged.

Highlights

1.1 image animation. Applies the updated 1.1 behavior to a supplied still, creating a more refined moving interpretation anchored to its recognizable content.

Stable subject development. Preserves important identity and composition cues as characters, objects, atmosphere, and camera motion develop over time.

Coherent spatial motion. Adds movement with attention to scene depth and relationships between visual elements, avoiding the feel of a simple slideshow effect.

Higher-confidence key-art extension. Serves campaigns and storyboards that need an approved frame translated into footage while retaining its art direction.

Pricing

Duration360p720p1080p4K
3 seconds$0.09$0.30$0.45$0.90
4 seconds$0.12$0.40$0.60$1.20
5 seconds$0.15$0.50$0.75$1.50
6 seconds$0.18$0.60$0.90$1.80
7 seconds$0.21$0.70$1.05$2.10
8 seconds$0.24$0.80$1.20$2.40
9 seconds$0.27$0.90$1.35$2.70
10 seconds$0.30$1.00$1.50$3.00

Billed per second of requested output duration.

When to Use

✅ Good fit❌ Consider alternatives
A still image should anchor the video's subject and compositionThe whole scene should be invented from text alone
A managed asynchronous result is suitable for the production pipelineA synchronous, interactive editor is essential
The documented controls cover the required duration, framing, or formatThe project needs controls outside this endpoint's schema
Creative iteration benefits from a repeatable request structureExact deterministic pixels, frames, geometry, or samples are mandatory
A finished downloadable media asset is the desired deliverableEditable source layers or a native project file are required

Prompt Guide

For image-conditioned generation, state the intended result first, then add the subject or source treatment, progression, style, and delivery constraints. Keep one creative variable per phrase, use the documented field names for controls, and change one setting at a time when comparing results.

{
  "aspect_ratio": "21:9",
  "duration": 8,
  "image": "https://example.com/reference.png",
  "prompt": "The dog turns its head and wags its tail in warm sunlight.",
  "resolution": "720p"
}

Technical Specs

SpecValue
Model IDgoogle/gemini-omni-flash/1.1/image-to-video
Inputsaspect_ratio, duration, end_image, image, prompt, resolution
Required inputsprompt, image
Output fieldscontent_type, url
ExecutionAsync (submit, then poll for result)
Resolution360p / 720p / 1080p / 4k
Aspect Ratio21:9 / 16:9 / 3:2 / 4:3 / 5:4 / 1:1 / 4:5 / 3:4 / 2:3 / 9:16

Related Models

Start building

Send your first request

OpenAI-compatible endpoint with unified authentication and usage tracking.

Production API
Unified Run API endpoint
https://api.sandbase.ai/v1/run
Model ID
google/gemini-omni-flash/1.1/image-to-video
# The quoted heredoc keeps Unicode and shell metacharacters unchanged.
result=$(curl --fail-with-body --silent \
  -X POST "https://api.sandbase.ai/v1/run" \
  -H "Authorization: Bearer $SANDBASE_API_KEY" \
  -H "Content-Type: application/json" \
  --data-binary @- <<'SANDBASE_JSON'
{
  "model": "google/gemini-omni-flash/1.1/image-to-video",
  "image": "https://storage.googleapis.com/falserverless/example_inputs/dog.png",
  "prompt": "The dog turns its head and wags its tail in warm sunlight.",
  "duration": 8,
  "resolution": "720p",
  "aspect_ratio": "16:9"
}
SANDBASE_JSON
)
run_id=$(printf '%s' "$result" | jq -r .id)
for attempt in $(seq 1 120); do
  status=$(printf '%s' "$result" | jq -r .status)
  case "$status" in completed|failed|timeout) break ;; esac
  sleep 2
  result=$(curl --fail-with-body --silent \
    -H "Authorization: Bearer $SANDBASE_API_KEY" \
    "https://api.sandbase.ai/v1/run/$run_id")
done
status=$(printf '%s' "$result" | jq -r .status)
[ "$status" = completed ] || { echo "Generation ended: $status" >&2; exit 1; }
printf '%s\n' "$result"

Choose your model

Compare models

All language models
ModelContextInput / 1MOutput / 1MReleased
Google———Aug 27, 2026
Google———Aug 27, 2026
Google———Aug 27, 2026
Google———Aug 27, 2026
Google———Jun 30, 2026
Google———Jun 30, 2026