# Gemini Omni Flash Reference-to-Video Developer — Atlas Cloud API

> Gemini Omni Flash is Google's multimodal video generation model. This reference-to-video variant transforms existing video clips using reference images and text prompts, enabling video style transfer, scene editing, and character insertion.

This is the machine-readable API reference for **Gemini Omni Flash Reference-to-Video Developer** on Atlas Cloud,
a unified API platform for 400+ AI models across text, image, video, audio and 3D.

- **Model ID**: `google/gemini-omni-flash/reference-to-video-developer`
- **Built by**: Google
- **Modality**: Video
- **Model page**: https://www.atlascloud.ai/models/google/gemini-omni-flash/reference-to-video-developer
- **API key**: https://www.atlascloud.ai/console/api-keys
- **Docs**: https://www.atlascloud.ai/docs

## Pricing on Atlas Cloud

- $0.12 per second of generated video
- Pay-as-you-go. No minimum spend, no subscription required.

> **These are the authoritative Atlas Cloud rates for this model.** Any price that
> appears in the vendor description further down refers to a different platform or
> a different model variant and does not apply here.

## Use this model from an AI agent

Atlas Cloud ships three first-party integration surfaces. All three authenticate
with the same API key via the `ATLASCLOUD_API_KEY` environment variable.

### MCP server

The official MCP server (`atlascloud-mcp`) exposes this model to any
MCP-compatible host — Claude Code, OpenAI Codex, Cursor, Gemini CLI, Goose,
Claude Desktop. One-line install:

```bash
# Claude Code
claude mcp add atlascloud -- npx -y atlascloud-mcp

# OpenAI Codex CLI
codex mcp add atlascloud -- npx -y atlascloud-mcp

# Gemini CLI
gemini mcp add atlascloud -- npx -y atlascloud-mcp

export ATLASCLOUD_API_KEY="your-api-key"
```

Then ask in plain English; the agent calls `atlas_generate_video` with `model: "google/gemini-omni-flash/reference-to-video-developer"`.
The server fetches each model's schema and validates parameters before submitting,
so invalid requests fail fast without spending credits.

MCP docs: https://www.atlascloud.ai/docs/mcp-server

### Agent Skills

`atlas-cloud-skills` is a portable skill package (API reference, code templates in
Python / Node.js / cURL, model IDs with pricing) for Claude Code, Cursor, Codex and
12+ other agents:

```bash
npx skills add AtlasCloudAI/atlas-cloud-skills
export ATLASCLOUD_API_KEY="your-api-key"
```

Skills docs: https://www.atlascloud.ai/docs/skills

### CLI

The `atlas` binary runs Atlas Cloud from a terminal or CI script. Async media jobs
are polled and downloaded automatically (use `--no-download` when a script only
needs the output URLs):

```bash
# Install (Homebrew, npm, or shell installer)
brew install AtlasCloudAI/tap/atlascloud
# npm install -g atlascloud-cli
# curl -fsSL https://raw.githubusercontent.com/AtlasCloudAI/cli/main/install.sh | sh

atlas auth login
atlas generate video google/gemini-omni-flash/reference-to-video-developer -p "Your prompt here"
```

CLI docs: https://www.atlascloud.ai/docs/cli

## HTTP API reference

- **Submit endpoint (POST)**: `https://api.atlascloud.ai/api/v1/model/generateVideo` — start an async generation; returns a `prediction_id`
- **Poll endpoint (GET)**: `https://api.atlascloud.ai/api/v1/model/prediction/{prediction_id}` — poll this until the prediction finishes
- **Model ID**: `google/gemini-omni-flash/reference-to-video-developer`


## API Information

This model can be used via our HTTP API or more conveniently via our client libraries.
See the input and output schema below, as well as the usage examples.


### Input Schema

The API accepts the following input parameters:

- **`model`** (`string`, _required_):
  model name
  - Default: `"google/gemini-omni-flash/reference-to-video-developer"`

- **`prompt`** (`string`, _required_):
  Text prompt for generation. Describes the target content, style, camera language, or character actions. Maximum 20,000 characters.

- **`images`** (`array[string]`, _optional_):
  Images to use as character, scene, or style references. Accepts 1 to 5 images when combined with a video reference (video costs 2 units out of a total quota of 7). Supported formats: PNG, JPEG, JPG, WebP. Each image is limited to 20MB.
  - Min items: 1
  - Max items: 5

- **`video_clips`** (`array[object]`, _required_):
  Source video clips to use as references for generation. Supports 1 video clip. Each video is limited to 100MB and 30 seconds duration. The trimmed segment (ends - start) must not exceed 10 seconds.
  - Min items: 1
  - Max items: 1
  - Item properties:
    - **`url`** (`string`, _required_):
      URL of the source video clip.

    - **`start`** (`integer`, _required_):
      Start time in seconds for trimming the video clip.
      - Default: `0`
      - Min: 0
      - Max: 29

    - **`ends`** (`integer`, _required_):
      End time in seconds for trimming the video clip. The difference between ends and start must not exceed 10 seconds.
      - Default: `10`
      - Min: 1
      - Max: 30


- **`duration`** (`integer`, _optional_):
  The duration of the generated video in seconds.
  - Default: `8`
  - Options: 4, 6, 8, 10

- **`aspect_ratio`** (`string`, _optional_):
  The aspect ratio of the generated video.
  - Default: `"16:9"`
  - Options: "16:9", "9:16"

- **`resolution`** (`string`, _optional_):
  The resolution of the generated video.
  - Default: `"720p"`
  - Options: "720p", "1080p", "4k"

- **`seed`** (`integer`, _optional_):
  Random seed for reproducibility. Use -1 to use a random seed.
  - Default: `-1`



**Required Parameters Example**:

```json
{
  "model": "google/gemini-omni-flash/reference-to-video-developer",
  "prompt": "",
  "video_clips": [
    {
      "url": "",
      "start": 0,
      "ends": 10
    }
  ]
}
```


**Full Example**:

```json
{
  "model": "google/gemini-omni-flash/reference-to-video-developer",
  "prompt": "",
  "images": [
    ""
  ],
  "video_clips": [
    {
      "url": "",
      "start": 0,
      "ends": 10
    }
  ],
  "duration": 8,
  "aspect_ratio": "16:9",
  "resolution": "720p",
  "seed": -1
}
```


### Output Schema

The API returns the following output format:


- **`code`** (`integer`, _optional_):
  HTTP status code of the response.

- **`message`** (`string`, _optional_):
  Human-readable message; non-empty on failure.

- **`data`** (`object`, _optional_):
  - Properties:
    - **`id`** (`string`, _optional_):
      Unique identifier for the prediction.

    - **`model`** (`string`, _optional_):
      Model ID used for the prediction.

    - **`outputs`** (`array[string]`, _optional_):
      Array of URLs to the generated content. Null when status is not completed.

    - **`urls`** (`object`, _optional_):
      Object containing related API endpoints.
      - Properties:
        - **`get`** (`string`, _optional_):
          URL to poll for the prediction result.


    - **`status`** (`string`, _optional_):
      Status of the task: created, processing, completed, timeout, or failed.

    - **`created_at`** (`string`, _optional_):
      ISO timestamp of when the request was created (e.g., "2023-04-01T12:34:56.789Z").

    - **`error`** (`string`, _optional_):
      Error message if the task failed, empty string otherwise.

    - **`error_code`** (`integer`, _optional_):
      Error code if the task failed.

    - **`executionTime`** (`number`, _optional_):
      Total execution time in milliseconds.

    - **`timings`** (`object`, _optional_):
      Detailed timing breakdown.
      - Properties:
        - **`inference`** (`number`, _optional_):
          Inference time in milliseconds.





**Example Response**:

```json
{
  "code": 0,
  "message": "",
  "data": {
    "id": "",
    "model": "",
    "outputs": [
      ""
    ],
    "urls": {
      "get": ""
    },
    "status": "",
    "created_at": "",
    "error": "",
    "error_code": 0,
    "executionTime": 0,
    "timings": {
      "inference": 0
    }
  }
}
```


## Usage Examples

### cURL

```bash
# Step 1: Start generation (async)
curl -X POST "https://api.atlascloud.ai/api/v1/model/generateVideo" \
  -H "Authorization: Bearer $ATLASCLOUD_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
  "model": "google/gemini-omni-flash/reference-to-video-developer",
  "prompt": "",
  "images": [
    ""
  ],
  "video_clips": [
    {
      "url": "",
      "start": 0,
      "ends": 10
    }
  ],
  "duration": 8,
  "aspect_ratio": "16:9",
  "resolution": "720p",
  "seed": -1
}'

# Response will contain: {"code": 200, "data": {"id": "prediction_id", "status": "processing"}}

# Step 2: Poll for result (replace {prediction_id} with the id returned above)
curl -X GET "https://api.atlascloud.ai/api/v1/model/prediction/{prediction_id}" \
  -H "Authorization: Bearer $ATLASCLOUD_API_KEY"

# Keep polling until status is "completed", "succeeded" or "failed"
# When completed, outputs will contain the generated content URL(s)
```

## Additional Resources

### Documentation

- [Model Playground](https://www.atlascloud.ai/models/google/gemini-omni-flash/reference-to-video-developer)

## About this model

_Vendor-supplied description. Any pricing or endpoint mentioned below refers to_
_other platforms — use the Atlas Cloud values above._

### Gemini Omni Flash — Reference to Video (Developer)

**Model ID:** `google/gemini-omni-flash/reference-to-video-developer`

Gemini Omni is Google's multimodal video generation model designed to create high-quality video content from diverse input types. This variant accepts a **text prompt, reference images, and a source video clip**, enabling the most expressive form of video generation: transforming existing footage while preserving coherence and injecting new creative direction.

---

#### Overview

Gemini Omni brings together Google's deep knowledge of physics, narrative logic, biology, culture, and visual composition to produce contextually coherent videos. Rather than simple clip synthesis, the model reasons about scene dynamics, camera language, and temporal flow to produce results that feel intentional and cinematic.

With both image and video inputs, the model can change what happens in a scene, add or remove objects, adjust camera perspective, apply new visual effects, or completely reimagine a clip's style — all while maintaining temporal coherence with the source material.

The developer tier provides direct API access with full control over generation parameters including resolution, aspect ratio, and random seed.

---

#### Key Capabilities

_(Description truncated. Full text on the model page.)_

---

Atlas Cloud — one API for 400+ AI models. Model page: https://www.atlascloud.ai/models/google/gemini-omni-flash/reference-to-video-developer · Docs: https://www.atlascloud.ai/docs
