google/gemini-omni-flash/video-edit

video-in-video

Gemini Omni Flash Video Edit API by Google

google/gemini-omni-flash/video-edit

Video-edit

A natively multimodal Google DeepMind model that edits an existing video from a text prompt with optional reference images, applying scene-consistent changes and native audio while preserving the untouched footage.

INPUT

Prompt *

Video *

Puoi trascinare un file qui o fare clic per caricarlo

MAX:1

Immagini(0/5)

Puoi trascinare un file qui o fare clic per caricarlo

MAX:5

Thinking level

Risoluzione

Seed

OUTPUT

In attesa

I video generati appariranno qui

Configura le impostazioni e clicca Esegui per iniziare

La tua richiesta costerà $0.14 per esecuzione. Con $10 puoi eseguire questo modello circa 71 volte.

Ecco cosa puoi fare dopo:

Seedance 2.0 Kling v3 Vidu Wan2.7

Parametri

Esempio di codice
import requests
import time

# Step 1: Start video generation
generate_url = "https://api.atlascloud.ai/api/v1/model/generateVideo"
headers = {
    "Content-Type": "application/json",
    "Authorization": "Bearer $ATLASCLOUD_API_KEY"
}
data = {
    "model": "google/gemini-omni-flash/video-edit",  # Required. model name
    "video": "example_value",  # Required. The source video to edit
    "prompt": "A beautiful sunset over the ocean with gentle waves",  # Required. Text prompt describing the edit to apply to the source video (e
    "images": [
        "https://example.com/image1.jpg"
    ],  # Images to use as character, scene, or style references
    "resolution": "720p",  # The resolution of the generated video. options: 720p
    "thinking_level": "default",  # Controls the amount of internal reasoning the model performs before generating a response. options: default | high | low
    "seed": -1,  # The random seed to use for the generation
}

generate_response = requests.post(generate_url, headers=headers, json=data)
generate_result = generate_response.json()
prediction_id = generate_result["data"]["id"]

# Step 2: Poll for result
poll_url = f"https://api.atlascloud.ai/api/v1/model/prediction/{prediction_id}"

def check_status():
    while True:
        response = requests.get(poll_url, headers={"Authorization": "Bearer $ATLASCLOUD_API_KEY"})
        result = response.json()

        if result["data"]["status"] in ["completed", "succeeded"]:
            print("Generated video:", result["data"]["outputs"][0])
            return result["data"]["outputs"][0]
        elif result["data"]["status"] == "failed":
            raise Exception(result["data"]["error"] or "Generation failed")
        else:
            # Still processing, wait 2 seconds
            time.sleep(2)

video_url = check_status()

Installa

Installa il pacchetto di dipendenze richiesto.

pip install requests

Autenticazione

Tutte le richieste API richiedono l'autenticazione tramite una chiave API. Puoi ottenere la tua chiave API dalla dashboard di Atlas Cloud.

export ATLASCLOUD_API_KEY="your-api-key-here"

Header HTTP

import os

API_KEY = os.environ.get("ATLASCLOUD_API_KEY")
headers = {
    "Content-Type": "application/json",
    "Authorization": f"Bearer {API_KEY}"
}

Proteggi la tua chiave API

Non esporre mai la tua chiave API nel codice lato client o nei repository pubblici. Utilizza invece variabili d'ambiente o un proxy backend.

Invia una richiesta

import requests

url = "https://api.atlascloud.ai/api/v1/model/generateVideo"
headers = {
    "Content-Type": "application/json",
    "Authorization": "Bearer $ATLASCLOUD_API_KEY"
}
data = {
    "model": "your-model",
    "prompt": "A beautiful landscape"
}

response = requests.post(url, headers=headers, json=data)
print(response.json())

Invia una richiesta

Invia una richiesta di generazione asincrona. L'API restituisce un ID di previsione che puoi usare per controllare lo stato e recuperare il risultato.

POST/api/v1/model/generateVideo

Corpo della richiesta

import requests

url = "https://api.atlascloud.ai/api/v1/model/generateVideo"
headers = {
    "Content-Type": "application/json",
    "Authorization": "Bearer $ATLASCLOUD_API_KEY"
}

data = {
    "model": "google/gemini-omni-flash/video-edit",
    "prompt": "A beautiful sunset over the ocean with gentle waves"
}

response = requests.post(url, headers=headers, json=data)
result = response.json()

print(f"Prediction ID: {result['data']['id']}")
print(f"Status: {result['data']['status']}")

Risposta

{
  "code": 200,
  "data": {
    "id": "pred_abc123",
    "status": "processing",
    "model": "model-name",
    "created_at": "2025-01-01T00:00:00Z"
  }
}

Controlla lo stato

Interroga l'endpoint di previsione per verificare lo stato attuale della tua richiesta.

GET/api/v1/model/prediction/{prediction_id}

Esempio di polling

import requests
import time

prediction_id = "pred_abc123"
url = f"https://api.atlascloud.ai/api/v1/model/prediction/{prediction_id}"
headers = { "Authorization": "Bearer $ATLASCLOUD_API_KEY" }

while True:
    response = requests.get(url, headers=headers)
    result = response.json()
    status = result["data"]["status"]
    print(f"Status: {status}")

    if status in ["completed", "succeeded"]:
        output_url = result["data"]["outputs"][0]
        print(f"Output URL: {output_url}")
        break
    elif status == "failed":
        print(f"Error: {result['data'].get('error', 'Unknown')}")
        break

    time.sleep(3)

Valori di stato

processingLa richiesta è ancora in fase di elaborazione.

completedGenerazione completata. Gli output sono disponibili.

succeededGenerazione riuscita. Gli output sono disponibili.

failedLa generazione è fallita. Controlla il campo errore.

Risposta completata

{
  "data": {
    "id": "pred_abc123",
    "status": "completed",
    "outputs": [
      "https://storage.atlascloud.ai/outputs/result.mp4"
    ],
    "metrics": {
      "predict_time": 45.2
    },
    "created_at": "2025-01-01T00:00:00Z",
    "completed_at": "2025-01-01T00:00:10Z"
  }
}

Carica file

Carica file nello storage Atlas Cloud e ottieni un URL utilizzabile nelle tue richieste API. Usa multipart/form-data per il caricamento.

POST/api/v1/model/uploadMedia

Esempio di caricamento

import requests

url = "https://api.atlascloud.ai/api/v1/model/uploadMedia"
headers = { "Authorization": "Bearer $ATLASCLOUD_API_KEY" }

with open("image.png", "rb") as f:
    files = {"file": ("image.png", f, "image/png")}
    response = requests.post(url, headers=headers, files=files)

result = response.json()
download_url = result["data"]["download_url"]
print(f"File URL: {download_url}")

Risposta

{
  "data": {
    "download_url": "https://storage.atlascloud.ai/uploads/abc123/image.png",
    "file_name": "image.png",
    "content_type": "image/png",
    "size": 1024000
  }
}

Schema di input

I seguenti parametri sono accettati nel corpo della richiesta.

Totale: 7Obbligatorio: 3Opzionale: 4

modelstringrequired

model name

Default: "google/gemini-omni-flash/video-edit"

videostringrequired

The source video to edit. Limited to 100MB and 30 seconds duration.

Format: uri

promptstringrequired

Text prompt describing the edit to apply to the source video (e.g., add, remove, or transform elements). Maximum 20,000 characters.

imagesarray[string]

Images to use as character, scene, or style references. Accepts 1 to 5 images when combined with a video reference. Supported formats: PNG, JPEG, JPG, WebP. Each image is limited to 20MB. Supports both a public URL and a base64-encoded image for each item.

Min items: 1Max items: 5

resolutionstring

The resolution of the generated video.

Default: "720p"

720p

thinking_levelstring

Controls the amount of internal reasoning the model performs before generating a response. Higher levels may improve quality on complex tasks but increase latency.

Default: "default"

defaulthighlow

seedinteger

The random seed to use for the generation. -1 means a random seed will be used.

Default: -1

Esempio di corpo della richiesta

{
  "model": "google/gemini-omni-flash/video-edit",
  "video": "example_video",
  "prompt": "A beautiful landscape",
  "resolution": "720p",
  "thinking_level": "default",
  "seed": -1
}

Schema di output

L'API restituisce una risposta di previsione con gli URL degli output generati.

codeinteger

HTTP status code of the response.

messagestring

Human-readable message; non-empty on failure.

dataobject

Esempio di risposta

{
  "id": "pred_abc123",
  "status": "completed",
  "model": "model-name",
  "outputs": [
    "https://storage.atlascloud.ai/outputs/result.mp4"
  ],
  "metrics": {
    "predict_time": 45.2
  },
  "created_at": "2025-01-01T00:00:00Z",
  "completed_at": "2025-01-01T00:00:10Z"
}

Atlas Cloud Skills

Atlas Cloud Skills integra oltre 400 modelli di IA direttamente nel tuo assistente di codifica IA. Un comando per installare, poi usa il linguaggio naturale per generare immagini, video e chattare con LLM.

Client supportati

Claude Code

OpenAI Codex

Gemini CLI

Cursor

Windsurf

VS Code

Trae

GitHub Copilot

Cline

Roo Code

Amp

Goose

Replit

40+ client supportati

Installa

npx skills add AtlasCloudAI/atlas-cloud-skills

Configura chiave API

Ottieni la tua chiave API dalla dashboard di Atlas Cloud e impostala come variabile d'ambiente.

export ATLASCLOUD_API_KEY="your-api-key-here"

Funzionalità

Una volta installato, puoi usare il linguaggio naturale nel tuo assistente IA per accedere a tutti i modelli Atlas Cloud.

Generazione di immaginiGenera immagini con modelli come Nano Banana 2, Z-Image e altri.

Creazione di videoCrea video da testo o immagini con Kling, Vidu, Veo, ecc.

Chat LLMChatta con Qwen, DeepSeek e altri grandi modelli linguistici.

Caricamento mediaCarica file locali per la modifica di immagini e flussi di lavoro da immagine a video.

Scopri di più

github.com/AtlasCloudAI/atlas-cloud-skills

Server MCP

Il server MCP di Atlas Cloud collega il tuo IDE con oltre 400 modelli di IA tramite il Model Context Protocol. Funziona con qualsiasi client compatibile MCP.

Client supportati

Cursor

VS Code

Windsurf

Claude Code

OpenAI Codex

Gemini CLI

Cline

Roo Code

100+ client supportati

Installa

npx -y atlascloud-mcp

Configurazione

Aggiungi la seguente configurazione al file delle impostazioni MCP del tuo IDE.

{
  "mcpServers": {
    "atlascloud": {
      "command": "npx",
      "args": [
        "-y",
        "atlascloud-mcp"
      ],
      "env": {
        "ATLASCLOUD_API_KEY": "your-api-key-here"
      }
    }
  }
}

Strumenti disponibili

atlas_generate_imageGenera immagini da prompt testuali.

atlas_generate_videoCrea video da testo o immagini.

atlas_chatChatta con grandi modelli linguistici.

atlas_list_modelsEsplora oltre 400 modelli di IA disponibili.

atlas_quick_generateCreazione di contenuti in un solo passaggio con selezione automatica del modello.

atlas_upload_mediaCarica file locali per i flussi di lavoro API.

Scopri di più

github.com/AtlasCloudAI/mcp-server

API Schema

{
  "info": {
    "title": "AtlasCloud API",
    "version": "1.0.0",
    "description": "The AtlasCloud API."
  },
  "openapi": "3.0.0",
  "paths": {
    "/api/v1/model/generateVideo": {
      "post": {
        "requestBody": {
          "content": {
            "application/json": {
              "schema": {
                "$ref": "#/components/schemas/Input"
              }
            }
          },
          "required": true
        },
        "responses": {
          "200": {
            "content": {
              "application/json": {
                "schema": {
                  "$ref": "#/components/schemas/PredictionResponse"
                }
              }
            },
            "description": "The request status."
          }
        }
      },
      "x-api-name": "model_run"
    },
    "/api/v1/model/prediction/{request_id}": {
      "get": {
        "parameters": [
          {
            "in": "path",
            "name": "request_id",
            "required": true,
            "schema": {
              "description": "Request ID",
              "type": "string"
            }
          }
        ],
        "responses": {
          "200": {
            "content": {
              "application/json": {
                "schema": {
                  "$ref": "#/components/schemas/PredictionResponse"
                }
              }
            },
            "description": "Result of the request."
          }
        }
      },
      "x-api-name": "model_result"
    }
  },
  "components": {
    "schemas": {
      "Input": {
        "properties": {
          "model": {
            "type": "string",
            "description": "model name",
            "default": "google/gemini-omni-flash/video-edit"
          },
          "video": {
            "description": "The source video to edit. Limited to 100MB and 30 seconds duration.",
            "type": "string",
            "format": "uri",
            "x-ui-component": "uploader"
          },
          "prompt": {
            "description": "Text prompt describing the edit to apply to the source video (e.g., add, remove, or transform elements). Maximum 20,000 characters.",
            "type": "string"
          },
          "images": {
            "description": "Images to use as character, scene, or style references. Accepts 1 to 5 images when combined with a video reference. Supported formats: PNG, JPEG, JPG, WebP. Each image is limited to 20MB. Supports both a public URL and a base64-encoded image for each item.",
            "items": {
              "type": "string",
              "format": "uri"
            },
            "maxItems": 5,
            "minItems": 1,
            "type": "array",
            "x-ui-component": "uploaders"
          },
          "resolution": {
            "default": "720p",
            "description": "The resolution of the generated video.",
            "enum": [
              "720p"
            ],
            "type": "string",
            "x-placeholder": "Select one",
            "x-ui-component": "select"
          },
          "thinking_level": {
            "description": "Controls the amount of internal reasoning the model performs before generating a response. Higher levels may improve quality on complex tasks but increase latency.",
            "default": "default",
            "enum": [
              "default",
              "high",
              "low"
            ],
            "type": "string"
          },
          "seed": {
            "default": -1,
            "description": "The random seed to use for the generation. -1 means a random seed will be used.",
            "type": "integer"
          }
        },
        "required": [
          "model",
          "video",
          "prompt"
        ],
        "type": "object",
        "x-order-properties": [
          "model",
          "prompt",
          "video",
          "images",
          "thinking_level",
          "resolution",
          "seed"
        ]
      },
      "PredictionResponse": {
        "type": "object",
        "properties": {
          "code": {
            "description": "HTTP status code of the response.",
            "type": "integer"
          },
          "message": {
            "description": "Human-readable message; non-empty on failure.",
            "type": "string"
          },
          "data": {
            "type": "object",
            "properties": {
              "id": {
                "description": "Unique identifier for the prediction.",
                "type": "string"
              },
              "model": {
                "description": "Model ID used for the prediction.",
                "type": "string"
              },
              "outputs": {
                "description": "Array of URLs to the generated content. Null when status is not completed.",
                "type": "array",
                "items": {
                  "type": "string"
                },
                "nullable": true
              },
              "urls": {
                "description": "Object containing related API endpoints.",
                "type": "object",
                "properties": {
                  "get": {
                    "description": "URL to poll for the prediction result.",
                    "type": "string",
                    "format": "uri"
                  }
                }
              },
              "has_nsfw_contents": {
                "description": "Array of boolean values indicating NSFW detection for each output. Null if not applicable.",
                "type": "array",
                "items": {
                  "type": "boolean"
                },
                "nullable": true
              },
              "status": {
                "description": "Status of the task: created, processing, completed, timeout, or failed.",
                "type": "string"
              },
              "created_at": {
                "description": "ISO timestamp of when the request was created (e.g., \"2023-04-01T12:34:56.789Z\").",
                "format": "date-time",
                "type": "string"
              },
              "error": {
                "description": "Error message if the task failed, empty string otherwise.",
                "type": "string"
              },
              "error_code": {
                "description": "Error code if the task failed.",
                "type": "integer"
              },
              "executionTime": {
                "description": "Total execution time in milliseconds.",
                "type": "number"
              },
              "timings": {
                "description": "Detailed timing breakdown.",
                "type": "object",
                "properties": {
                  "inference": {
                    "description": "Inference time in milliseconds.",
                    "type": "number"
                  }
                }
              }
            }
          }
        }
      }
    },
    "securitySchemes": {
      "apiKeyAuth": {
        "in": "header",
        "name": "Authorization",
        "type": "apiKey"
      }
    }
  },
  "servers": [
    {
      "url": "https://api.atlascloud.ai"
    }
  ]
}

Template di prompt ottimizzato per LLM

# google/gemini-omni-flash/video-edit

> A natively multimodal Google DeepMind model that edits an existing video from a text prompt with optional reference images, applying scene-consistent changes and native audio while preserving the untouched footage.


## Overview

- **Submit endpoint (POST)**: `https://api.atlascloud.ai/api/v1/model/generateVideo` — start an async generation; returns a `prediction_id`
- **Poll endpoint (GET)**: `https://api.atlascloud.ai/api/v1/model/prediction/{prediction_id}` — poll this until the prediction finishes
- **Model ID**: `google/gemini-omni-flash/video-edit`


## API Information

This model can be used via our HTTP API or more conveniently via our client libraries.
See the input and output schema below, as well as the usage examples.


### Input Schema

The API accepts the following input parameters:

- **`model`** (`string`, _required_):
  model name
  - Default: `"google/gemini-omni-flash/video-edit"`

- **`prompt`** (`string`, _required_):
  Text prompt describing the edit to apply to the source video (e.g., add, remove, or transform elements). Maximum 20,000 characters.

- **`video`** (`string`, _required_):
  The source video to edit. Limited to 100MB and 30 seconds duration.

- **`images`** (`array[string]`, _optional_):
  Images to use as character, scene, or style references. Accepts 1 to 5 images when combined with a video reference. Supported formats: PNG, JPEG, JPG, WebP. Each image is limited to 20MB. Supports both a public URL and a base64-encoded image for each item.
  - Min items: 1
  - Max items: 5

- **`thinking_level`** (`string`, _optional_):
  Controls the amount of internal reasoning the model performs before generating a response. Higher levels may improve quality on complex tasks but increase latency.
  - Default: `"default"`
  - Options: "default", "high", "low"

- **`resolution`** (`string`, _optional_):
  The resolution of the generated video.
  - Default: `"720p"`
  - Options: "720p"

- **`seed`** (`integer`, _optional_):
  The random seed to use for the generation. -1 means a random seed will be used.
  - Default: `-1`



**Required Parameters Example**:

```json
{
  "model": "google/gemini-omni-flash/video-edit",
  "video": "",
  "prompt": ""
}
```


**Full Example**:

```json
{
  "model": "google/gemini-omni-flash/video-edit",
  "prompt": "",
  "video": "",
  "images": [
    ""
  ],
  "thinking_level": "default",
  "resolution": "720p",
  "seed": -1
}
```


### Output Schema

The API returns the following output format:


- **`code`** (`integer`, _optional_):
  HTTP status code of the response.

- **`message`** (`string`, _optional_):
  Human-readable message; non-empty on failure.

- **`data`** (`object`, _optional_):
  - Properties:
    - **`id`** (`string`, _optional_):
      Unique identifier for the prediction.

    - **`model`** (`string`, _optional_):
      Model ID used for the prediction.

    - **`outputs`** (`array[string]`, _optional_):
      Array of URLs to the generated content. Null when status is not completed.

    - **`urls`** (`object`, _optional_):
      Object containing related API endpoints.
      - Properties:
        - **`get`** (`string`, _optional_):
          URL to poll for the prediction result.


    - **`has_nsfw_contents`** (`array[boolean]`, _optional_):
      Array of boolean values indicating NSFW detection for each output. Null if not applicable.

    - **`status`** (`string`, _optional_):
      Status of the task: created, processing, completed, timeout, or failed.

    - **`created_at`** (`string`, _optional_):
      ISO timestamp of when the request was created (e.g., "2023-04-01T12:34:56.789Z").

    - **`error`** (`string`, _optional_):
      Error message if the task failed, empty string otherwise.

    - **`error_code`** (`integer`, _optional_):
      Error code if the task failed.

    - **`executionTime`** (`number`, _optional_):
      Total execution time in milliseconds.

    - **`timings`** (`object`, _optional_):
      Detailed timing breakdown.
      - Properties:
        - **`inference`** (`number`, _optional_):
          Inference time in milliseconds.





**Example Response**:

```json
{
  "code": 0,
  "message": "",
  "data": {
    "id": "",
    "model": "",
    "outputs": [
      ""
    ],
    "urls": {
      "get": ""
    },
    "has_nsfw_contents": [],
    "status": "",
    "created_at": "",
    "error": "",
    "error_code": 0,
    "executionTime": 0,
    "timings": {
      "inference": 0
    }
  }
}
```


## Usage Examples

### cURL

```bash
# Step 1: Start generation (async)
curl -X POST "https://api.atlascloud.ai/api/v1/model/generateVideo" \
  -H "Authorization: Bearer $ATLASCLOUD_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
  "model": "google/gemini-omni-flash/video-edit",
  "prompt": "",
  "video": "",
  "images": [
    ""
  ],
  "thinking_level": "default",
  "resolution": "720p",
  "seed": -1
}'

# Response will contain: {"code": 200, "data": {"id": "prediction_id", "status": "processing"}}

# Step 2: Poll for result (replace {prediction_id} with the id returned above)
curl -X GET "https://api.atlascloud.ai/api/v1/model/prediction/{prediction_id}" \
  -H "Authorization: Bearer $ATLASCLOUD_API_KEY"

# Keep polling until status is "completed", "succeeded" or "failed"
# When completed, outputs will contain the generated content URL(s)
```

## Additional Resources

### Documentation

- [Model Playground](https://www.atlascloud.ai/models/google/gemini-omni-flash/video-edit)

Change the overall color of the headphones to red.

Caricamento...

Gemini Omni Flash — Video Edit

Model ID: google/gemini-omni-flash/video-edit

Gemini Omni Flash is Google DeepMind's high-performance, natively multimodal model built for high-speed video generation, editing, and cinematic control. This variant accepts a source video plus a text prompt (and, optionally, reference images), transforming an existing clip according to your instructions.

Overview

Gemini Omni Flash (gemini-omni-flash-preview) was introduced by Google alongside Nano Banana 2 Lite as a new generation of multimodal media models. Unlike traditional pipelines that stitch modalities together, Omni Flash is a single transformer that processes text, images, audio, and video simultaneously, producing output that is more cohesive, consistent, and controllable.

What sets it apart from earlier video models (such as the Veo family) is that it natively generates audio with every video — dialogue, ambience, music, and sound design are produced together with the picture rather than added afterward. The model is grounded in Gemini's real-world knowledge, so it reasons about physics, narrative logic, culture, and visual composition to produce results that feel intentional and cinematic. Generated media carries an invisible SynthID watermark.

AtlasCloud exposes Gemini Omni Flash through four endpoints — text-to-video, image-to-video, reference-to-video, and video-edit. All four route to the same gemini-omni-flash-preview model and differ only by the input modality they accept, corresponding to the model's task parameter (text_to_video, image_to_video, reference_to_video, edit). This endpoint maps to edit.

Inputs

This variant takes a source video and a text prompt, with optional reference images. The prompt describes the edit to apply — adding, removing, or transforming elements, restyling, or changing the audio — while Omni Flash preserves the rest of the clip. Because the model understands the whole scene, edits stay consistent with the surrounding footage rather than looking pasted on.

Video — the source clip to edit. Up to 100 MB and 30 seconds in duration.
Prompt — Natural-language description of the edit to apply (up to 20,000 characters).
Images (optional) — 1 to 5 reference images to guide the edit (e.g. a subject, object, or style to introduce). PNG/JPEG/JPG/WebP, ≤20 MB each. URL or base64.

Key Capabilities

Instruction-driven editing — Add, remove, replace, or restyle elements of a clip from a plain-language description.
Scene-consistent results — Edits blend into the existing footage, preserving untouched regions, lighting, and motion.
Reference-guided edits — Optionally supply up to 5 images to introduce a specific subject, object, or style.
Native audio generation — Edits can regenerate or adjust the accompanying soundtrack alongside the picture.
World-grounded realism — Physics, motion, and scene dynamics informed by Gemini's real-world knowledge.
Adjustable reasoning — The thinking_level control trades latency for quality on complex edits.
Reproducible results — Set a fixed seed to reproduce or iterate on a specific generation.

Input Parameters

Parameter	Type	Required	Default	Description
`model`	string	Yes	`google/gemini-omni-flash/video-edit`	Model identifier
`prompt`	string	Yes	—	Text description of the edit to apply. Max 20,000 characters.
`video`	string (uri)	Yes	—	Source video to edit. ≤100 MB and ≤30 seconds.
`images`	array of string (uri)	No	—	1–5 optional reference images for character, scene, or style. PNG/JPEG/JPG/WebP, ≤20 MB each. URL or base64.
`thinking_level`	string	No	`default`	Internal reasoning effort. Enum: `default`, `high`, `low`.
`resolution`	string	No	`720p`	Output resolution. Enum: `720p`.
`seed`	integer	No	`-1`	Random seed for reproducibility. `-1` uses a random seed.

Output duration and aspect ratio follow the source video, so this variant has no duration or aspect_ratio parameter.

Use Cases

Element edits — Add, remove, or replace objects, characters, or backgrounds in existing footage.
Restyling — Transform the look, color grade, or mood of a clip while keeping its motion.
Localization & cleanup — Swap on-screen elements or refresh assets without reshooting.
Reference-driven insertion — Bring a specific product or character (supplied as images) into an existing scene.
Iterative refinement — Apply successive edits to converge on a desired result.

Pricing

Billing is based on the duration of the source video, charged at a flat per-second rate.

SKU	Rate
Per second of source video	$0.14

Formula: clamp(video_duration, 3, 30) × $0.14

Billing follows the source video's duration, clamped to a 3-second minimum and a 30-second maximum.
Example: a 10-second source video costs 10 × $0.14 = $1.40.
Example: a 30-second source video costs 30 × $0.14 = $4.20.

Gemini Omni Flash Video Edit API by Google

INPUT

OUTPUT

Parametri

Esempio di codice

Installa

Autenticazione

Header HTTP

Invia una richiesta

Invia una richiesta

Corpo della richiesta

Risposta

Controlla lo stato

Esempio di polling

Valori di stato

Risposta completata

Carica file

Esempio di caricamento

Risposta

Schema di input

Esempio di corpo della richiesta

Schema di output

Esempio di risposta

Atlas Cloud Skills

Client supportati

Installa

Configura chiave API

Funzionalità

Server MCP

Client supportati

Installa

Configurazione

Strumenti disponibili

API Schema

Template di prompt ottimizzato per LLM

Gemini Omni Flash — Video Edit

Overview

Inputs

Key Capabilities

Input Parameters

Use Cases

Pricing

Esplora Modelli Simili

Gemini Omni Flash Image-to-Video Developer

Gemini Omni Flash Text-to-Video Developer

Veo 3.1 Lite Text-to-video

Veo 3.1 Lite Start-End Frame to Video

Veo 3.1 Lite Image-to-video

Veo3.1 Fast Image-to-video

Veo3.1 Fast Text-to-video

Veo3.1 Image-to-video

Veo3.1 Reference-to-video

Veo3.1 Text-to-video

Gemini Omni Flash Reference-to-Video

Gemini Omni Flash Image-to-Video

Gemini Omni Flash Text-to-Video

Gemini Omni Flash Reference-to-Video Developer

Sync.so Lipsync v3

VEED Lipsync

Un'unica API per tutta l'IA multimediale.

Join our Discord community

INPUT

OUTPUT

Parametri

Esempio di codice

Installa

Autenticazione

Header HTTP

Invia una richiesta

Invia una richiesta

Corpo della richiesta

Risposta

Controlla lo stato

Esempio di polling

Valori di stato

Risposta completata

Carica file

Esempio di caricamento

Risposta

Schema di input