
The LTX-2 API brings Lightricks' open-weight video family into production, generating video and synchronized audio in a single pass. Prompt from text, animate a still image, or extend an existing clip, and every result arrives with matched lip sync, foley, and ambient sound. On Atlas Cloud you reach it through one OpenAI-compatible key with transparent pay-as-you-go pricing and Day-0 access. Start building today.
Atlas Cloud provides you with the latest industry-leading creative models.
Compare every Ltx-2 API modality in one view, from prompt-driven generation to animating stills and continuing existing clips, each with optional synchronized audio.
| Modality | Description |
|---|---|
| Ltx-2 T2V API (Text to Video) | Feed a text prompt of up to 2,000 characters and the Ltx-2 API returns a fully rendered video, with optional synchronized audio when generate_audio is enabled. Clip length scales from 25 up to 257 frames, and six resolution presets cover square, portrait 9:16, and landscape 16:9 outputs. This endpoint suits ad concepts, social shorts, and storyboard previews built entirely from a written description, priced at a standard $0.002 base rate. |
| Ltx-2 I2V API (Image to Video) | Supply a still image in jpg, png, webp, gif, or avif alongside a motion prompt, and this endpoint animates it into a moving clip with optional audio generation. The same 25 to 257 frame range and six aspect ratio presets apply, so a single source frame becomes a landscape or vertical video. Product shots, character portraits, and concept art are natural inputs, all at the standard $0.002 base price. |
| Ltx-2 Extend Video API | When an existing clip needs more runtime, this endpoint continues it either forward or backward from a supplied mp4, mov, webm, m4v, or gif source. It stretches sequences up to 481 frames, preserves the original pace through match_input_fps, and can layer in generated audio. Reach for it to lengthen b-roll, loop scenes, or bridge cuts, billed at the same $0.002 base rate. |
The Ltx-2 API pairs text, image, and extend generation modes with native synchronized audio, controllable frame counts, and six aspect ratio presets, all reachable through one OpenAI-compatible key on transparent pay-as-you-go pricing.
Describe a scene in up to 2,000 characters and the Ltx-2 API turns pure text into a moving, sounding clip with no reference asset required. Your prompt drives subject, motion, lighting, and camera behavior, while an optional seed locks the result for repeatable iteration. Because everything starts from language, it suits storyboarding, concept pitches, and rapid creative exploration when you have an idea but no footage yet.
Feed a single still in jpg, png, webp, gif, or avif and the model animates it into video with matching audio. The prompt guides how the frame should move, from a slow push-in to a full action beat, while the original composition stays intact. Photographers, designers, and product teams lean on this to turn one hero image into shareable motion without reshooting anything.
Upload an existing clip and the Ltx-2 API continues it either forward or backward, generating up to 481 additional frames that inherit the original style and pace. The match_input_fps option keeps new footage locked to the source frame rate, so the joins stay invisible. This is how short takes grow into long-form sequences, and how you rescue a scene that ended a beat too early.
Frame count is fully controllable from 25 up to 257 frames in text and image modes, always set in multiples of eight plus one for clean sampling. Six resolution presets cover square HD at 1024 by 1024, a 512 square, portrait 3:4, portrait 9:16, landscape 4:3, and landscape 16:9, so a single model serves cinema, social, and vertical feeds. Choose the shape first, then let one request return the right crop.
Every request travels through one OpenAI-compatible key on Atlas Cloud, with transparent pay-as-you-go pricing and no subscription. Pass a fixed seed alongside the same prompt and the model returns the same clip, which turns unpredictable generation into a controllable, versionable step. Day-0 access means new Ltx-2 API releases are callable the moment they ship, so teams can standardize a video pipeline and scale it without managing GPUs.
Send a single identical prompt to the Ltx-2 API and two of the most popular video models on Atlas Cloud, then compare how each one handles motion, synchronized audio, and cinematic detail from the exact same instructions.
A curious macaque monkey darts through a crowded Marrakech spice market at golden hour and snatches a merchant's bright red fez off a stall. It bounds across open sacks of saffron and paprika, kicking up swirling orange dust clouds, then freezes on top of a canopy pole and peeks back just as the laughing merchant gives chase. Open on a low-angle tracking shot weaving behind the monkey through the crowd, whip pan to the merchant's startled face, then rise into a sweeping crane shot high above the stalls as dust glitters in the warm light. Realistic cinematic style, shallow depth of field, rich earthy color grade. Audio: lively market chatter, clinking brass, and rhythmic hand-drum percussion that builds into the chase. 16:9 aspect ratio.
Generated with Ltx 2.3 Quality on Atlas Cloud
Generated with Veo3.1 on Atlas Cloud
Generated with Seedance 2.0 on Atlas Cloud
A young sky-courier in a flowing crimson cloak leaps off the edge of a floating stone temple and rides a mechanical paper crane through canyon clouds, weaving between drifting glowing lanterns. She banks hard around a jagged rock spire, snatches a shimmering letter out of the air, then tucks into a steep dive as a rival glider bursts through the mist right behind her. Begin on a first-person POV over her shoulder as the crane launches, cut to a drone shot racing alongside the dive, then drop to a low-angle upward shot as she pulls out of the descent into blinding sunlight. Vibrant anime style, cel-shaded, luminous magic-hour palette. Audio: rushing wind, fluttering paper wings, and a soaring orchestral string swell timed to the pull-up. 16:9 aspect ratio.
Generated with Ltx 2.3 Quality on Atlas Cloud
Generated with Veo3.1 on Atlas Cloud
Generated with Seedance 2.0 on Atlas Cloud
From text and image prompts to extended cuts and synchronized sound, the Ltx-2 API covers the full path from first draft to 4K delivery across marketing, e-commerce, and creative production.
A single written prompt lets the Ltx-2 API output native 4K video with audio generated inline. Studios and ad teams convert campaign briefs into finished social spots without cameras or manual editing.
Upload a single still and image-to-video generation sets it in motion, preserving the source composition while adding lifelike camera movement. E-commerce sellers and illustrators bring static catalogs and portfolios to life for scroll-stopping feeds.
Need a longer take? The extend-video endpoint appends fresh frames to an existing clip while keeping motion, style, and audio continuous. Creators stitch short generations into seamless narratives that outrun a single request.
When sound matters, the Ltx-2 API synthesizes dialogue, foley, and ambient audio in one pass with the visuals, locked to motion. It fits explainer videos, trailers, and scenes needing sound on cue.
For final delivery, the Ltx-2 API renders native 4K frames with clean motion and sharp textures rather than upscaled previews. Agencies and studios export broadcast-ready masters straight from the endpoint into their editing pipelines.
Because LTX-2 is an efficient open model, high-volume jobs run on transparent pay-as-you-go pricing while you switch between performance modes per request. Product teams prototype cheaply, then raise fidelity for release.
Weigh the Ltx-2 API against comparable text and image driven video models on Atlas Cloud across generation modes, resolution, native audio, clip length, and starting price.
| Model | Generation Modes | Max Resolution | Synchronized Audio | Max Clip Length | Starting Price |
|---|---|---|---|---|---|
| LTX-2.3 Quality Text-to-Video | Text to video | 1024×1024 presets (16:9 default) | √ (optional toggle) | Up to 257 frames | $0.002 |
| LTX-2.3 Quality Image-to-Video | Image to video, extend | 1024×1024 presets (16:9 default) | √ (optional toggle) | Up to 257 frames | $0.002 |
| Seedance 2.0 Text-to-Video | Text to video | Up to 2K | √ native | 4 to 15 s | $0.112 |
| Veo 3.1 Text-to-Video | Text to video | Up to 4K | √ native | 8 s (extendable) | $0.2 |
| Kling Video O3 4K Text-to-Video | Text and image to video | Native 4K (3840×2160) | √ native | 3 to 15 s | $0.42 |
Get started in minutes — follow these simple steps to integrate and deploy models through Atlas Cloud's platform.
Sign up at atlascloud.ai and complete verification. New users receive free credits to explore the platform and test models.
Combining the advanced Ltx-2 models with Atlas Cloud's GPU-accelerated platform provides unmatched performance, scalability, and developer experience.
Low Latency:
GPU-optimized inference for real-time reasoning.
Unified API:
Run Ltx-2, GPT, Gemini, and DeepSeek with one integration.
Transparent Pricing:
Predictable per-token billing with serverless options.
Developer Experience:
SDKs, analytics, fine-tuning tools, and templates.
Reliability:
99.99% uptime, RBAC, and compliance-ready logging.
Security & Compliance:
SOC 2 Type II, HIPAA alignment, data sovereignty in US.
The Ltx-2 API gives developers programmatic access to LTX-2, the open-weight audio-video foundation model from Lightricks. It turns text or image prompts into synchronized video and audio in a single generation pass, and on Atlas Cloud you reach it through one OpenAI-compatible endpoint with pay-as-you-go pricing.
LTX-2 covers text-to-video, image-to-video, and clip extension, so you can turn a written scene into motion, animate a still image, or lengthen a shot you already have. Every mode can render accompanying audio, which makes it a fit for short films, ads, product demos, and social content.
Yes. Audio and video are produced in one synchronized pass rather than stitched together afterward, spanning dialogue lip sync, sound effects, and ambient sound. On Atlas Cloud you enable it with the generate_audio parameter, which is off by default.
The LTX-2 model natively reaches up to 4K resolution at up to 50 fps, with clips running as long as 20 seconds across frame rates of 24, 25, 48, and 50. On Atlas Cloud the quality endpoints let you choose the aspect ratio and set clip length through a frame count between 25 and 257 frames.
Send a standard request to the Atlas Cloud endpoint with your prompt, a model such as ltx-2.3-quality/text-to-video, an aspect ratio, a frame count, and an optional seed for reproducible output. Because the endpoint is OpenAI-compatible, a single key works across the Ltx-2 API and every other model on the platform, so integration usually takes minutes.
Absolutely. The ltx-2.3-quality/extend-video model continues an existing LTX-2 clip while keeping motion and audio consistent beyond the original length. This helps when a single generation runs shorter than the duration a scene actually needs.
LTX-2 ships with open weights and is efficient enough to run on consumer-grade GPUs, so self-hosting is a genuine option. If you would rather skip GPU provisioning, queue management, and scaling, the hosted API delivers the same model with no infrastructure to maintain.
Atlas Cloud bills the Ltx-2 API on a pay-as-you-go basis with transparent per-call pricing and no subscription or minimum commitment. You pay only for what you generate, and the same key gives you Day-0 access to new model releases. Start building today.
LTX-2.3 is the current iteration of the LTX-2 family and the version Atlas Cloud serves through its quality text-to-video, image-to-video, and extend-video endpoints. It refines the original release while keeping the same synchronized audio-video approach, so you get the latest quality without changing how you call the model.