
Kling 1.6 is Kuaishou's video generation model, built to turn prompts and reference images into coherent short-form video. It improves responsiveness to motion, temporal actions, and camera movement while strengthening subject consistency, color accuracy, lighting dynamics, and detailed rendering. Access the family through Atlas Cloud with one OpenAI-compatible key and transparent pay-as-you-go pricing. Start building today.
Kling 1.6 is developed by Kuaishou. Atlas Cloud (operated by Atlas Cloud AI LLC) provides access to it and does not own it. All trademarks belong to their respective owners.
See how each Kling 1.6 endpoint handles text or image inputs and where its verified strengths in motion, coherence, detail, and efficiency fit.
| Modality | Description |
|---|---|
| Kling 1.6 Multi I2V Pro API (Multi Image To Video) | Convert images into video featuring multiple subjects with improved coherence and advanced motion tracking accuracy. This Pro endpoint suits complex scenes where coordinated subject movement and visual consistency matter. |
| Kling 1.6 Multi I2V Standard API (Multi Image To Video) | For cost efficient multi subject generation, this endpoint turns images into video while balancing speed and detail. It fits basic scene animation and routine content production with multiple subjects. |
| Kling 1.6 T2V Standard API (Text To Video) | Text prompts become short form videos with stable motion and dependable prompt alignment. Use this entry level endpoint for concept visualization, social content, and straightforward prompt driven sequences. |
| Kling 1.6 I2V Pro API (Image To Video) | Starting from a still image, the Pro endpoint produces video with smoother motion blending and improved texture realism. It works well for polished product visuals, character shots, and cinematic image animation. |
| Kling 1.6 I2V Standard API (Image To Video) | Choose this lightweight endpoint to transform images into video through a foundational generation workflow. Its minimal cost positioning makes it suitable for simple animations, early creative tests, and higher volume production. |
Kling 1.6 combines text, single image, and up to four image reference workflows with 5 or 10 second outputs, controllable prompt adherence, optional negative prompts, selected aspect ratios, and Standard or Pro access through one Atlas Cloud API.
Kling 1.6 covers text to video and image to video generation through dedicated variants. Start with a written scene or animate a source image, then guide action with prompts up to 2,500 characters. Standard models prioritize cost efficiency, while the Pro image variant emphasizes smoother motion blending and more realistic textures. This range fits concept testing, social clips, and polished visual sequences.
Upload between one and four reference images with the Multi I2V Standard or Pro variant. The model uses them as visual references while a prompt directs subjects, movement, and scene development. Pro is tuned for stronger multi subject coherence and more accurate motion tracking. It suits character interactions, product combinations, and scenes that bring several visual elements together.
Set both a first image and an optional end image with the I2V Pro variant to shape where a clip begins and lands. Each image can be JPG, JPEG, or PNG, up to 10 MB and at least 300 by 300 pixels. Add motion instructions between those visual anchors. This control is especially useful for planned transitions, pose changes, and product reveals.
Smooth motion is a defining focus of the Pro image variant, with upgraded blending and improved texture realism recorded in its Atlas Cloud profile. A guidance scale from 0 to 1 adjusts prompt adherence, while negative prompts help steer away from unwanted elements. Use these controls to balance natural movement against tightly directed action. This combination suits demanding character, fabric, and camera movement.
Choose 5 or 10 second output across the Kling 1.6 family. Text and multi image variants also expose 16:9, 9:16, and 1:1 aspect ratios, so one workflow can target widescreen, vertical, or square placements. A 2,500 character prompt ceiling leaves room for subject, action, lighting, and camera direction. These options help teams plan platform specific assets without changing model families.
Run every listed Kling 1.6 variant through one Atlas Cloud video generation API with pay as you go billing. Standard variants use a verified original base price of $0.056 per run, while Pro variants use $0.098 per run. Pick the tier that matches the required balance of cost, detail, and motion quality. This setup supports quick experiments and production pipelines without separate provider integrations.
See how Kling 1.6 and two alternative video models interpret identical prompts across realistic action and stylized storytelling.
A cinematic 8–10 second miniature live-action sequence inside a glassblowing workshop at midnight: a young artisan continuously rotates a blowpipe as a white-hot glass bubble rapidly expands, its perfectly round form anchoring the composition. Begin with an extreme macro orbit around the spinning molten glass, capturing transparent amber filaments stretching, viscous surface tension, heat shimmer, tiny sparks, and realistic internal refraction. The swelling bubble suddenly slips from the pipe, strikes the silver-gray metal table, and rolls fast; drop into a table-level high-speed tracking shot alongside it as it wobbles, deforms, sheds glowing threads, and reflects cobalt-blue moonlight against the furnace-orange glow. Whip-pan to the alarmed artisan lunging across the bench and catching the runaway glass with a wet wooden paddle at the last instant—an explosive hiss sends physically accurate steam swirling through the frame. As the steam clears, reveal the glass miraculously frozen into a small transparent pufferfish with delicate glass fins and consistent internal bubbles; it gently puffs its cheeks once, a playful final beat. Seamless continuous motion, coherent subject transformation, realistic glass viscosity, collisions, thermal glow, sparks, refraction, caustics, heat distortion, and steam physics; layered amber, cobalt blue, and silver-gray palette; furnace orange key light, cool moon-blue rim light, shallow depth of field, tactile high-end practical miniature filmmaking, photoreal cinematic texture, no slow motion, no static filler, no screens, software interfaces, dashboards, progress bars, charts, captions, text, logos, or watermarks. Synchronized audio: roaring furnace, rotating pipe scrape, sharp metallic clink, rolling glass rattle, urgent footstep, loud wet hiss, then a tiny crystalline puff; tense percussive rhythm ending on a whimsical glass chime. 16:9 aspect ratio.
Generated with Kling v1.6 Multi i2v Pro on Atlas Cloud
Generated with Seedance 2.0 Image-to-Video on Atlas Cloud
Generated with Kling v1.6 i2v Standard on Atlas Cloud
A tense 9-second micro-story in the cramped back kitchen of a late-night Hong Kong noodle shop: a young chef urgently rescues a ramen order, moving continuously with precise, believable hand choreography. Begin with an extreme low-angle lateral tracking shot skimming across the flour-dusted cutting board as he snaps his wrists and throws a long bundle of noodles high into the air; chase the twisting strands through dense steam, with individual noodles flexing, stretching, and naturally occluding his hands and hanging cookware. Whip-pan to a tight stove-side angle as the noodle bundle nearly drops into a licking gas flame; at the last instant he lunges forward and catches it cleanly in a wire skimmer, the mesh bending under its weight, then pivots in one fluid motion and plunges the noodles into violently boiling broth, sending realistic droplets, bubbles, and oily ripples across the pot. Snap to an overhead top-down shot as the noodles unfurl into a neat blooming spiral in the soup; through the service hatch, waiting diners lean in and burst into delighted applause while the chef flashes a breathless grin. Maintain exact continuity of the same chef, clothing, utensils, noodle bundle, kitchen geography, and motion across every cut. Tight deep staging constantly coordinates the chef, noodles, flames, pots, and foreground utensils; tactile layers of airborne flour, wet tile reflections, glistening oil, condensation, and rolling steam. Photorealistic Hong Kong cinema aesthetic, handheld kinetic energy, crisp natural motion blur, warm tungsten-orange practical lights clashing with cool cyan ceramic tiles, rich contrast, subtle 35mm film grain, realistic skin and food texture. Synchronized sound: knife-board clatter, gas flame roar, rushing steam, skimmer clang, boiling broth splash, then a sharp burst of applause; fast percussive kitchen rhythm, no dialogue. No slow motion, frozen poses, empty establishing shots, montage gaps, jumpy continuity, extra fingers, malformed hands, duplicated limbs, broken utensils, clipping, teleporting noodles, rubbery motion, impossible fluid behavior, floating objects, UI, screens, dashboards, progress bars, charts, captions, subtitles, logos, or visible text. 16:9 aspect ratio.
Generated with Kling v1.6 Multi i2v Pro on Atlas Cloud
Generated with Seedance 2.0 Image-to-Video on Atlas Cloud
Generated with Kling v1.6 i2v Standard on Atlas Cloud
Kling 1.6 turns prompts and source images into short video concepts for storyboards, product campaigns, multi-subject scenes, animated artwork, social variations, and character narratives.
Turn written concepts into short-form video drafts with stable motion and dependable prompt alignment. Directors, agencies, and product teams can preview pacing, action, and visual direction before committing to full production.
Animate product stills with smoother motion blending and more realistic textures. Marketing teams can turn catalog imagery into polished launch clips, feature reveals, and campaign assets without arranging a new shoot.
Combine subjects from separate images while preserving stronger scene coherence and tracking their motion accurately. Build character interactions, ensemble moments, or lifestyle narratives for branded content and social campaigns at scale.
Animate static artwork with the Pro image-to-video variant's smoother motion blending and improved texture realism. Artists and game teams can create animated concepts, character beats, or atmospheric scene studies from existing visuals.
Start with a prompt or source image to produce short video concepts with stable or smoothly blended motion. Creators can develop multiple hooks, visual treatments, and campaign directions for social channels.
When several subjects must share one scene, the Multi variants prioritize coherence and advanced motion tracking. Use them for character pairings, pet interactions, group moments, or narrative tests built from images.
Compare Kling 1.6 variants with other video models available on Atlas Cloud across input workflows, clip duration, native audio, and standard pricing.
| Model | Input Workflow | Output Duration | Native Audio | Standard Price |
|---|---|---|---|---|
| Kling v1.6 Multi i2v Pro | Prompt + 1 to 4 images | 5 or 10 seconds | - | $0.098/run |
| Kling v1.6 Multi i2v Standard | Prompt + 1 to 4 images | 5 or 10 seconds | - | $0.056/run |
| Kling v1.6 t2v Standard | Text prompt | 5 or 10 seconds | - | $0.056/run |
| Kling v1.6 i2v Pro | Prompt + start frame + optional end frame | 5 or 10 seconds | - | $0.098/run |
| Kling v1.6 i2v Standard | Prompt + start frame | 5 or 10 seconds | - | $0.056/run |
| Wan-3.0 Image-to-video | Prompt + start frame + optional end frame | 2 to 30 seconds | √ | $0.05/second |
| MiniMax H3 Image-to-Video | Prompt + start frame + optional end frame | 4 to 15 seconds | √ | $0.038/second |
Get started in minutes — follow these simple steps to integrate and deploy models through Atlas Cloud's platform.
Sign up at atlascloud.ai and complete verification. New users receive free credits to explore the platform and test models.
Combining the advanced Kling 1.6 models with Atlas Cloud's GPU-accelerated platform provides unmatched performance, scalability, and developer experience.
Low Latency:
GPU-optimized inference for real-time reasoning.
Unified API:
Run Kling 1.6, GPT, Gemini, and DeepSeek with one integration.
Transparent Pricing:
Predictable per-token billing with serverless options.
Developer Experience:
SDKs, analytics, fine-tuning tools, and templates.
Reliability:
99.99% uptime, RBAC, and compliance-ready logging.
Security & Compliance:
SOC 2 Type II, HIPAA alignment, data sovereignty in US.
Kling 1.6 is a video generation model family developed by Kuaishou. On Atlas Cloud, it covers text to video, single image to video, and multi image to video through five Standard and Pro endpoints. Its variants offer stable motion, prompt alignment, smoother motion blending, texture realism, and multi subject coherence.
Turn text prompts into original video scenes, animate a single starting image, or combine multiple reference images in a multi subject composition. Depending on the endpoint, you can generate clips lasting 5 or 10 seconds.
Choose T2V Standard for text driven generation and I2V Standard for cost efficient animation from one image. I2V Pro prioritizes smoother motion blending and more realistic textures. For several subjects or reference images, select Multi I2V Standard or Multi I2V Pro according to your cost and coherence requirements.
Create an API key in the Atlas Cloud dashboard, then send a JSON request to POST /api/v1/model/generateVideo with your Bearer token. Include the exact model identifier, a prompt, and the required image or images for image based endpoints. Save the returned prediction ID and use it to retrieve the asynchronous result.
Every endpoint requires a model identifier and a prompt of up to 2,500 characters. Text to video supports aspect ratio, duration, guidance scale, and negative prompts, while single image endpoints require a JPG, JPEG, or PNG starting image. Multi image endpoints accept one to four reference images and also provide duration, aspect ratio, and negative prompt controls.
All five Atlas Cloud endpoints support 5 or 10 second generation, with 5 seconds as the documented default. Text to video and multi image endpoints accept 16:9, 9:16, or 1:1. Single image workflows derive their framing from the supplied source image instead of exposing the same aspect ratio parameter.
At standard list pricing, the T2V Standard, I2V Standard, and Multi I2V Standard endpoints cost $0.056 per generated video second. I2V Pro and Multi I2V Pro cost $0.098 per generated video second, with pay as you go billing.
First verify the Bearer token, exact model identifier, required prompt, and any required image fields. Check the documented prompt length, image format, file size, resolution, duration, and parameter ranges before resubmitting an invalid request. Because generation is asynchronous, retain the prediction ID and continue querying the result endpoint while its status is processing.
Guides, tutorials, and product updates to help you get the most out of Atlas Cloud.