
Kling v2.6 is Kuaishou's video model family for cinematic text-to-video and image-to-video creation, AI avatars, and reference-driven motion transfer. It supports flexible aspect ratios, enhanced dynamics, stable identity, and temporal consistency across specialized workflows. Atlas Cloud provides Pro and Standard options through consistent endpoints with transparent pay-as-you-go pricing. Start building today.
Kling 2.6 is developed by Kuaishou. Atlas Cloud (operated by Atlas Cloud AI LLC) provides access to it and does not own it. All trademarks belong to their respective owners.
Match each Kling v2.6 endpoint to its supported input workflow, production strengths, and intended video use cases.
| Modality | Description |
|---|---|
| Kling v2.6 Pro Avatar API (Avatar To Video) | Designed for premium avatar generation, this endpoint produces high-quality videos with clean detail, stable motion, and strong identity consistency. It suits polished profile videos, introductions, and social content. |
| Kling v2.6 Std Avatar API (Avatar To Video) | Choose the Standard Avatar endpoint to create AI avatar videos with clean detail, cinematic motion, and reliable prompt adherence. Its capabilities fit profiles, short introductions, and recurring social media production. |
| Kling v2.6 Pro Motion Control API (Reference Motion To Video) | Reference motion clips transfer dance, action, or gestures to a character image or source video while preserving identity and temporal consistency. Use it for smooth character animation that follows specific recorded movements. |
| Kling v2.6 Std Motion Control API (Image And Motion To Video) | Animate a still character image by supplying a motion clip containing a dance, action, or gesture. The endpoint extracts that movement and generates smooth, realistic video for accessible motion transfer workflows. |
| Kling v2.6 Pro API (Text To Video) | Text prompts become cinematic videos with generated sound and flexible aspect ratios through this Kuaishou model. It supports concept development, narrative scenes, advertisements, and other prompt-driven video projects. |
| Kling v2.6 Pro API (Image To Video) | Starting from an image, this endpoint creates cinematic video with sound generation and enhanced dynamics. Bring product visuals, illustrations, or creative stills to life when controlled visual continuity matters. |
Kling v2.6 combines native audiovisual generation, text and image workflows, reference driven motion control, audio driven avatars, flexible parameters, and pay as you go access across Standard and Pro endpoints.
Kling v2.6 Pro can generate video and sound together from either a text prompt or a source image. The sound switch is enabled by default on Atlas Cloud's text to video and image to video endpoints. Dialogue, effects, and ambience can therefore follow the scene as it unfolds. This is a strong fit for short narrative clips that need a coherent audiovisual result without a separate sound pass.
Start from a written scene or animate a still image with the Pro generation endpoints. Text prompts support 1:1, 9:16, and 16:9 outputs, while both paths offer 5 or 10 second durations, a sound toggle, negative prompts, and CFG control from 0 to 1. Image inputs accept JPG, JPEG, or PNG files up to 10 MB, giving developers practical control over social, product, and cinematic assets.
Bring a character image and a reference motion clip to Kling v2.6 Motion Control, which transfers full body dance, action, or gesture sequences while protecting identity and temporal consistency. Reference videos can span 3 to 30 seconds. Choose whether composition follows the image or the motion video, and decide whether to retain source audio. The result suits continuous performances, character animation, and action driven campaigns.
An image and an audio track are the required inputs for the Standard and Pro Avatar endpoints, with an optional prompt to guide the result. Pro prioritizes clean detail, stable motion, and strong identity consistency, while Standard pairs cinematic motion with reliable prompt adherence. Use this path when a recognizable subject must deliver prepared audio in profile videos, branded introductions, or recurring social content.
Choose Standard or Pro endpoints by workload, then pay only for the calls you make through Atlas Cloud. Standard Avatar starts at $0.056 per call, Standard Motion Control at $0.07, Pro Text to Video and Image to Video at $0.07, and Pro Avatar or Motion Control at $0.112. One account spans all six endpoints, helping developers balance visual quality, motion needs, and production cost without a subscription commitment.
See how Kling v2.6 and two Atlas Cloud alternatives interpret identical prompts across realistic action and stylized storytelling.
An 8–10 second hyperreal fashion-commercial film set at noon on a vast white salt lake: six high-speed skaters in iridescent mirrored capes form a precise geometric “kite salt-harvesting team,” each pulling taut cobalt-blue and fluorescent-orange lines connected to a gigantic manta-ray kite. Open with a vertical top-down drone shot as the synchronized formation carves razor-sharp patterns through the salt crust; the drone suddenly dives and levels out inches above the ground, accelerating behind the leader as the team slaloms through fluttering flag gates, capes snapping and ropes flexing with realistic tension. Whip-pan into a fast 360-degree orbit when a sudden dust devil erupts beside them, spiraling loose salt crystals into the air; the six riders react in perfect unison, sharply reel in their lines, and the manta kite folds into a dramatic power dive, blasting through the crystal vortex and scattering glittering salt like a dazzling snowfall over the speeding team. Continuous coordinated action from beginning to end, coherent identities and formation across every shot, convincing skating momentum, physically accurate rope tension, aerodynamic fabric motion, granular salt particles, wind turbulence, and grounded shadows. Brutal overhead sunlight, crisp high-contrast top lighting, prismatic rainbow reflections across the mirrored capes and salt surface, layered snow-white, cobalt-blue, and fluorescent-orange composition, premium surreal high-fashion advertising aesthetic, razor-sharp cinematic detail, energetic percussion synchronized with every carve and camera transition, rushing wind, scraping blades, snapping fabric, tightening ropes, then a bright crystalline hiss at the finale. No slow motion, no static filler shots, no cuts that break spatial continuity, no deformed bodies, no duplicated people, no tangled or disconnected ropes, no inconsistent costumes, no floating objects, no screens, no software interfaces, no dashboards, no progress bars, no charts, no captions, no logos, no watermarks, no readable text. 16:9 aspect ratio.
Generated with Kling v2.6 Pro Text-to-Video on Atlas Cloud
Generated with Seedance 2.0 Text-to-Video on Atlas Cloud
Generated with Seedance 2.0 Fast Text-to-Video on Atlas Cloud
An 8–10 second continuous cinematic deep-sea chase inside the steeply tilted ballroom of a decaying shipwreck: begin with an extreme macro shot of a semi-transparent mimic octopus tentacle shimmering through cracked chandelier crystals as it snatches a glowing “pearl”; rapidly dolly-zoom backward to reveal the octopus swinging from a corroded railing over overturned dining tables while moray eels weave through chairs and pursue it, silverware, glass shards, sediment, and chains of bubbles rising and colliding with believable underwater physics. Transition into a tight 360-degree orbiting chase shot as the octopus continuously shifts its skin texture and translucency to match velvet curtains, tarnished brass, and barnacled wood, then blasts a dense ink cloud that curls into a convincing octopus-shaped decoy; the eels strike the false silhouette as the real octopus slips past them. Whip-pan and roll the camera into a dramatically canted top-down view: the stolen “pearl” suddenly opens a tiny eye, its amber glow intensifies, and the enormous camouflaged face of an anglerfish emerges beneath the banquet table—the pearl is its lure; the octopus freezes for one comic beat as the anglerfish lunges toward camera. Photorealistic deep-sea cinematic realism, cold cyan ambient light contrasted with warm amber bioluminescence, volumetric water, drifting particles, high-density soft-body tentacle motion, fluid simulation, floating-object collisions, seamless subject continuity, physically credible momentum, sharp readable action, no slow motion. Immersive synchronized audio: muffled hull groans, crystal clinks, rushing bubbles, eel snaps, pulsing chase percussion, a wet ink burst, then a sudden bass sting at the reveal. No screens, interfaces, dashboards, progress bars, charts, captions, subtitles, logos, watermarks, or explanatory text. 16:9 aspect ratio
Generated with Kling v2.6 Pro Text-to-Video on Atlas Cloud
Generated with Seedance 2.0 Text-to-Video on Atlas Cloud
Generated with Seedance 2.0 Fast Text-to-Video on Atlas Cloud
From cinematic prompts and animated campaign stills to consistent avatars and motion guided performances, Kling v2.6 supports developers creating films, advertisements, social content, and branded characters.
Turn written concepts into cinematic video with Kling v2.6 Pro Text to Video, using flexible aspect ratios and generated sound. Filmmakers can prototype trailers, dramatic beats, and visual pitches from a single prompt.
Bring a product image or campaign still to life with enhanced dynamics, cinematic quality, and generated sound. Marketing teams can create launch teasers, visual demonstrations, and social promotions from existing artwork.
Create polished avatar videos with clean detail, stable motion, and strong identity consistency through the Pro Avatar model. Use them for profile introductions, presenter segments, and recurring branded social personalities.
Map dance, action, or gesture footage onto a character image while preserving identity and temporal consistency. Creators can produce choreography tests, animated mascots, and performance driven social clips from recorded movement.
Give a character image or source video the movement from a reference clip through Pro Motion Control. Production teams can explore alternate performers, stylized action sequences, and consistent character animation.
Start with a text prompt and generate cinematic video with sound in flexible aspect ratios. Developers can automate short campaign concepts, creator content, and platform specific visuals without supplying an opening image.
Compare Kling v2.6 with leading text-to-video models across provider, clip duration, native audio, and standard base pricing.
| Model | Provider | Output Duration | Native Audio | Standard Base Price |
|---|---|---|---|---|
| Kling v2.6 Pro Text-to-Video | Kuaishou | 5 or 10 seconds | √ | $0.07/second |
| Seedance 2.0 Text-to-Video | ByteDance | 4 to 15 seconds | √ | $0.112/second |
| Wan-2.7 Text-to-video | Qwen | 2 to 15 seconds | √ | $0.10/second |
| Veo 3.1 Lite Text-to-video | 4, 6, or 8 seconds | √ | $0.05/second |
Get started in minutes — follow these simple steps to integrate and deploy models through Atlas Cloud's platform.
Sign up at atlascloud.ai and complete verification. New users receive free credits to explore the platform and test models.
Combining the advanced Kling 2.6 models with Atlas Cloud's GPU-accelerated platform provides unmatched performance, scalability, and developer experience.
Low Latency:
GPU-optimized inference for real-time reasoning.
Unified API:
Run Kling 2.6, GPT, Gemini, and DeepSeek with one integration.
Transparent Pricing:
Predictable per-token billing with serverless options.
Developer Experience:
SDKs, analytics, fine-tuning tools, and templates.
Reliability:
99.99% uptime, RBAC, and compliance-ready logging.
Security & Compliance:
SOC 2 Type II, HIPAA alignment, data sovereignty in US.
Kling v2.6 is a Kuaishou AI video model family available through Atlas Cloud. It includes Pro text-to-video and image-to-video generation, plus Standard and Pro endpoints for avatars and motion control.
Create videos from prompts, animate source images, produce avatar videos from an image and audio, or transfer dance, action, and gesture sequences from reference clips. The Pro text-to-video and image-to-video endpoints can generate sound alongside the visuals.
Choose Text-to-Video when starting from a prompt and Image-to-Video when animating a still image. Avatar endpoints pair an image with supplied audio, while Motion Control applies movement from a reference video to a character.
Send a POST request to `/api/v1/model/generateVideo` with your Atlas Cloud API key, the exact model ID, and the endpoint-specific inputs. The asynchronous response provides a prediction ID that you can check through `/api/v1/model/prediction/{id}` until the task completes.
Controls vary by endpoint. Text-to-video supports prompts, negative prompts, 1:1, 9:16, and 16:9 aspect ratios, 5 or 10 second durations, sound, and guidance scale, while the other workflows add image, audio, motion video, orientation, or sound-retention inputs as applicable.
For Pro text-to-video and image-to-video, the `sound` parameter controls simultaneous sound generation. Avatar models instead require an audio input with the character image, while Motion Control can retain the reference video's original sound.
Atlas Cloud lists standard rates of $0.07 per second for Pro Text-to-Video, Pro Image-to-Video, and Standard Motion Control. Standard Avatar is $0.056 per second, while Pro Avatar and Pro Motion Control are $0.112 per second, using original rates rather than promotional prices.
Prepare a JPG, JPEG, or PNG character image and an MP4 or MOV motion video. Each file must be no larger than 10 MB, measure at least 300 pixels in both dimensions, and use an aspect ratio between 1:2.5 and 2.5:1.
First, verify the model ID, Bearer authentication, required fields, and endpoint-specific media constraints. If the task was accepted, poll its prediction ID until it completes or fails and inspect the returned error. For visual inconsistencies, simplify the scene and use clear source media, especially when transferring facial motion or interactions with complex objects.
Guides, tutorials, and product updates to help you get the most out of Atlas Cloud.