




The Youchuan API now covers both the V8.1 and V8.2 generations, spanning stills and motion. Prompt-driven modes return four variations at once, images render at native 2K HD with no upscaling pass, blends fuse up to five references, and a single still animates into five-second clips at 480p or 720p. Atlas Cloud serves it through one OpenAI-compatible key with Day-0 access and per-call pricing from $0.086. Start building today.
Atlas Cloud provides you with the latest industry-leading creative models.
Compare every modality in the Youchuan API lineup side by side, with the V8.1 and V8.2 endpoints, output counts, resolutions, and standard per-call rates in a single view.
| Modality | Description |
|---|---|
| Youchuan T2I API (Text to Image) | Text prompts go in and four finished images come back, with optional native 2K HD, a style reference, and aspect-ratio, stylize, chaos, and weird controls on every run. The V8.1 and V8.2 endpoints both cost $0.086 per call, which suits concept art, campaign visuals, and fast moodboards. |
| Youchuan I2I API (Image to Image) | Already have a frame worth building on? This endpoint re-imagines an input image under the guidance of a text prompt and returns four variations, carrying over native 2K HD, style reference, and the same stylize, chaos, and weird controls. V8.1 and V8.2 are each priced at $0.086 per call, so iterating on a layout or testing alternate art directions stays cheap. |
| Youchuan I2V API (Image to Video) | One still image animates into four five-second videos at either 480p or 720p. Four takes per run mean you can keep the motion that actually fits the shot, which works for social loops, film pre-visualization, and adding movement to static key art. Both versions run at $0.086 per call. |
| Youchuan Blend API (Multi-Image Fusion) | Supply two to five source images and the model fuses them into four combined results at native 2K HD, with an optional prompt to steer the direction of the mix. Compositing, mood mixing, and pulling scattered references into one frame all belong here, at $0.086 per call on V8.1 and V8.2 alike. |
| Youchuan Style Transfer API (Image Retexture) | Retexture rewrites the artistic style of an input image while holding its composition in place, returning four restyled variations. At $0.129 per call on either version, it fits rebranding passes, art-direction tests, and giving a mixed photo set one shared look. |
| Youchuan Remove Background API (Background Removal) | Background removal is the most direct call in the family, turning any input image into a single result with the subject isolated on transparency. Product listings, marketing layouts, and compositing pipelines take that output directly, priced at $0.086 per call on both V8.1 and V8.2. |
Text-to-image, image-to-image, blending, retexture, background removal, and five-second animation all run through one Youchuan API key, now spanning both the V8.1 and V8.2 checkpoints, with native 2K HD, nine aspect ratios, and stylize, chaos, and weird controls billed per call.
Each prompt of up to 1024 characters returns four renders in one call, all carrying the family's signature saturated color and dramatic composition. Art directors compare several strong directions before locking a final frame.
Stylize spans 0 to 1000, chaos reaches 100, weird climbs to 3000, and quality toggles between 1 and 4, with seeds locked for reproducibility. Brand-safe output and wild experiments come from the same endpoint.
Flip the HD flag and the Youchuan API renders native 2K at 1.5x the base rate across nine aspect ratios from 1:1 to 21:9. Nothing is interpolated, so fine textures survive large-format print.
Blend fuses two to five source images into four cohesive results at native 2K, with an optional prompt steering how color, lighting, and subjects combine. Concept teams collapse scattered references into one on-brand frame.
Composition stays locked while retexture restyles an image into four variations, shifting only color, texture, and mood to match the prompt. At $0.129 per call, one approved layout can serve several campaign themes.
One still image and a short motion prompt yield four five-second clips at 480p or 720p, with motion set low or high. Static key art turns into social teasers or quick shot previews.
Each comparison below runs one identical prompt through the Youchuan API and two other image-to-video models available on Atlas Cloud, so differences in motion, scene coherence, and visual style show up in the footage itself rather than on a spec sheet.
A bike courier on a fixed-gear threads through a rain-soaked Hong Kong night market, neon signs bleeding across the wet asphalt. He clips a fruit stall and oranges scatter and bounce everywhere as the vendor throws up his hands; the courier skids sideways, plants a foot, kicks off hard, and shoots a narrow gap between two passing trams while sparks trail off the rails behind him. Open on a low tracking shot chasing the rear wheel through the puddles, whip pan to catch the oranges tumbling mid-air, then cut to a drone that cranes up over the market roof as the bike bursts out the far end onto an empty road. Volumetric neon light, mirror reflections in every puddle, shallow depth of field, gritty cinematic realism. Ambient market chatter, tram bells, and a driving synth pulse that swells at the escape. 16:9 aspect ratio.
Generated with Youchuan V8.2 on Atlas Cloud
Generated with Kling v3.0 Pro on Atlas Cloud
Generated with seedance 2.0 Mini on Atlas Cloud
A young sky-sailor grips the rigging of a small wooden airship as it dives through a canyon of floating islands at golden hour. A flock of crystal-winged birds scatters ahead of the bow, then one bird darts back and snatches the brass compass straight out of her hand. She lunges along the boom, catches the bird by a tailfeather, and the two of them tumble off the deck together before her glider snaps open and she banks hard around a waterfall spilling into open sky. Begin with a sweeping orbit around the airship, cut to a first-person POV as she lunges down the boom, then drop to a low-angle upshot as the glider unfurls against the sun. Painted anime style with hand-drawn cel shading and lush watercolor skies. Whooshing wind, creaking timber, and a soaring orchestral swell at the moment the glider releases. 16:9 aspect ratio.
Generated with Youchuan V8.2 on Atlas Cloud
Generated with Seedance 2.0 Mini on Atlas Cloud
Generated with Grok Imagine Video on Atlas Cloud
Campaign hero art, product cutouts, pre-visualization frames, restyled assets, and five-second social clips all come out of the same Youchuan API, whether you are prompting from scratch or animating an approved still.
One prompt returns four renders, and the HD flag pushes them to native 2K while stylize, chaos, and aspect ratio steer the mood. Marketing teams walk into reviews with several directions instead of one.
When a product shot has to sit on any background, Remove Background isolates the subject onto transparency for $0.086 per call. Storefronts and marketplace feeds get consistent cutouts without manual masking.
Feed two to five reference images into Blend and the model fuses them into four cohesive 2K frames, guided by an optional prompt. Art departments pitch a whole world before anything gets built.
Style Transfer repaints an approved image into four variations at $0.129 per call while its composition stays untouched. That lets one asset carry a summer look, a holiday look, and everything between.
Why restart a layout that almost works? Image-to-Image turns your draft into four refined variations, steered by a text prompt and an optional style reference, so designers explore directions while keeping what already works.
Any still can become four five-second clips at 480p or 720p, with motion set low or high and a short prompt describing the camera move. Social teams fill feeds that reward movement.
Line up the Youchuan API beside Kling, Seedance, Veo, and Grok Imagine on accepted inputs, resolution, clip length, audio, and list price, then pick the endpoint that fits the shot you need.
| Model | Inputs Accepted | Resolution | Clip Length | Native Audio | Atlas Cloud List Price |
|---|---|---|---|---|---|
| Youchuan V8.2 Image-to-Video | A start image plus an optional motion prompt such as a camera move | 480p or 720p, where 720p costs 3x the 480p rate | Four 5-second clips come back from a single call | - | $0.086 |
| Youchuan V8.1 Image-to-Video | Same pairing of start image and optional prompt as V8.2 | 480p or 720p at the identical 3x ratio for the higher tier | Also four takes of 5 seconds each per request | - | $0.086 |
| Kling V3.0 Turbo Image-to-Video | One JPG or PNG start frame up to 10MB, guided by a positive prompt | 720p or 1080p, with 1080p as the default | Anywhere from 3 to 15 seconds | - | $0.112 |
| Seedance 2.0 Fast Image-to-Video | First frame, an optional last frame, and a motion prompt | 480p and 720p, rising to 720p-SR, 1080p-SR, and 1440p-SR | 4 to 15 seconds, or -1 to let the model choose | √ | $0.09 |
| Veo3.1 Fast Image-to-video | Start image, optional end frame, prompt and negative prompt | 720p, 1080p, or 4K output | 4, 6, or 8 seconds, though 1080p and 4K require 8 | √ | $0.08 |
| Grok Imagine Video v1.5 Image-to-Video | A starting-frame HTTPS URL or data URI with a natural language prompt | 480p, 720p, or 1080p | 1 to 15 seconds, defaulting to 8 | - | $0.08 |
Get started in minutes — follow these simple steps to integrate and deploy models through Atlas Cloud's platform.
Sign up at atlascloud.ai and complete verification. New users receive free credits to explore the platform and test models.
Combining the advanced Youchuan models with Atlas Cloud's GPU-accelerated platform provides unmatched performance, scalability, and developer experience.
Low Latency:
GPU-optimized inference for real-time reasoning.
Unified API:
Run Youchuan, GPT, Gemini, and DeepSeek with one integration.
Transparent Pricing:
Predictable per-token billing with serverless options.
Developer Experience:
SDKs, analytics, fine-tuning tools, and templates.
Reliability:
99.99% uptime, RBAC, and compliance-ready logging.
Security & Compliance:
SOC 2 Type II, HIPAA alignment, data sovereignty in US.
The Youchuan API gives developers programmatic access to Youchuan, Midjourney's official Chinese platform, through a single Atlas Cloud key. It covers six generation modes across two model generations, V8.1 and V8.2, spanning text-to-image, image-to-image, image-to-video, multi-image blending, style transfer, and background removal. Billing is pay-as-you-go per call, so there is no subscription to sign before your first request.
Six modes ship in the family, and each one is available on both V8.1 and V8.2. Text-to-image, image-to-image, and blend return four candidates per run, style transfer repaints an input while holding its composition, image-to-video animates a still into four five-second clips, and background removal returns one transparent cutout. Blend accepts two to five source images in a single call.
V8.1 shipped on April 30, 2026 as the fastest model in the lineup, roughly four to five times quicker than earlier versions, and it introduced native 2K rendering through the hd flag. V8.2 followed on July 27, 2026 with an aesthetics and image quality pass: renders skew bolder and more refined, and low-quality outliers appear far less often. Both versions remain callable, so you can pin V8.1 for consistency with existing work or move to V8.2 for the newer look.
Create an Atlas Cloud account, generate one API key, then send a request naming the model ID for the mode and version you want, such as youchuan/v8.2/text-to-image. Rates are transparent and charged per call, and Day-0 access means each new version is callable as soon as it lands. Start building today.
Image modes default to 1:1 and support nine aspect ratios, from square through cinematic 21:9 and tall 9:21. Setting hd to true produces native 2K output instead of an upscaling pass, billed at 1.5x the base rate. On the motion side, image-to-video returns four five-second clips at either 480p or 720p, with motion set to low or high.
Prompt-driven modes expose the controls Midjourney users already know: aspect_ratio, stylize from 0 to 1000, chaos up to 100, weird up to 3000, and quality at either 1 or 4. Prompts accept up to 1024 characters, and an optional sref image URL steers the aesthetic without dictating composition. Style transfer and background removal are deliberately simpler, taking an input image plus a target style prompt for the former.
Generation is asynchronous. Your call returns a request_id, and polling the prediction endpoint reports status as created, processing, completed, or failed, with the outputs array holding result URLs once the job finishes. Set enable_base64_output to true when you would rather receive the payload inline than fetch hosted files.
Most modes on the Youchuan API run $0.086 per call, covering text-to-image, image-to-image, blend, background removal, and image-to-video, while style transfer is priced at $0.129 per call. Two multipliers sit on top of those rates: hd images bill at 1.5x the base rate, and 720p video bills at 3x the 480p rate. Nothing is bundled into a plan, so you pay only for the calls you make.
Yes. Fix seed to a specific integer and the same prompt with the same settings reproduces that render exactly, whereas the default value of -1 randomizes every run. For a shared look across a campaign, pass one sref reference image with a steady stylize value and vary only the prompt.
Output shapes are fixed: clips run five seconds, blend takes at most five images, and prompts cut off at 1024 characters. Style transfer exposes no aspect ratio, hd, or stylize parameters, so resolution and framing follow the input image. Since jobs are queued and polled rather than streamed, build status checks and retries into your pipeline instead of expecting a synchronous response.
Guides, tutorials, and product updates to help you get the most out of Atlas Cloud.