
The HiDream O1 1.5 Image API brings HiDream.ai's unified foundation model to your stack, running text-to-image, single-image editing, and subject-driven personalization on one pixel-level system. Tune guidance and inference steps for strong prompt fidelity across six aspect ratio presets. Atlas Cloud delivers it through one OpenAI-compatible endpoint with transparent pay-as-you-go pricing at $0.044 per image. Start building today.
Compare what each route of the HiDream O1 1.5 Image API takes in, renders out, and charges per call.
| Modality | Description |
|---|---|
| HiDream O1 1.5 Text-to-Image API (Text To Image) | Turn a written prompt of up to 2,500 characters into a fully composed image across six presets, from a 512x512 square to 16:9 landscape, with PNG, JPEG, or WebP output. Denoising steps range from 1 to 100 and guidance scale from 1.0 to 20.0, so each request can trade speed against how tightly the result follows your prompt. At $0.044 per image, it fits e-commerce mockups, advertising concepts, and game art produced at volume. |
| HiDream O1 1.5 Edit API (Image Editing) | Feed one reference image URL alongside your instruction and this endpoint rewrites that image, or pass several URLs for subject-driven personalization across a set. It shares the same six size presets, 1 to 100 inference steps, and 1.0 to 20.0 guidance range as the text-to-image route, returning PNG, JPEG, or WebP. Billed at $0.044 per image, it handles product retouching, background swaps, and consistent character edits. |
The HiDream O1 1.5 Image API unifies text-to-image generation, instruction-based editing, and subject-driven personalization inside one pixel-native model that renders accurate bilingual text and hands developers direct control over guidance, sampling steps, and output format.

Send a prompt of up to 2,500 characters and the model renders it as a finished image through a single pixel-native transformer that encodes pixels, text, and task conditions in one shared space. Because no external VAE or separate text encoder sits in the path, fine detail and composition stay stable across dense, multi-clause descriptions. This makes it a dependable base for concept art, marketing visuals, and product mockups.

Few image models place legible words inside a composition, yet HiDream O1 1.5 renders Chinese, English, mixed-language strings, and numerical data cleanly enough to skip manual retouching. The pixel-native design handles multi-region layouts, keeping headlines, captions, and labels sharp where latent-space models often blur or garble type. Designers can draft posters, packaging, and social graphics whose text is ready to ship.

When you pass one reference image URL with a plain-language instruction such as remove the earphones, the edit endpoint applies the change while preserving the surrounding composition. The same model that generates also edits, so lighting, style, and untouched regions stay consistent rather than being rebuilt from scratch. Teams reach for it to iterate on approved visuals without a full redesign.

Multiple reference image URLs let the model lock onto a subject and carry its identity across entirely new scenes, poses, and backgrounds. This subject-driven mode keeps a character, product, or brand mascot recognizable from one generation to the next without any per-image fine-tuning. It suits campaigns, storyboards, and game assets where the same figure has to appear everywhere.

How much control do you actually need? Tune guidance_scale from 1.0 to 20.0 and inference steps from 1 to 100, choose one of six aspect presets, and export as PNG, JPEG, or WebP. Every call runs through one OpenAI-compatible endpoint at a transparent $0.044 per image with pay-as-you-go billing and no subscription. Start building today.
Send one identical prompt through the HiDream O1 1.5 Image API alongside two rival image models, then compare how each reads the same words into composition, lighting, and fine detail.
A bustling morning fish market in a Mediterranean harbor town, wooden stalls lined with hand-chalked price boards spelling out the day's fresh catch, a young fishmonger in a striped apron laughing mid-gesture as she tosses a silver sardine into the air, low golden side light raking across wet cobblestones and glistening fish scales, deep telephoto compression stacking the stalls into a soft misty harbor behind, palette of teal shutters against warm terracotta walls and cold silver fish, crisp chalk lettering and weathered wood grain, candid documentary reportage photography, 35mm, wide 16:9 aspect ratio, full-bleed

Generated with HiDream O1 1.5 on Atlas Cloud

Generated with Nano Banana Pro on Atlas Cloud

Generated with Seedream v4.5 on Atlas Cloud
A pair of scarlet macaws caught mid-squabble over a fruiting cecropia branch, wings flared into a burst of crimson and cobalt, one bird tumbling upside down mid-flap, backlit by soft overcast jungle light glowing through translucent feathers, shot on a 400mm telephoto that compresses layered misty rainforest into the background, generous negative space of pale sky filling the right third, complementary red plumage read against deep emerald foliage, feather barbs and beak texture rendered razor sharp, natural-history wildlife photography, wide 16:9 aspect ratio, full-bleed

Generated with HiDream O1 1.5 on Atlas Cloud

Generated with Nano Banana Pro on Atlas Cloud

Generated with Seedream v4.5 on Atlas Cloud
Across e-commerce, advertising, game art, and social campaigns, the HiDream O1 1.5 Image API turns one prompt or a set of references into generation, editing, and subject-consistent personalization at a flat $0.044 per image.
Retail teams generate product shots and lifestyle scenes from a text prompt at $0.044 per image, choosing from six aspect ratio presets. Catalog visuals ship without a photo shoot or studio turnaround.
Craft campaign posters and banners rendered as rigorously composed, cinematically lit layouts across landscape, portrait, and square framings. Agencies iterate on hero creative in one sitting, then hand production-ready art to clients.
One reference image plus an editing prompt lets the model restyle, retouch, or recompose a photo while preserving its structure and lighting. Designers fix backgrounds or swap elements without a full editor.
Feed several reference images and the model keeps a character, product, or mascot consistent across entirely new scenes. Studios build reusable brand assets and campaign series that stay on model.
When a game team needs environments, props, or character concepts, the model returns detailed art tuned by guidance scale and inference steps. Art directors explore visual directions before committing studio time.
Running a busy content calendar? Marketers spin up scroll-stopping graphics for posts, stories, and thumbnails across square, portrait, and landscape presets, each rendered at a flat and predictable $0.044 per image.
See where the HiDream O1 1.5 Image API stands next to Alibaba and ByteDance image models on built-in reasoning, bilingual text, open weights, and per-image cost.
| Model | Provider | Reasoning Prompt Agent | Bilingual Text Rendering | Open Weights | Price (per image) |
|---|---|---|---|---|---|
| HiDream O1 1.5 Text-to-Image | HiDream.ai | √ | √ | √ | $0.044 |
| HiDream O1 1.5 Edit | HiDream.ai | √ | √ | √ | $0.044 |
| Qwen Image 2.0 | Alibaba (Qwen) | - | √ | - | $0.035 |
| Seedream v4.5 | ByteDance | - | √ | - | $0.04 |
Get started in minutes — follow these simple steps to integrate and deploy models through Atlas Cloud's platform.
Sign up at atlascloud.ai and complete verification. New users receive free credits to explore the platform and test models.
Combining the advanced HiDream models with Atlas Cloud's GPU-accelerated platform provides unmatched performance, scalability, and developer experience.
Low Latency:
GPU-optimized inference for real-time reasoning.
Unified API:
Run HiDream, GPT, Gemini, and DeepSeek with one integration.
Transparent Pricing:
Predictable per-token billing with serverless options.
Developer Experience:
SDKs, analytics, fine-tuning tools, and templates.
Reliability:
99.99% uptime, RBAC, and compliance-ready logging.
Security & Compliance:
SOC 2 Type II, HIPAA alignment, data sovereignty in US.
The HiDream O1 1.5 Image API gives developers programmatic access to HiDream's unified image generation model through a single OpenAI-compatible endpoint on Atlas Cloud. Built on a pixel-level unified transformer, it delivers text-to-image, editing, and subject-driven personalization from one model instead of a stack of separate tools. Access is Day-0 with pay-as-you-go, transparent per-call pricing.
Beyond straightforward text-to-image generation, the model handles instruction-based editing, subject-driven personalization across multiple reference images, and accurate long-text rendering for posters and commercial graphics. Teams reach for it in e-commerce product visuals, advertising creative, and game art, where tight composition and legible on-image text both matter.
Yes. HiDream O1 1.5 was trained to interpret nuanced prompts in both Chinese and English, and it renders multilingual on-image text with strong accuracy. That makes it a practical fit for teams shipping localized visuals without switching between models.
You call the HiDream O1 1.5 Image API with one OpenAI-compatible key, so most existing SDKs work once you point them at the Atlas Cloud endpoint. Send a request with your prompt and any optional parameters to the hidream-o1-1.5/text-to-image model, then read back the generated image. No separate model hosting or GPU infrastructure is required on your side.
Prompts can run up to 2,500 characters, and you pick from preset sizes including square_hd at 1024x1024, square at 512x512, plus portrait and landscape options in 4:3 and 16:9. You can also tune num_inference_steps from 1 to 100 with a default of 50, set guidance_scale between 1.0 and 20.0 with a default of 5.0, and return PNG, JPEG, or WebP.
Pass a single URL in reference_image_urls to run instruction-based editing on an existing image, or supply multiple URLs to drive personalization that keeps a consistent subject across scenes. Leave the field empty for standard text-to-image generation. A dedicated hidream-o1-1.5/edit model is available for editing workflows at the same per-image rate.
The HiDream O1 1.5 Image API is priced at $0.044 per image on Atlas Cloud, and the text-to-image and edit models share that same rate. Billing is pay-as-you-go with transparent per-call pricing, so you pay only for the images you generate with no subscription. Start building today.
On Atlas Cloud you choose a preset size such as square_hd at 1024x1024, and the model synthesizes each image directly from raw pixels through its unified transformer rather than compressing into a latent space. Because detail and on-image text are generated instead of upscaled from a bottleneck, HiDream is known for clean typography and crisp edges in posters and product graphics.
Join the Discord community for the latest model updates, prompts, and support.