Image o1
Model Overview
A multimodal image generation model supporting up to 10 reference images for visual consistency, detailed element editing, style control, and series generation. Ideal for character IPs, comic artwork, and branded content.
Setup your API Key
If you don’t have an API key for the AI/ML API yet, feel free to use our Quickstart guide.
API Schema
The text prompt describing the content, style, or composition of the image to be generated.
List of URLs or local Base64 encoded images to edit.
The aspect ratio of the generated image.
16:9Possible values: The resolution of the output image.
1KPossible values: The format of the generated image.
pngPossible values: The number of images to generate.
1Provider routing override. kling (alias klingai) runs native Kling with no fallback; fal runs the fal.ai mirror; auto (default) uses the Kling -> fal.ai fallback chain. Case-insensitive.
autocurl -L \
--request POST \
--url 'https://api.aimlapi.com/v1/images/generations' \
--header 'Authorization: Bearer <YOUR_AIMLAPI_KEY>' \
--header 'Content-Type: application/json' \
--data '{
"model": "klingai/image-o1",
"prompt": "Combine the images so the T-Rex is wearing a business suit, sitting in a cozy small café, drinking from the mug. Blur the background slightly to create a bokeh effect.",
"image_urls": [
"https://raw.githubusercontent.com/aimlapi/api-docs/main/reference-files/t-rex.png",
"https://raw.githubusercontent.com/aimlapi/api-docs/main/reference-files/blue-mug.jpg"
]
}'{
"data": [
{
"url": "https://cdn.aimlapi.com/generations/hedgehog/1749730923700-29fe35d2-4aef-4bc5-a911-6c39884d16a8.png",
"b64_json": null
}
],
"meta": {
"usage": {
"credits_used": 120000
}
}
}Quick Example
Let's generate an image using two input images and a prompt that defines how they should be edited.



Last updated
Was this helpful?