MiniMax Image-01 generates high-quality images from text or edits an uploaded image with a prompt. Supports flexible sizes, prompt optimization, seeded reproducibility, and batch generation.
Added Jan 21, 2026
Approx. Price
$0.005 per image
Model Type
both
Settings
Generation controls available for this model.
Images Per Run
Up to 9
Output images
Input Images
Supported
Reference/edit images accepted • Route max 30 MB
Output Sizes
Number of Images
Default
1
Prompt Optimizer
Default
No
Automatically enhance prompts for better results.
Resolution
Default
1024*1024
Options (9)
1024*1024 (Square (1024x1024)), 1280*720 (Widescreen (1280x720)), 1152*864 (Standard (1152x864)), 1248*832 (Photo (1248x832)) +5 more
Seed
Default
-1
Control reproducibility (-1 for random).
Benchmarks
Benchmarks
No benchmark data is available yet for this model.
Examples
Loading examples…
Related image models
Compare MiniMax Image-01 with similar models from the same provider or model family.
MiniMax H3 Image Edit
wavespeed-ai/minimax-h3/image-editEdit images from up to nine references, preserving identity while changing scenes, outfits, or styles. Supports 1K and 2K output.
MiniMax H3 Image
wavespeed-ai/minimax-h3/text-to-imageGenerate photorealistic and cinematic images from text at 1K or 2K resolution, with fifteen aspect ratios.
Qwen Image 2.1 Edit
wavespeed-ai/qwen-image-2.1/editEdit from up to ten reference images with natural-language instructions, strong subject preservation, flexible framing, and native output up to 2K.
Qwen Image 2.1 Edit LoRA
wavespeed-ai/qwen-image-2.1/edit-loraEdit from up to ten references while applying up to three custom LoRAs for precise identity, product, or style control at up to 2K.
Qwen Image 2.1
wavespeed-ai/qwen-image-2.1/text-to-imageCreate high-quality images with strong prompt following, multilingual text rendering, flexible framing, and native output up to 2K.
Qwen Image 2.1 LoRA
wavespeed-ai/qwen-image-2.1/text-to-image-loraGenerate Qwen Image 2.1 artwork with up to three custom character, product, or style LoRAs and native output up to 2K.