Model library
All 65 models — Image, Editing, Video, Avatar, Audio, 3D, Architecture, and LoRA training — ready in the studio.
Image
FLUX.1
State-of-the-art photorealistic text-to-image. Best overall quality.
Flux Klein 9B
Distilled FLUX variant matching leading quality with only 8 NFEs.
Flux Klein 9B I2I
Transform an existing image with Flux Klein 9B.
Flux.2 Dev
High-performance developer-focused text-to-image model.
Flux.2 Dev I2I
Supply an input image plus a prompt and Flux.2 Dev applies the edits.
Flux Kontext Dev
Context-guided image-to-image transformation with FLUX.
Flux Image To Image
Transform an existing image using FLUX.
Krea 2 Turbo
Fast text-to-image for stunning, high-quality visuals in seconds.
Hidream O1
Advanced text-to-image with excellent prompt understanding and stylistic diversity.
Z Image Turbo
Distilled Z-Image matching leading quality with only 8 NFEs.
Z Image Turbo I2I
Transform an existing image with Z-Image Turbo.
Ghibli Art Style
Generate Ghibli-styled images via ControlNet.
ControlNet
Structure-guided image generation using ControlNet (canny, depth, pose, and more).
QR Code Generator
Generate artistic QR codes from a prompt via ControlNet.
Editing
Qwen Image Edit 2511
Powerful prompt-based image editing with strong identity preservation.
Lazymix+ V4 Inpaint
Replace objects in specific regions of an image with a mask.
Flux Headshot
Generate professional headshots from photos with FLUX.
SDXL Headshot
Generate professional headshots from photos with SDXL.
Ultra Resolution 2x
Upscale any image 2x with AI detail enhancement.
Video
LTX 2.3 T2V
Text-to-video with LTX 2.3 — up to 12s clips.
LTX 2.3 I2V
Animate a still image into a video clip with LTX 2.3 (up to 12s).
Wan 2.2 T2V
High quality text-to-video with Wan 2.2 (up to ~7s).
Wan 2.2 I2V
Animate a still image into a video clip with Wan 2.2 (up to ~7s).
Wan 2.2 T2V (Ultra)
Ultra text-to-video generation (up to ~7s).
Wan 2.2 I2V (Ultra)
Ultra image-to-video generation (up to ~7s).
Avatar
Talking Avatar
Animate a portrait with lip-synced speech from a text script (D-ID Talks).
Lipsync from Audio
Drive a portrait with your own audio for precise lipsync (D-ID Talks).
Audio
Text To Speech
Natural multilingual text-to-speech with emotion control.
Qwen Voice Design
Create and customize any AI voice from a text prompt.
AI Music Generator
Generate original music compositions with AI.
3D
Architecture
Interior
Generate interior designs from a prompt or image.
Interior Mixer
Combine interior objects and design elements into one image.
LoRA
Z Image Turbo LoRA Trainer
Fast-train custom Z-Image models with optimized pipelines.