AI Image and Video API Platform
Compare models, inspect endpoints, and access docs in one place.
API Reference
Public docs
Endpoints, schemas, and examples.
OpenAPI 3.1
Typed integrations
SDK generation and machine-readable specs.
LLM Docs
Agent-ready docs
`llm.txt` and agent integration help.
API Keys
Authentication
Create keys for apps and agents.
Usage Dashboard
Requests and credits
Monitor volume, logs, and balance.
Buy Credits
Prepaid capacity
Top up for image and video usage.
On sale now
Happy Horse 40% off! · through August 25 · applied automatically to every API call and tool run
Available Models
154 AI models for image and video generation behind one async control plane
Boogu Image
Boogu Image — bilingual (EN/ZH) text-to-image generation with crisp detail and 2K output.
/api/v1/models/boogu-image/runBria 3.2
Bria 3.2 — text-to-image with 9 aspect ratio presets at 1K resolution, optional image and prompt enhancement, and photography/art medium hints.
/api/v1/models/bria-3-2/runErnie
Baidu Ernie text-to-image (fal.ai). Multilingual prompts and built-in prompt expansion.
/api/v1/models/ernie/runFLUX
FLUX family on Replicate. Schnell, Dev, Pro, Kontext, Ultra, and LoRA remix variants in one entrypoint.
/api/v1/models/flux/runFlux 2 Flex
Max-quality with up to 10 reference images
/api/v1/models/flux-2-flex/runFlux 2 Klein 4B
Very fast generation and editing with up to 5 reference images
/api/v1/models/flux-2-klein-4b/runFlux 2 Klein 9B
4-step distilled FLUX.2 [klein] foundation model for flexible control
/api/v1/models/flux-2-klein-9b/runFlux 2 Pro
High-quality with up to 8 reference images
/api/v1/models/flux-2-pro/runFlux 2 Max
The highest fidelity image model from Black Forest Labs
/api/v1/models/flux-2-max/runFlux 2 Dev
Fast quality with up to 4 reference images
/api/v1/models/flux-2-dev/runFlux 2 Dev + LoRA
Dev model with custom LoRA support
/api/v1/models/flux-2-lora/runFlux Dev
High-quality development model with configurable steps and guidance. For LoRAs, use Flux Dev Multi LoRA.
/api/v1/models/flux-dev/runFlux Krea Dev
Photorealistic generation that avoids the oversaturated AI look. Supports a single LoRA via lora_weights.
/api/v1/models/flux-krea-dev/runFlux Dev Multi LoRA
Flux Dev with multiple stacked custom LoRAs for complex style combinations.
/api/v1/models/flux-dev-multi-lora/runFlux 1.1 Pro
Latest pro model with enhanced quality and strong prompt adherence.
/api/v1/models/flux-1.1-pro/runFlux 1.1 Pro Ultra
Highest quality Flux model with raw mode for natural-looking images.
/api/v1/models/flux-1.1-pro-ultra/runFlux Kontext Pro
Advanced model with state-of-the-art performance for both generation and editing.
/api/v1/models/flux-kontext-pro/runFlux Kontext Max
Premium model with maximum performance and improved typography for generation and editing.
/api/v1/models/flux-kontext-max/runGoogle Gemini Flash
Fast generation with Gemini 2.5 Flash
/api/v1/models/gemini-flash/runGoogle Nano Banana Pro
SOTA with accurate typography and reasoning
/api/v1/models/nano-banana-pro/runGoogle Nano Banana 2
Next-generation SOTA model with stronger consistency
/api/v1/models/nano-banana-2/runGoogle Nano Banana 2 Lite
Faster, lower-cost Nano Banana 2 at fixed 1K resolution
/api/v1/models/nano-banana-2-lite/runGPT Image 2
OpenAI GPT Image 2 via fal.ai — next-generation image model with 4K rendering and sharper text fidelity.
/api/v1/models/gpt-image-2/runHunyuan Image 3
Hunyuan Image 3.0 — Tencent's 80B-parameter MoE text-to-image model. High-fidelity generation with seven aspect ratio presets and a fast mode toggle.
/api/v1/models/hunyuan-image-3/runIdeogram 4 Turbo
Fastest and cheapest Ideogram 4.0. Same stunning realism, creative designs, and text rendering — tuned for speed and iteration.
/api/v1/models/ideogram-v4-turbo/runIdeogram 4 Balanced
The sweet spot. Balances speed, quality, and cost — a great default for most graphic design, marketing, and poster work.
/api/v1/models/ideogram-v4-balanced/runIdeogram 4 Quality
The highest-quality Ideogram 4.0. Slowest but best for hero images, print-ready work, and detailed text-heavy layouts.
/api/v1/models/ideogram-v4-quality/runIdeogram 4 Fast
Ideogram 4.0 on a speed-tuned backbone. Same great text rendering and design at a lower cost — ideal for quick iteration and drafts.
/api/v1/models/ideogram-v4-fast/runIdeogram 4 Instant
The fastest, most affordable Ideogram 4.0. Near-instant text-to-image for rapid drafts and high-volume work.
/api/v1/models/ideogram-v4-instant/runImagineArt
ImagineArt family: 1.5, 1.5 Pro, and the 2.0 preview.
/api/v1/models/imagineart/runKling Image
Kling Image V3 (fal.ai). High-quality text-to-image with flexible aspect ratios.
/api/v1/models/kling-image/runKrea Image
Krea's aesthetic text-to-image, three tiers in one tool. Turbo for fast, cheap (0.5 credit), spicy-capable generation with optional custom LoRAs; Medium and Large for higher fidelity with a creativity control and optional style-reference images.
/api/v1/models/krea-v2/runLuma UNI 1
Luma UNI 1 (Standard + MAX) via Runware. Text-to-image and reference-guided image editing with one prompt, two quality tiers.
/api/v1/models/luma-uni-1/runMAI Image
Microsoft MAI Image 2.5 — text-to-image with strong prompt adherence, natural lighting, and clean detail across 11 aspect ratios.
/api/v1/models/mai-image/runP-Image
Pruna P-Image. Sub-second text-to-image with optional custom dimensions.
/api/v1/models/p-image/runP-Image Ideogram
Ideogram-class text rendering on the Pruna backbone. Pick a thinking budget from 0.1 to 1 credit per image.
/api/v1/models/p-image-ideogram/runPony Realism
Pony Realism - Stylized anime generation
/api/v1/models/ponyxl-ponyrealism-v23/runPony NAI
Pony NAI - Stylized anime generation
/api/v1/models/ponyxl-tponynai3-v7/runWai ANI
Wai ANI - Stylized anime generation
/api/v1/models/ponyxl-waianinsfwponyxl-v140/runQwen 3 Pro
Qwen Image 3.0 Pro — text-to-image generation and image editing in one model. Generate from a prompt, or supply 1-3 reference images plus an instruction.
/api/v1/models/qwen-3-pro/runQWEN Image Plus
Fast generation with excellent quality
/api/v1/models/qwen-image-plus/runQWEN Image Max
Highest quality output
/api/v1/models/qwen-image-max/runQWEN Image 2.0
Fast, balanced image generation and editing
/api/v1/models/qwen-image-2.0/runQWEN Image 2.0 Pro
Enhanced text rendering, realistic textures, and semantic adherence
/api/v1/models/qwen-image-2.0-pro/runRecraft V4.1
Recraft's latest image model. Better photorealism, smoother gradients, and improved text rendering vs V4. ~1024px output.
/api/v1/models/recraft-v4.1/runRecraft V4.1 Pro
V4.1 at ~2048px resolution. Print-ready and large-scale work with the same prompt accuracy and design taste as standard.
/api/v1/models/recraft-v4.1-pro/runRecraft V4.1 SVG
Production-ready SVG vector output. Clean geometry, structured layers, editable paths — V4.1's design taste applied to vector.
/api/v1/models/recraft-v4.1-svg/runRecraft V4.1 Pro SVG
Detailed SVG vector graphics with finer paths and more geometric detail than standard SVG.
/api/v1/models/recraft-v4.1-pro-svg/runReve 2.1
Reve 2.1 — generate from text, edit a single image, or remix up to 8 reference images with frame-level prompt control.
/api/v1/models/reve/runRiverflow
Sourceful Riverflow 2.0 (Fast + Pro) via Runware. Text-to-image with optional reference image guidance — references steer style and composition, the model generates a fresh frame from your prompt. Two quality tiers.
/api/v1/models/riverflow/runSeedream 4.5
ByteDance Seedream 4.5 — new-generation image creation with superior aesthetics, text rendering, and up to 4K resolution.
/api/v1/models/seedream-4/runSeedream 5
Seedream 5 Pro — the flagship Seedream image model with sharper realism, stronger prompt adherence, and best-in-class text rendering.
/api/v1/models/seedream-5/runSeedream 5 Lite
ByteDance Seedream 5.0 Lite — fast, high-quality image generation and editing with strong aesthetics and text rendering.
/api/v1/models/seedream-5-lite/runWAN 2.6
Alibaba WAN 2.6 text-to-image with prompt enhancement and multi-image output.
/api/v1/models/wan-2.6-image/runWAN 2.7 Standard
Faster Wan 2.7 image generation and editing
/api/v1/models/wan-2.7-image/runWAN 2.7 Pro
Higher quality Wan 2.7 tier with 4K support for text-to-image
/api/v1/models/wan-2.7-image-pro/runWAN Image
Fast cinematic image generation (3-6 seconds) with up to 2MP output and optional LoRA support.
/api/v1/models/wan-image/runGrok Image
xAI Grok Imagine. Fast tier for quick iteration, Quality tier for higher fidelity at 1k or 2k.
/api/v1/models/xai-image/runZ Image Spicy
Z Image Spicy text-to-image. Square / portrait / landscape compositions, 256–1536px on each side.
/api/v1/models/z-image-spicy/runZ Image Turbo
Super-fast 6B parameter text-to-image with great text rendering and LoRA support.
/api/v1/models/z-image-turbo/runQuick start:
https://pixeldojo.ai/api/v1