AI Image and Video API Platform
Compare models, inspect endpoints, and access docs in one place.
API Reference
Public docs
Endpoints, schemas, and examples.
OpenAPI 3.1
Typed integrations
SDK generation and machine-readable specs.
LLM Docs
Agent-ready docs
`llm.txt` and agent integration help.
API Keys
Authentication
Create keys for apps and agents.
Usage Dashboard
Requests and credits
Monitor volume, logs, and balance.
Buy Credits
Prepaid capacity
Top up for image and video usage.
On sale now
Happy Horse 40% off! · through August 25 · applied automatically to every API call and tool run
Available Models
154 AI models for image and video generation behind one async control plane
Flux 3 Video
Flux 3 Video generation. Text-to-video or image-to-video up to 20 seconds with synchronized audio, or extend and transform an existing clip.
/api/v1/models/flux-3-video/runGemini Omni Flash
Google Gemini Omni Flash: text, image, or video into 3–10s 720p clips with native audio. Image-to-video, reference images, and video editing.
/api/v1/models/google-gemini-omni-flash/runHailuo Standard
Premium quality text-to-video and image-to-video
/api/v1/models/hailuo-standard/runHappy Horse 1.0 Text-to-Video
Text-to-video with 720p/1080p output and 2-15 second durations
/api/v1/models/happyhorse-1.0-t2v/runKling 2.6 Pro
Kling Video v2.6 Pro (fal.ai). Text-to-video or image-to-video, 5 or 10 seconds, with audio generation.
/api/v1/models/kling-v2-6/runKling Video v3 Standard (Text)
Standard text-to-video with native audio
/api/v1/models/kling-video-v3-standard-text/runKling Video v3 Pro (Text)
Pro text-to-video with cinematic quality and native audio
/api/v1/models/kling-video-v3-pro-text/runLTX-2.5 Text-to-Video
Text-to-video with synced audio (6-20s, 1080p-2160p). Self-hosted open weights.
/api/v1/models/ltx-2-fast-t2v/runLTX-2.5 Text-to-Video (alias)
Alias of ltx-2-fast-t2v. LTX-2.5 has one schedule, so pro and fast are the same engine.
/api/v1/models/ltx-2-pro-t2v/runMiniMax H3
MiniMax H3 — text-to-video, image-to-video, and reference-to-video (images, video and audio references) in one model. 2K or 768p output, 5-15s, native synced audio.
/api/v1/models/minimax-h3/runMiniMax H3 Spicy
MiniMax H3 open weights, self-hosted. Text, image, or reference to video at 480p or 720p, with turbo / standard / high compute tiers and synced audio on every one.
/api/v1/models/minimax-h3-spicy/runP Video
Pruna P-Video — video generation with text/image/audio conditioning, draft mode, and 720p/1080p outputs.
/api/v1/models/p-video/runPixverse
Pixverse v5.6 video generation via Replicate — text-to-video or image-to-video with optional audio, at 360p–1080p.
/api/v1/models/pixverse/runPixVerse V6
Pixverse V6 video generation via Runware. Text-to-video, image-to-video (start frame), or multi-clip (start + end frame).
/api/v1/models/pixverse-v6/runSeedance 1.5
ByteDance Seedance 1 video generation. Text-to-video or image-to-video with optional end frame.
/api/v1/models/seedance-1.5/runSeedance 2 High
Higher-quality Seedance 2.0 video generation (supports 1080p)
/api/v1/models/seedance-2-high/runSeedance 2 Mini
Cost-effective Seedance 2.0 Mini — same creation flow at roughly half the credits (480p / 720p)
/api/v1/models/seedance-2-mini/runSeedance 2.5
Seedance 2.5 — text-to-video, first/last frame, and multimodal reference-to-video. Up to 30 seconds in one request with native audio in 11 languages.
/api/v1/models/seedance-2-5/runSeedance 2 Video Edit
Edit source videos with Seedance 2.0 using prompted changes, optional reference images, and 480p, 720p, or 1080p output.
/api/v1/models/seedance-video-edit/runVEO 3.1 Fast
Faster generation at 3 credits per second
/api/v1/models/veo-3.1-fast/runVEO 3.1 Standard
Higher quality at 8 credits per second
/api/v1/models/veo-3.1-standard/runVEO 3.1 Lite
Runware-powered Lite variant at 1.5 credits/sec for 720p and 2 credits/sec for 1080p. No reference images, no audio generation, no 1:1 aspect ratio.
/api/v1/models/veo-3.1-lite/runRunway Aleph
Runway Aleph 2.0 via Replicate. Transform up to 30 seconds of video with a prompt.
/api/v1/models/video-transform/runVidu Q3
Vidu Q3 — text-to-video and image-to-video at 360p, 540p, 720p, or 1080p with optional synchronized audio.
/api/v1/models/vidu-q3/runWAN 2.1 Video
WAN 2.1 (14B) text & image to video with LoRA support. 480p/720p, 1-5 second clips.
/api/v1/models/wan-2.1-video/runWAN 2.2 Standard
Premium quality with enhanced detail
/api/v1/models/wan-2.2-standard/runWAN 2.2 Extended
fal.ai WAN 2.2 with up to 10-second videos and dual LoRA support
/api/v1/models/wan-2.2-extended/runWAN 2.6 Standard
Higher quality, 720p/1080p support
/api/v1/models/wan-2.6-standard/runWAN 2.7 Text-to-Video
Text-to-video with audio sync, 720p/1080p output, and 2-15 second durations
/api/v1/models/wan-2.7-t2v/runWAN 3.0
Alibaba WAN 3.0 video. Text-to-video with optional reference images or first/last frame control, native audio, up to 30 seconds.
/api/v1/models/wan-3-0-video/runGrok Video
xAI Grok Imagine video. Text-to-video or image-to-video, 1-15 seconds at 480p or 720p. Image-to-video can use the Grok Imagine 1.5 backbone for natively-synchronized audio.
/api/v1/models/xai-video/runQuick start:
https://pixeldojo.ai/api/v1