Skip to main content

On sale now

Happy Horse 40% off! · through August 25 · applied automatically to every API call and tool run

Available Models

154 AI models for image and video generation behind one async control plane

86Images|64Videos
Reset
Flux 3 Video example output
video

Flux 3 Video

25 credits

Flux 3 Video generation. Text-to-video or image-to-video up to 20 seconds with synchronized audio, or extend and transform an existing clip.

/api/v1/models/flux-3-video/run
text-to-videoimage-to-videovideo-extend
View Docs
Gemini Omni Flash example output
video

Gemini Omni Flash

16 credits

Google Gemini Omni Flash: text, image, or video into 3–10s 720p clips with native audio. Image-to-video, reference images, and video editing.

/api/v1/models/google-gemini-omni-flash/run
text-to-videoimage-to-video
View Docs
Grok R2V example output
video

Grok R2V

10 credits

xAI Grok Imagine reference-to-video via Replicate. 1 to 7 reference images plus prompt for 1 to 10 second clips at 480p or 720p.

/api/v1/models/grok-r2v/run
image-to-video
View Docs
Hailuo Standard example output
video

Hailuo Standard

8 credits

Premium quality text-to-video and image-to-video

/api/v1/models/hailuo-standard/run
text-to-videoimage-to-video
View Docs
Hailuo Fast example output
video

Hailuo Fast

4 credits

Fast image-to-video generation

/api/v1/models/hailuo-fast/run
image-to-video
View Docs
Happy Horse Reference example output
video

Happy Horse Reference

4 credits/sec

Alibaba Happy Horse reference-to-video (1.0 or 1.1) — multi-reference image input that preserves subject characters, driven by a text prompt. 720p / 1080p, 3-15 second clips. Version 1.1 runs at a lower per-second credit rate.

/api/v1/models/happyhorse-1.0-r2v/run
image-to-video
View Docs
Happy Horse 1.0 Image-to-Video example output
40% off through August 25
video

Happy Horse 1.0 Image-to-Video

Original price 4 credits. Sale price 2.4 credits. 40 percent off.

Image-to-video animation with 720p/1080p output and 2-15 second durations

/api/v1/models/happyhorse-1.0-i2v/run
image-to-video
View Docs
Happy Horse Video Edit example output
video

Happy Horse Video Edit

4 credits/sec

Alibaba Happy Horse 1.0 video edit — apply style transfer or local replacement to a source video using text prompts and optional reference images. 720p / 1080p, 3-15 second output.

/api/v1/models/happyhorse-1.0-video-edit/run
image-to-video
View Docs
Heygen Avatar example output
video

Heygen Avatar

2 credits/sec

Heygen Avatar 4 via fal.ai. Animate a portrait with prompt-driven speech or an audio track, with optional background and captions. Costs 2 credits per second of output video: an estimate is held at submission and settled to the actual duration when the job completes.

/api/v1/models/heygen-avatar/run
image-to-video
View Docs
Kling Motion Control v3 Standard example output
video

Kling Motion Control v3 Standard

3 credits/sec

Kling Video v3 Standard motion control endpoint

/api/v1/models/kling-motion-control/run
image-to-video
View Docs
Kling Motion Control v3 Pro example output
video

Kling Motion Control v3 Pro

4 credits/sec

Kling Video v3 Pro motion control endpoint

/api/v1/models/kling-motion-control-pro/run
image-to-video
View Docs
Kling Reference to Video example output
video

Kling Reference to Video

15 credits

Kling O3 reference-driven video generation. Image or video references, Standard or Pro tier.

/api/v1/models/kling-reference-to-video/run
image-to-videovideo-extend
View Docs
Kling 2.6 Pro example output
video

Kling 2.6 Pro

15 credits

Kling Video v2.6 Pro (fal.ai). Text-to-video or image-to-video, 5 or 10 seconds, with audio generation.

/api/v1/models/kling-v2-6/run
text-to-videoimage-to-video
View Docs
Kling Video v3 Standard (Image) example output
video

Kling Video v3 Standard (Image)

6 credits/sec

Standard image-to-video with native audio

/api/v1/models/kling-video-v3-standard-image/run
image-to-video
View Docs
Kling Video v3 Pro (Image) example output
video

Kling Video v3 Pro (Image)

8 credits/sec

Pro image-to-video with cinematic quality and native audio

/api/v1/models/kling-video-v3-pro-image/run
image-to-video
View Docs
LTX-2.5 Image-to-Video example output
video

LTX-2.5 Image-to-Video

2 credits/sec

Image-to-video with synced audio (6-20s, 1080p-2160p). Self-hosted open weights.

/api/v1/models/ltx-2-fast-i2v/run
image-to-video
View Docs
LTX-2.5 Image-to-Video (alias) example output
video

LTX-2.5 Image-to-Video (alias)

2 credits/sec

Alias of ltx-2-fast-i2v. LTX-2.5 has one schedule, so pro and fast are the same engine.

/api/v1/models/ltx-2-pro-i2v/run
image-to-video
View Docs
MiniMax H3 example output
video

MiniMax H3

15 credits

MiniMax H3 — text-to-video, image-to-video, and reference-to-video (images, video and audio references) in one model. 2K or 768p output, 5-15s, native synced audio.

/api/v1/models/minimax-h3/run
text-to-videoimage-to-video
View Docs
video

MiniMax H3 Spicy

7.5 credits

MiniMax H3 open weights, self-hosted. Text, image, or reference to video at 480p or 720p, with turbo / standard / high compute tiers and synced audio on every one.

/api/v1/models/minimax-h3-spicy/run
text-to-videoimage-to-video
View Docs
OmniHuman example output
video

OmniHuman

45 credits

ByteDance OmniHuman 1.5 via Replicate. Audio-driven talking-head video with lip sync.

/api/v1/models/omnihuman/run
audio-to-videoimage-to-video
View Docs
P Video example output
video

P Video

0.5 credits/sec

Pruna P-Video — video generation with text/image/audio conditioning, draft mode, and 720p/1080p outputs.

/api/v1/models/p-video/run
text-to-videoimage-to-videoaudio-to-video
View Docs
P Video Avatar example output
video

P Video Avatar

1 credit/sec

Pruna P Video Avatar — animate a portrait into a talking avatar from a script or an audio file. 30 voices, 10 languages, 720p / 1080p.

/api/v1/models/p-video-avatar/run
image-to-video
View Docs
Pixverse example output
video

Pixverse

7.5 credits

Pixverse v5.6 video generation via Replicate — text-to-video or image-to-video with optional audio, at 360p–1080p.

/api/v1/models/pixverse/run
text-to-videoimage-to-video
View Docs
PixVerse V6 example output
video

PixVerse V6

10 credits

Pixverse V6 video generation via Runware. Text-to-video, image-to-video (start frame), or multi-clip (start + end frame).

/api/v1/models/pixverse-v6/run
text-to-videoimage-to-video
View Docs
Seedance 1.5 example output
video

Seedance 1.5

8 credits

ByteDance Seedance 1 video generation. Text-to-video or image-to-video with optional end frame.

/api/v1/models/seedance-1.5/run
text-to-videoimage-to-video
View Docs
Seedance 2 High example output
video

Seedance 2 High

4 credits/sec

Higher-quality Seedance 2.0 video generation (supports 1080p)

/api/v1/models/seedance-2-high/run
text-to-videoimage-to-video
View Docs
video

Seedance 2 Mini

2 credits/sec

Cost-effective Seedance 2.0 Mini — same creation flow at roughly half the credits (480p / 720p)

/api/v1/models/seedance-2-mini/run
text-to-videoimage-to-video
View Docs
video

Seedance 2.5

25 credits

Seedance 2.5 — text-to-video, first/last frame, and multimodal reference-to-video. Up to 30 seconds in one request with native audio in 11 languages.

/api/v1/models/seedance-2-5/run
text-to-videoimage-to-video
View Docs
video

Seedance 2.5 Video Edit

3 credits/sec

Edit or extend an existing video with Seedance 2.5. Replace a subject, add or remove objects, or continue a clip forward or backward, with up to 50 reference materials.

/api/v1/models/seedance-2-5-video-edit/run
image-to-video
View Docs
Seedance 2 Reference example output
video

Seedance 2 Reference

20 credits

Seedance 2.0 multimodal reference-to-video. Combine up to 9 images, 3 video clips, and 3 audio tracks to guide characters, motion, and sound.

/api/v1/models/seedance-2-reference/run
image-to-video
View Docs
Seedance 2 Video Edit example output
video

Seedance 2 Video Edit

25 credits

Edit source videos with Seedance 2.0 using prompted changes, optional reference images, and 480p, 720p, or 1080p output.

/api/v1/models/seedance-video-edit/run
text-to-videoimage-to-video
View Docs
VEO 3.1 Fast example output
video

VEO 3.1 Fast

3 credits/sec

Faster generation at 3 credits per second

/api/v1/models/veo-3.1-fast/run
text-to-videoimage-to-video
View Docs
VEO 3.1 Standard example output
video

VEO 3.1 Standard

8 credits/sec

Higher quality at 8 credits per second

/api/v1/models/veo-3.1-standard/run
text-to-videoimage-to-video
View Docs
VEO 3.1 Lite example output
video

VEO 3.1 Lite

2 credits/sec

Runware-powered Lite variant at 1.5 credits/sec for 720p and 2 credits/sec for 1080p. No reference images, no audio generation, no 1:1 aspect ratio.

/api/v1/models/veo-3.1-lite/run
text-to-videoimage-to-video
View Docs
Runway Aleph example output
video

Runway Aleph

7 credits/sec

Runway Aleph 2.0 via Replicate. Transform up to 30 seconds of video with a prompt.

/api/v1/models/video-transform/run
text-to-videoimage-to-video
View Docs
Vidu Q3 example output
video

Vidu Q3

10 credits

Vidu Q3 — text-to-video and image-to-video at 360p, 540p, 720p, or 1080p with optional synchronized audio.

/api/v1/models/vidu-q3/run
text-to-videoimage-to-video
View Docs
WAN 2.1 Video example output
video

WAN 2.1 Video

1.5 credits/sec

WAN 2.1 (14B) text & image to video with LoRA support. 480p/720p, 1-5 second clips.

/api/v1/models/wan-2.1-video/run
text-to-videoimage-to-video
View Docs
WAN 2.2 Standard example output
video

WAN 2.2 Standard

3 credits

Premium quality with enhanced detail

/api/v1/models/wan-2.2-standard/run
text-to-videoimage-to-video
View Docs
WAN 2.2 Plus example output
video

WAN 2.2 Plus

10 credits

Official Alibaba model with 1080p support

/api/v1/models/wan-2.2-plus/run
image-to-video
View Docs
WAN 2.2 Extended example output
video

WAN 2.2 Extended

1.2 credits/sec

fal.ai WAN 2.2 with up to 10-second videos and dual LoRA support

/api/v1/models/wan-2.2-extended/run
text-to-videoimage-to-video
View Docs
WAN 2.2 Animate example output
video

WAN 2.2 Animate

10 credits

WAN 2.2 video animation. Drive a character image with a motion reference video.

/api/v1/models/wan-2.2-animate/run
image-to-video
View Docs
WAN 2.2 Spicy Image-to-Video example output
video

WAN 2.2 Spicy Image-to-Video

10 credits

Image-to-video with WAN 2.2 Spicy. Animate a starting image. 480p or 720p, 5s or 8s clips.

/api/v1/models/wan-2.2-i2v-spicy/run
image-to-video
View Docs
WAN 2.2 Replace example output
video

WAN 2.2 Replace

10 credits

WAN 2.2 character replacement. Swap a character in a source video while preserving scene and motion.

/api/v1/models/wan-2.2-replace/run
image-to-video
View Docs
WAN 2.6 Standard example output
video

WAN 2.6 Standard

2.5 credits/sec

Higher quality, 720p/1080p support

/api/v1/models/wan-2.6-standard/run
text-to-videoimage-to-video
View Docs
WAN 2.6 Flash example output
video

WAN 2.6 Flash

1 credit/sec

Fast and affordable image-to-video

/api/v1/models/wan-2.6-flash/run
image-to-video
View Docs
WAN 2.7 Spicy Image-to-Video example output
30% off through September 2
video

WAN 2.7 Spicy Image-to-Video

Original price 20 credits. Sale price 14 credits. 30 percent off.

Image-to-video with WAN 2.7 Spicy. Animate a starting image with optional driving audio. 720p or 1080p, 2–15 second clips.

/api/v1/models/wan-2.7-i2v-spicy/run
image-to-video
View Docs
WAN 2.7 Image-to-Video example output
video

WAN 2.7 Image-to-Video

2.5 credits/sec

Image-to-video and video continuation with optional last-frame control and audio sync

/api/v1/models/wan-2.7-i2v/run
image-to-videovideo-extend
View Docs
WAN 3.0 example output
video

WAN 3.0

12.5 credits

Alibaba WAN 3.0 video. Text-to-video with optional reference images or first/last frame control, native audio, up to 30 seconds.

/api/v1/models/wan-3-0-video/run
text-to-videoimage-to-video
View Docs
WAN Reference to Video example output
video

WAN Reference to Video

6 credits

Alibaba WAN reference-to-video. Up to 5 image/video references with multi-shot support.

/api/v1/models/wan-reference-to-video/run
image-to-video
View Docs
WAN Video Character Swap example output
video

WAN Video Character Swap

20 credits

Alibaba WAN character swap. Combine a character image with a reference video to produce a new clip.

/api/v1/models/wan-video-character-swap/run
image-to-video
View Docs
Grok Video example output
video

Grok Video

10 credits

xAI Grok Imagine video. Text-to-video or image-to-video, 1-15 seconds at 480p or 720p. Image-to-video can use the Grok Imagine 1.5 backbone for natively-synchronized audio.

/api/v1/models/xai-video/run
text-to-videoimage-to-video
View Docs

Quick start:

1. Get API key2. Buy credits3. Read docs4. Pick a model above and ship
https://pixeldojo.ai/api/v1