Skip to main content

Price drop on WAN and Seedream! WAN 3.0 is now 2 credits a second at 720p, WAN 3.0 Fast is 3 at 720p and 6 at 1080p, and WAN 2.7 Image to Video is 3 a second at 720p. Seedream 5 Pro is now 2.5 credits per image (1.25 at 1K), and Seedream 4.5 and Seedream 5 Lite are back to 1 credit. Details: pixeldojo.ai/changelog

Veo 3.1 20% off through October 8

On sale now

Veo 3.1 20% off · through October 8 · applied automatically to every API call and tool run

Available Models

150 AI models for image, video and audio generation behind one async control plane

86Images|60Videos
Reset
FLUX 3 Video example output
video

FLUX 3 Video

25 credits

FLUX 3 Video generation. Text-to-video or image-to-video up to 20 seconds with synchronized audio, or extend and transform an existing clip.

/api/v1/models/flux-3-video/run
text-to-videoimage-to-videovideo-extend
View Docs
Gemini Omni Flash 1.1 example output
video

Gemini Omni Flash 1.1

24 credits

Google Gemini Omni Flash 1.1: text, frames, references, or a source video into 3–10s clips with native audio at 360p to 4K. First/last frame, reference images and videos, edit or extend a clip.

/api/v1/models/google-gemini-omni-flash/run
text-to-videoimage-to-video
View Docs
Grok R2V example output
video

Grok R2V

10 credits

xAI Grok Imagine reference-to-video via Replicate. 1 to 7 reference images plus prompt for 1 to 10 second clips at 480p or 720p.

/api/v1/models/grok-r2v/run
image-to-video
View Docs
Grok Video Extend example output
video

Grok Video Extend

1.5 credits/sec

xAI Grok Imagine video extension. Continue an existing MP4 with a prompt-directed extension (2 to 10 seconds).

/api/v1/models/grok-video-extend/run
video-extend
View Docs
Hailuo Standard example output
video

Hailuo Standard

8 credits

Premium quality text-to-video and image-to-video

/api/v1/models/hailuo-standard/run
text-to-videoimage-to-video
View Docs
Hailuo Fast example output
video

Hailuo Fast

4 credits

Fast image-to-video generation

/api/v1/models/hailuo-fast/run
image-to-video
View Docs
Happy Horse Reference example output
video

Happy Horse Reference

4 credits/sec

Alibaba Happy Horse reference-to-video (1.0 or 1.1): multi-reference image input that preserves subject characters, driven by a text prompt. 720p / 1080p, 3-15 second clips. Version 1.1 runs at a lower per-second credit rate.

/api/v1/models/happyhorse-1.0-r2v/run
image-to-video
View Docs
Happy Horse 1.0 Text-to-Video example output
video

Happy Horse 1.0 Text-to-Video

4 credits/sec

Text-to-video with 720p/1080p output and 3-15 second durations

/api/v1/models/happyhorse-1.0-t2v/run
text-to-video
View Docs
Happy Horse 1.0 Image-to-Video example output
video

Happy Horse 1.0 Image-to-Video

4 credits/sec

Image-to-video animation with 720p/1080p output and 3-15 second durations

/api/v1/models/happyhorse-1.0-i2v/run
image-to-video
View Docs
Happy Horse Video Edit example output
video

Happy Horse Video Edit

5 credits/sec

Alibaba Happy Horse 1.0 video edit: apply style transfer or local replacement to a source video using text prompts and optional reference images. 720p / 1080p, 3-15 second output.

/api/v1/models/happyhorse-1.0-video-edit/run
image-to-video
View Docs
video

Kling Avatar

2 credits/sec

Kling Avatar V2: turn a portrait plus an audio file into a talking avatar with audio-driven lip sync. Works on realistic humans, stylized characters, cartoons, and animals.

/api/v1/models/kling-avatar/run
image-to-video
View Docs
Kling Motion Control v3 Standard example output
video

Kling Motion Control v3 Standard

3 credits/sec

Kling Video v3 Standard motion control (720p)

/api/v1/models/kling-motion-control/run
image-to-video
View Docs
Kling Motion Control v3 Pro example output
video

Kling Motion Control v3 Pro

4 credits/sec

Kling Video v3 Pro motion control (1080p)

/api/v1/models/kling-motion-control-pro/run
image-to-video
View Docs
Kling Reference to Video example output
video

Kling Reference to Video

15 credits

Kling reference-driven video generation. Image or video references, Standard or Pro tier.

/api/v1/models/kling-reference-to-video/run
image-to-videovideo-extend
View Docs
Kling 2.6 Pro example output
video

Kling 2.6 Pro

25 credits

Kling Video v2.6 Pro. Text-to-video or image-to-video, 5 or 10 seconds, with audio generation.

/api/v1/models/kling-v2-6/run
text-to-videoimage-to-video
View Docs
Kling Video v3 Standard (Text) example output
video

Kling Video v3 Standard (Text)

4 credits/sec

Standard text-to-video with native audio

/api/v1/models/kling-video-v3-standard-text/run
text-to-video
View Docs
Kling Video v3 Standard (Image) example output
video

Kling Video v3 Standard (Image)

4 credits/sec

Standard image-to-video with native audio

/api/v1/models/kling-video-v3-standard-image/run
image-to-video
View Docs
Kling Video v3 Pro (Text) example output
video

Kling Video v3 Pro (Text)

5.5 credits/sec

Pro text-to-video with cinematic quality and native audio

/api/v1/models/kling-video-v3-pro-text/run
text-to-video
View Docs
Kling Video v3 Pro (Image) example output
video

Kling Video v3 Pro (Image)

5.5 credits/sec

Pro image-to-video with cinematic quality and native audio

/api/v1/models/kling-video-v3-pro-image/run
image-to-video
View Docs
Kling Video Edit example output
video

Kling Video Edit

28 credits

Kling video-to-video edit. Standard or Pro, with optional reference images and audio preservation.

/api/v1/models/kling-video-edit/run
video-extend
View Docs
MiniMax H3 example output
video

MiniMax H3

4 credits/sec

MiniMax H3: text-to-video, image-to-video, and reference-to-video (images, video and audio references) in one model. 2K or 768p output, 5-15s, native synced audio. Also carries H3 Max (cheaper 768p/480p, frames only), H3 Max Turbo (faster, tiered 768p/480p, frames only) and H3 Fast (480p, references).

/api/v1/models/minimax-h3/run
text-to-videoimage-to-video
View Docs
video

MiniMax H3 Max

2 credits/sec

MiniMax H3 Max: the cheaper H3 sibling. Text-to-video and image-to-video at 768p (2 credits/sec) or 480p (1.5 credits/sec), 5-15s, native synced audio. No reference inputs.

/api/v1/models/minimax-h3-max/run
text-to-videoimage-to-video
View Docs
video

MiniMax H3 Turbo

7.5 credits

MiniMax H3 open weights, self-hosted with a step-distillation Turbo LoRA for fast, low-cost generation. Text, image, or reference to video at 480p or 720p with synced audio.

/api/v1/models/minimax-h3-turbo/run
text-to-videoimage-to-video
View Docs
OmniHuman example output
video

OmniHuman

5 credits/sec

ByteDance OmniHuman 1.5 via Replicate. Audio-driven talking-head video with lip sync.

/api/v1/models/omnihuman/run
audio-to-videoimage-to-video
View Docs
P-Video example output
video

P-Video

0.75 credits/sec

Pruna P-Video 2 (and P-Video 1): video generation with text/image/audio conditioning, draft mode, and 720p/1080p outputs.

/api/v1/models/p-video/run
text-to-videoimage-to-videoaudio-to-video
View Docs
P-Video Avatar example output
video

P-Video Avatar

1 credit/sec

Pruna P-Video Avatar: animate a portrait into a talking avatar from a script or an audio file. 30 voices, 10 languages, 720p / 1080p.

/api/v1/models/p-video-avatar/run
image-to-video
View Docs
PixVerse V5.6 example output
video

PixVerse V5.6

11.25 credits

PixVerse v5.6 video generation via Replicate: text-to-video or image-to-video with optional audio, at 360p–1080p.

/api/v1/models/pixverse/run
text-to-videoimage-to-video
View Docs
PixVerse V6 example output
video

PixVerse V6

10 credits

Pixverse V6 video generation via Runware. Text-to-video, image-to-video (start frame), or multi-clip (start + end frame).

/api/v1/models/pixverse-v6/run
text-to-videoimage-to-video
View Docs
Seedance 1.5 example output
video

Seedance 1.5

8 credits

ByteDance Seedance 1 video generation. Text-to-video or image-to-video with optional end frame.

/api/v1/models/seedance-1.5/run
text-to-videoimage-to-video
View Docs
Seedance 2 High example output
video

Seedance 2 High

6 credits/sec

Higher-quality Seedance 2.0 video generation (supports 1080p)

/api/v1/models/seedance-2-high/run
text-to-videoimage-to-video
View Docs
video

Seedance 2 Mini

2 credits/sec

Cost-effective Seedance 2.0 Mini: same creation flow at roughly half the credits (480p / 720p)

/api/v1/models/seedance-2-mini/run
text-to-videoimage-to-video
View Docs
video

Seedance 2.5

6 credits/sec

Seedance 2.5: text-to-video, first/last frame, and multimodal reference-to-video. Up to 30 seconds in one request with native audio in 11 languages.

/api/v1/models/seedance-2-5/run
text-to-videoimage-to-video
View Docs
video

Seedance 2.5 Video Edit

4.5 credits/sec

Edit or extend an existing video with Seedance 2.5. Replace a subject, add or remove objects, or continue a clip forward or backward, with up to 50 reference materials.

/api/v1/models/seedance-2-5-video-edit/run
image-to-video
View Docs
Seedance 2 Reference example output
20% off through October 8
video

Seedance 2 Reference

Original price 20 credits. Sale price 16 credits. 20 percent off.

Seedance 2.0 multimodal reference-to-video. Combine up to 9 images, 3 video clips, and 3 audio tracks to guide characters, motion, and sound.

/api/v1/models/seedance-2-reference/run
image-to-video
View Docs
Seedance 2 Video Edit example output
video

Seedance 2 Video Edit

25 credits

Edit source videos with Seedance 2.0 using prompted changes, optional reference images, and 480p, 720p, or 1080p output.

/api/v1/models/seedance-video-edit/run
text-to-videoimage-to-video
View Docs
Veo 3.1 Fast example output
20% off through October 8
video

Veo 3.1 Fast

Original price 4.5 credits. Sale price 3.6 credits. 20 percent off.

Faster generation at 4.5 credits per second with native audio (3 with generate_audio false)

/api/v1/models/veo-3.1-fast/run
text-to-videoimage-to-video
View Docs
Veo 3.1 Standard example output
20% off through October 8
video

Veo 3.1 Standard

Original price 12 credits. Sale price 9.6 credits. 20 percent off.

Higher quality at 12 credits per second with native audio (8 with generate_audio false)

/api/v1/models/veo-3.1-standard/run
text-to-videoimage-to-video
View Docs
Veo 3.1 Lite example output
20% off through October 8
video

Veo 3.1 Lite

Original price 2 credits. Sale price 1.6 credits. 20 percent off.

Runware-powered Lite variant at 1.5 credits/sec for 720p and 2 credits/sec for 1080p. No reference images, no audio generation, no 1:1 aspect ratio.

/api/v1/models/veo-3.1-lite/run
text-to-videoimage-to-video
View Docs
Video Autocaption example output
video

Video Autocaption

5 credits

TikTok-style auto-captioning via Replicate.

/api/v1/models/video-autocaption/run
View Docs
Video Reframe example output
video

Video Reframe

8 credits

Luma Reframe Video via Replicate. Change a video's aspect ratio intelligently.

/api/v1/models/video-reframe/run
View Docs
Runway Aleph example output
video

Runway Aleph

10 credits/sec

Runway Aleph 2.0 via Replicate. Transform up to 30 seconds of video with a prompt.

/api/v1/models/video-transform/run
text-to-videoimage-to-video
View Docs
Video Upscaler example output
video

Video Upscaler

5 credits

FlashVSR video upscaling. Enhance video resolution up to 4K.

/api/v1/models/video-upscaler/run
View Docs
Vidu Q3 example output
video

Vidu Q3

10 credits

Vidu Q3: text-to-video and image-to-video at 360p, 540p, 720p, or 1080p with optional synchronized audio.

/api/v1/models/vidu-q3/run
text-to-videoimage-to-video
View Docs
WAN 2.1 Video example output
video

WAN 2.1 Video

2.5 credits/sec

WAN 2.1 (14B) text & image to video with LoRA support. 480p/720p, 1-5 second clips.

/api/v1/models/wan-2.1-video/run
text-to-videoimage-to-video
View Docs
WAN 2.2 Standard example output
video

WAN 2.2 Standard

3 credits

Premium quality with enhanced detail

/api/v1/models/wan-2.2-standard/run
text-to-videoimage-to-video
View Docs
WAN 2.2 Plus example output
video

WAN 2.2 Plus

3 credits

Official Alibaba model with 1080p support

/api/v1/models/wan-2.2-plus/run
image-to-video
View Docs
WAN 2.2 Animate example output
video

WAN 2.2 Animate

10 credits

WAN 2.2 video animation. Drive a character image with a motion reference video.

/api/v1/models/wan-2.2-animate/run
image-to-video
View Docs
WAN 2.2 Image-to-Video example output
video

WAN 2.2 Image-to-Video

5 credits

Image-to-video with WAN 2.2. Animate a starting image. 480p or 720p, 5s or 8s clips.

/api/v1/models/wan-2.2-i2v-spicy/run
image-to-video
View Docs
WAN 2.2 Replace example output
video

WAN 2.2 Replace

10 credits

WAN 2.2 character replacement. Swap a character in a source video while preserving scene and motion.

/api/v1/models/wan-2.2-replace/run
image-to-video
View Docs
WAN 2.6 Standard example output
video

WAN 2.6 Standard

2 credits/sec

Higher quality, 720p/1080p support

/api/v1/models/wan-2.6-standard/run
text-to-videoimage-to-video
View Docs
WAN 2.6 Flash example output
video

WAN 2.6 Flash

1.5 credits/sec

Fast and affordable image-to-video

/api/v1/models/wan-2.6-flash/run
image-to-video
View Docs
WAN 2.7 Image-to-Video example output
video

WAN 2.7 Image-to-Video

12.5 credits

Image-to-video with WAN 2.7. Animate a starting image with optional driving audio. 720p or 1080p, 2–15 second clips.

/api/v1/models/wan-2.7-i2v-spicy/run
image-to-video
View Docs
WAN 2.7 Text-to-Video example output
video

WAN 2.7 Text-to-Video

2 credits/sec

Text-to-video with audio sync, 720p/1080p output, and 2-15 second durations

/api/v1/models/wan-2.7-t2v/run
text-to-video
View Docs
WAN 2.7 Image-to-Video example output
video

WAN 2.7 Image-to-Video

2 credits/sec

Image-to-video and video continuation with optional last-frame control and audio sync

/api/v1/models/wan-2.7-i2v/run
image-to-videovideo-extend
View Docs
WAN 3.0 example output
video

WAN 3.0

10 credits

Alibaba WAN 3.0 video. Text-to-video with optional reference images or first/last frame control, native audio, up to 30 seconds.

/api/v1/models/wan-3-0-video/run
text-to-videoimage-to-video
View Docs
WAN Reference to Video example output
video

WAN Reference to Video

4.5 credits/sec

Alibaba WAN reference-to-video. Up to 5 image/video references with multi-shot support.

/api/v1/models/wan-reference-to-video/run
image-to-video
View Docs
WAN Video Character Swap example output
video

WAN Video Character Swap

4 credits/sec

Alibaba WAN character swap. Combine a character image with a reference video to produce a new clip.

/api/v1/models/wan-video-character-swap/run
image-to-video
View Docs
WAN 2.7 Video Edit example output
video

WAN 2.7 Video Edit

6.5 credits/sec

Alibaba WAN 2.7 video editing. Modify an existing clip via prompt with optional reference images.

/api/v1/models/wan-video-edit/run
video-extend
View Docs
Grok Video example output
video

Grok Video

7.5 credits

xAI Grok Imagine video. Text-to-video or image-to-video, 1-15 seconds at 480p or 720p. Image-to-video can use the Grok Imagine 1.5 backbone for natively-synchronized audio.

/api/v1/models/xai-video/run
text-to-videoimage-to-video
View Docs
Grok Video Edit example output
video

Grok Video Edit

15 credits

xAI Grok Imagine Video edit. Transform short clips via Replicate.

/api/v1/models/xai-video-edit/run
video-extend
View Docs

Questions & Answers

Frequently Asked Questions

Authentication, pricing, async jobs, and output lifetime for the PixelDojo API

Quick start:

https://pixeldojo.ai/api/v1