PixelDojo API
AI Image and Video API Platform
Build with 150+ AI image, video and audio models through one REST API. Submit an async job, poll or use a webhook, and get output URLs back. It uses the same credits as the web app, with plans from $10/month.
API Reference
Public docs
Endpoints, schemas, and examples.
OpenAPI 3.1
Typed integrations
SDK generation and machine-readable specs.
LLM Docs
Agent-ready docs
llm.txt and agent integration help.
API Keys
Authentication
Create keys for apps and agents.
Usage Dashboard
Requests and credits
Monitor volume, logs, and balance.
Buy Credits
Prepaid capacity
Top up for image and video usage.
On sale now
Veo 3.1 20% off · through October 8 · applied automatically to every API call and tool run
Available Models
150 AI models for image, video and audio generation behind one async control plane
Change Camera Angle
Camera-aware editing via Qwen Image Edit with multi-angle LoRA. 360° orbit, tilt, and zoom.
/api/v1/models/change-camera-angle/runClarity Pro
Clarity Pro Upscaler via Replicate. Photorealistic upscaling with identity preservation and creative control: up to 16× and 64 megapixels.
/api/v1/models/clarity-pro-upscaler/runConsistent Characters
Generate consistent character variations with FLUX Kontext, Nano Banana Pro/2, Flux 2 Dev, Qwen Image 2 Pro, or Grok Imagine.
/api/v1/models/consistent-characters/runCreative Upscaler
Clarity Upscaler (creative upscale) via Replicate. Boost detail with stable-diffusion refinement.
/api/v1/models/creative-upscale/runPortrait Upscaler
Crystal Upscaler via Replicate. Face-detail preserving upscale, cost scales with output megapixels.
/api/v1/models/face-enhance/runFLUX
FLUX family on Replicate. Schnell, Dev, Pro, Kontext, Ultra, and LoRA remix variants in one entrypoint.
/api/v1/models/flux/runFlux 2 Flex
Max-quality with up to 10 reference images
/api/v1/models/flux-2-flex/runFlux 2 Klein 4B
Very fast generation and editing with up to 5 reference images
/api/v1/models/flux-2-klein-4b/runFlux 2 Klein 9B
4-step distilled FLUX.2 [klein] foundation model for flexible control
/api/v1/models/flux-2-klein-9b/runFlux 2 Pro
High-quality with up to 8 reference images
/api/v1/models/flux-2-pro/runFlux 2 Max
The highest fidelity image model from Black Forest Labs
/api/v1/models/flux-2-max/runFlux 2 Dev
Fast quality with up to 4 reference images
/api/v1/models/flux-2-dev/runFlux 2 Klein 9B + LoRA
Klein 9B with custom LoRA support and up to 5 reference images
/api/v1/models/flux-2-lora/runFlux 2 Dev + LoRA
Flux 2 Dev with your own FLUX.2 [dev] LoRAs and up to 4 reference images
/api/v1/models/flux-2-dev-lora/runFLUX 3 Image
Black Forest Labs FLUX 3: photoreal text-to-image up to 4K, edits with up to 10 reference images, and web-grounded generation for real products and places.
/api/v1/models/flux-3-image/runFlux Edit (Kontext)
Black Forest Labs FLUX.1 Kontext for text-driven image editing. Dev (open-weight), Pro (state-of-the-art), and Max (premium typography).
/api/v1/models/flux-edit/runFlux Kontext Pro
Advanced model with state-of-the-art performance for both generation and editing.
/api/v1/models/flux-kontext-pro/runFlux Kontext Max
Premium model with maximum performance and improved typography for generation and editing.
/api/v1/models/flux-kontext-max/runNano Banana Edit
Google Nano Banana image editing. Multi-image fusion + edit instruction with Standard/Pro/Pro-fal tiers; 1K/2K/4K on the Pro tiers, fixed 1K on Standard.
/api/v1/models/google-nano-banana/runGPT Image 2
OpenAI GPT Image 2: next-generation image model with 4K rendering and sharper text fidelity.
/api/v1/models/gpt-image-2/runGPT Image 2.5
OpenAI GPT Image 2.5 in two variants: Sunburst for premium reference-faithful renders and edits, Flare for the same quality tier at half the latency. Up to 4K output, optional reference images and masks.
/api/v1/models/gpt-image-2-5/runGPT Image 2 Edit
OpenAI GPT Image 2 image editing: supply 1-8 reference images plus an edit instruction. 4K-capable; pricing varies by quality + size.
/api/v1/models/gpt-image-2-edit/runGrok Video Extend
xAI Grok Imagine video extension. Continue an existing MP4 with a prompt-directed extension (2 to 10 seconds).
/api/v1/models/grok-video-extend/runHappy Horse Video Edit
Alibaba Happy Horse 1.0 video edit: apply style transfer or local replacement to a source video using text prompts and optional reference images. 720p / 1080p, 3-15 second output.
/api/v1/models/happyhorse-1.0-video-edit/runIdeogram 4.5
Ideogram 4.5: best-in-class text rendering, posters, logos and photoreal images, plus natural-language edits and Precise Edit that changes only what you ask.
/api/v1/models/ideogram-4-5/runIdeogram Character
Generate consistent characters from a single reference image in many styles.
/api/v1/models/ideogram-character/runCharacter Stylist
One-shot FLUX Kontext variants: filters, cartoonify, iconic locations, haircut swap, headshots, face-to-many, and more.
/api/v1/models/image-editor/runMagic Lighting
Relight images with Magic Lighting, Nano Banana Pro/2, or Qwen Image Edit: multi-provider routing with per-model credit rates.
/api/v1/models/image-relighting/runImage to Image
FLUX Dev LoRA image-to-image on Replicate. Prompt + source image + optional LoRA weights.
/api/v1/models/image-to-image-flux/runKling Image Edit
Kling Image V3 image-to-image editing with a text instruction.
/api/v1/models/kling-image-edit/runKling Video Edit
Kling video-to-video edit. Standard or Pro, with optional reference images and audio preservation.
/api/v1/models/kling-video-edit/runLuma UNI 1
Luma UNI 1 (Standard + MAX) via Runware. Text-to-image and reference-guided image editing with one prompt, two quality tiers.
/api/v1/models/luma-uni-1/runMagnific Upscaler
Freepik Magnific upscaler. Creative or precision mode, up to 16x.
/api/v1/models/magnific-upscaler/runImage Outpainting
Outpainting. Expand an image beyond its original edges.
/api/v1/models/outpaint/runP-Image Edit
Pruna P-Image Edit. Fast image editing with up to 5 reference images.
/api/v1/models/p-image-edit/runP-Image Upscale
Pruna P-Image Upscale. Fast image upscaling to a target megapixel size, with optional detail and realism enhancement.
/api/v1/models/p-image-upscale/runPony Realism
Pony Realism - Stylized anime generation
/api/v1/models/ponyxl-ponyrealism-v23/runPony NAI
Pony NAI - Stylized anime generation
/api/v1/models/ponyxl-tponynai3-v7/runWai ANI
Wai ANI - Stylized anime generation
/api/v1/models/ponyxl-waianinsfwponyxl-v140/runQwen 3 Pro
Qwen Image 3.0 Pro: text-to-image generation and image editing in one model. Generate from a prompt, or supply 1-3 reference images plus an instruction.
/api/v1/models/qwen-3-pro/runQwen Image Plus
Fast generation with excellent quality
/api/v1/models/qwen-image-plus/runQwen Image Max
Highest quality output
/api/v1/models/qwen-image-max/runQwen Image 2.0
Fast, balanced image generation and editing
/api/v1/models/qwen-image-2.0/runQwen Image 2.0 Pro
Enhanced text rendering, realistic textures, and semantic adherence
/api/v1/models/qwen-image-2.0-pro/runQwen Image 2 Edit
Alibaba DashScope Qwen Image 2 edit: supply 1-3 reference images plus an edit instruction. Standard and Pro variants.
/api/v1/models/qwen-image-2-edit/runQwen Image Edit
Alibaba DashScope Qwen Image edit: supply 1-3 reference images plus an edit instruction. Plus and Max model variants.
/api/v1/models/qwen-image-edit/runQwen Image Edit Realistic
Qwen Image Edit, Realistic mode. Add, remove, or modify elements in an existing image with text guidance.
/api/v1/models/qwen-image-edit-spicy/runFlux Redux
Black Forest Labs Flux Redux image variations: feed a source image, get stylistic riffs.
/api/v1/models/redux-flux/runSeedance 2.5 Video Edit
Edit or extend an existing video with Seedance 2.5. Replace a subject, add or remove objects, or continue a clip forward or backward, with up to 50 reference materials.
/api/v1/models/seedance-2-5-video-edit/runSeedance 2 Video Edit
Edit source videos with Seedance 2.0 using prompted changes, optional reference images, and 480p, 720p, or 1080p output.
/api/v1/models/seedance-video-edit/runSeedream 4.5
ByteDance Seedream 4.5: new-generation image creation with superior aesthetics, text rendering, and up to 4K resolution.
/api/v1/models/seedream-4/runSeedream 5
Seedream 5 Pro: the flagship Seedream image model with sharper realism, stronger prompt adherence, and best-in-class text rendering.
/api/v1/models/seedream-5/runSeedream 5 Flash
Seedream 5 Flash: the fast, low-cost Seedream 5 model for quick drafts and high-volume image generation and editing.
/api/v1/models/seedream-5-flash/runSeedream 5 Lite
ByteDance Seedream 5.0 Lite: fast, high-quality image generation and editing with strong aesthetics and text rendering.
/api/v1/models/seedream-5-lite/runWAN 2.2 Replace
WAN 2.2 character replacement. Swap a character in a source video while preserving scene and motion.
/api/v1/models/wan-2.2-replace/runWAN 2.6 Image Edit
Alibaba WAN 2.6 image editing. Up to 4 reference images.
/api/v1/models/wan-2.6-image-edit/runWAN 2.7 Standard
Faster Wan 2.7 image generation and editing
/api/v1/models/wan-2.7-image/runWAN 2.7 Pro
Higher quality Wan 2.7 tier with 4K support for text-to-image
/api/v1/models/wan-2.7-image-pro/runWAN 2.7 Image Edit
Alibaba WAN 2.7 image editing. Standard and Pro tiers, supports up to 9 input images for fusion edits.
/api/v1/models/wan-2.7-image-edit/runWAN 2.7 Video Edit
Alibaba WAN 2.7 video editing. Modify an existing clip via prompt with optional reference images.
/api/v1/models/wan-video-edit/runGrok Image Edit
xAI Grok image editing. Sync response (no polling). Provide an image URL and a text edit instruction. Optional quality tier for 1k/2k high-fidelity edits.
/api/v1/models/xai-image-edit/runGrok Video Edit
xAI Grok Imagine Video edit. Transform short clips via Replicate.
/api/v1/models/xai-video-edit/runQuestions & Answers
Frequently Asked Questions
Authentication, pricing, async jobs, and output lifetime for the PixelDojo API
Create an API key on the API Keys page (any signed-in PixelDojo account can) and send it as a Bearer token: Authorization: Bearer pd_your_api_key. The same key works on every model endpoint.
POST to /api/v1/models/{apiId}/run with the model input. You get back a jobId and a statusUrl. Poll GET /api/v1/jobs/{jobId} until the status is completed, or pass webhook_url on the run call to be notified when the job finishes. Every model publishes its JSON schema at /api/v1/models/{apiId}/schema.
Each model lists its credit cost per generation or per second on this page and in GET /api/v1/models. These are the same credits you use in the PixelDojo web app. Plans start at $10/month for 160 credits, and credit packs start at $5 for 80 credits with no subscription needed. Failed jobs are refunded automatically.
150+ image, video, and audio models, including Nano Banana, Flux 2, GPT Image 2, Seedream 5, Kling, Veo 3.1, WAN 2.7, and Seed Audio 1.0. Browse them above, or list them with GET /api/v1/models.
API outputs are available for one hour after a job completes, and every asset in the job response carries an expiresAt time. Download or copy anything you want to keep.
Yes. The OpenAPI 3.1 spec is at /api/openapi, an LLM-optimized reference is at /llm.txt, and the TypeScript SDK is @pixeldojo/sdk on npm.
Yes. Connect the hosted MCP server at https://pixeldojo.ai/mcp from Claude, Cursor, Codex, ChatGPT and other MCP clients, or run npx @pixeldojo/mcp init for a local install. See the Skills and MCP page for every tool.
Quick start:
https://pixeldojo.ai/api/v1