Catalogue

Every model, live pricing

What you see is what runs. Inputs, parameters, and credit cost are pulled directly from the production manifest — no roadmap models, no waitlist tease.

Text → Image

3 models

Alibaba/wan 2.5/text To Image

wavespeed · live

3 credits

Alibaba WAN 2.5 Text-to-Image turns text prompts into AI-generated images with the WAN 2.5 model for on-demand image creation. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

Inputs

  • prompt · text

Params

  • enable_prompt_expansion · bool
  • negative_prompt · textarea
  • seed · int
  • size · text

Output: image

Open in Studio →

Google/nano Banana Pro/text To Image

wavespeed · live

15 credits

Google's Nano Banana pro (Gemini 3.0 Pro Image) is a cutting-edge text-to-image model enabling high-res 4K image generation optimized for phones. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

Inputs

  • prompt · text

Params

  • aspect_ratio · 1:1 / 3:2 / 2:3 / 3:4 / 4:3 / 4:5 / 5:4 / 9:16 / 16:9 / 21:9
  • enable_base64_output · bool
  • enable_sync_mode · bool
  • output_format · png / jpeg
  • resolution · 1k / 2k / 4k

Output: image

Open in Studio →

Ideogram Ai/ideogram V3 Balanced

wavespeed · live

6 credits

Ideogram V3 Balanced delivers the highest-quality image generation with stunning realism, creative design, and consistent styles. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

Inputs

  • reference_images · audio (optional)
  • image · image (optional)
  • mask_image · image (optional)
  • prompt · text

Params

  • aspect_ratio · 1:1 / 16:9 / 9:16 / 4:3 / 3:4
  • enable_base64_output · bool
  • style · Auto / General / Realistic / Design

Output: image

Open in Studio →

Image → Image

2 models

Google/nano Banana Pro/edit

wavespeed · live

15 credits

Google Nano Banana Pro (Gemini 3.0 Pro Image) Edit enables image editing with 4K-capable output. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

Inputs

  • images · audio
  • prompt · text

Params

  • aspect_ratio · 1:1 / 3:2 / 2:3 / 3:4 / 4:3 / 4:5 / 5:4 / 9:16 / 16:9 / 21:9
  • enable_base64_output · bool
  • enable_sync_mode · bool
  • output_format · png / jpeg
  • resolution · 1k / 2k / 4k

Output: image

Open in Studio →

Kwaivgi/kling Image O1

wavespeed · live

3 credits

Kling Omni Image O1 is Kuaishou's multi-modal image generation model with MVL technology. Supports up to 10 reference images for feature consistency, precise detail editing (add/remove/modify), style control, and series content creation. Perfect for IP character design, comic panels, and brand merchandise. Ready-to-use REST API, best performance, no coldstarts, affordable pricing.

Inputs

  • images · audio (optional)
  • prompt · text

Params

  • aspect_ratio · 16:9 / 9:16 / 1:1 / 4:3 / 3:4 / 3:2 / 2:3 / 21:9 / auto
  • num_images · int · 1–9
  • resolution · 1k / 2k

Output: image

Open in Studio →

Text → Video

2 models

Bytedance/seedance 2.0/text To Video

wavespeed · live

60 credits

Seedance 2.0 (Text-to-Video) generates Hollywood-grade cinematic videos from text prompts with native audio-visual synchronization, director-level camera and lighting control, and exceptional motion stability. Built on ByteDance Seed's unified multimodal architecture, it leads on instruction adherence, motion quality, and visual aesthetics. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.

Inputs

  • reference_audios · audio (optional)
  • reference_images · audio (optional)
  • reference_videos · audio (optional)
  • prompt · text

Params

  • aspect_ratio · 16:9 / 9:16 / 4:3 / 3:4 / 1:1 / 21:9
  • duration · int · 4–15
  • enable_web_search · bool
  • generate_audio · bool
  • resolution · 480p / 720p / 1080p / 4k

Output: video

Open in Studio →

Bytedance/seedance V1 Lite T2v 720p

wavespeed · live

16 credits

ByteDance Seedance V1 Lite produces coherent multi-shot 720p videos with smooth motion and accurate following of detailed text prompts. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

Inputs

  • prompt · text

Params

  • aspect_ratio · 21:9 / 16:9 / 4:3 / 1:1 / 3:4 / 9:16
  • camera_fixed · bool
  • duration · 2 / 3 / 4 / 5 / 6 / 7 / 8 / 9 / 10 / 11 / 12
  • seed · int

Output: video

Open in Studio →

Image → Video

4 models

Alibaba/wan 2.5/image To Video

wavespeed · live

25 credits

Alibaba WAN 2.5 converts text or images into videos (480p/720p/1080p) with synced audio, faster and more affordable than Google Veo3. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

Inputs

  • audio · audio (optional)
  • image · image
  • prompt · text

Params

  • duration · 3 / 4 / 5 / 6 / 7 / 8 / 9 / 10
  • enable_prompt_expansion · bool
  • negative_prompt · textarea
  • resolution · 480p / 720p / 1080p
  • seed · int

Output: video

Open in Studio →

Alibaba/wan 2.6/image To Video

wavespeed · live

50 credits

Alibaba WAN 2.6 converts text or images into videos (720p/1080p) with synced audio, faster and more affordable than Google Veo3. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.

Inputs

  • audio · audio (optional)
  • image · image
  • prompt · text

Params

  • duration · 5 / 10 / 15
  • enable_prompt_expansion · bool
  • negative_prompt · textarea
  • resolution · 720p / 1080p
  • seed · int
  • shot_type · single / multi

Output: video

Open in Studio →

Bytedance/seedance V1 Lite I2v 720p

wavespeed · live

16 credits

ByteDance Seedance Lite i2v 720p creates coherent multi-shot image-to-video clips with smooth, stable motion and prompt fidelity. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

Inputs

  • last_image · image (optional)
  • image · image
  • prompt · text (optional)

Params

  • aspect_ratio · 21:9 / 16:9 / 4:3 / 1:1 / 3:4 / 9:16
  • camera_fixed · bool
  • duration · 2 / 3 / 4 / 5 / 6 / 7 / 8 / 9 / 10 / 11 / 12
  • seed · int

Output: video

Open in Studio →

Google/veo3.1 Fast/image To Video

wavespeed · live

120 credits

Google Veo 3.1 Fast is an Image-to-Video model with native 1080p output for high-detail videos from images and fast performance. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

Inputs

  • last_image · image (optional)
  • image · image
  • prompt · text

Params

  • aspect_ratio · 16:9 / 9:16
  • duration · 8 / 4 / 6
  • generate_audio · bool
  • negative_prompt · textarea
  • resolution · 720p / 1080p / 4k
  • seed · int

Output: video

Open in Studio →