Higgsfield API Alternative
The models developers look for in the Higgsfield API — Seedance 2.5 and 2.0, Kling 3.0 and O3, Wan 3.0, open-weights MiniMax H3 and Qwen Image 3.0 — plus Seedream 5.0, Nano Banana 2 and Pro, and GPT Image 2.5 and 2, on one WaveSpeedAI API key. Pay as you go: credits never expire, and there is no membership or subscription.
Evaluating the Higgsfield API? Every endpoint below is live on WaveSpeedAI today, with a playground, examples, and its current per-run price. Top up once and use your credits whenever you like — they never expire — on these models and 1,000+ other image, video, audio, 3D, and language models.
Overview
About the Higgsfield API Alternative
What you can call on WaveSpeedAI if you are evaluating the Higgsfield API.
The models developers look for in the Higgsfield API — Seedance 2.5 and 2.0, Kling 3.0 and O3, Wan 3.0, open-weights MiniMax H3 and Qwen Image 3.0 — plus Seedream 5.0, Nano Banana 2 and Pro, and GPT Image 2.5 and 2, on one WaveSpeedAI API key. Pay as you go: credits never expire, and there is no membership or subscription.
Evaluating the Higgsfield API? Every endpoint below is live on WaveSpeedAI today, with a playground, examples, and its current per-run price. Top up once and use your credits whenever you like — they never expire — on these models and 1,000+ other image, video, audio, 3D, and language models.
The endpoints below are 37 live WaveSpeedAI REST endpoints for these models, covering Text-To-Video, Image-To-Video, Lora-Support, Image-To-Image, Text-To-Image workflows. Each has its own playground, pricing, and examples, and every one is called the same way: POST JSON, get a prediction id back, poll or receive a webhook.
They share one API key, one balance, and one integration with the other 1,000+ models on WaveSpeedAI, from text-to-image and text-to-video through audio, 3D, upscaling, editing, and language models.
Endpoints
All WaveSpeedAI API endpoints
37 endpoints available now on WaveSpeedAI — pick the variant that matches your workflow.
/filters:quality(82)/media/images/1790244142136824775_5G4eMW5f.webp)
Seedance 2.5 Text To Video Turbo
Seedance 2.5 (Text-to-Video Turbo) generates cinematic 720p/1080p videos from text prompts —a faster, more affordable high-resolution tier with native audio-visual synchronization, director-level control, and exceptional motion stability.
/filters:quality(82)/media/images/1789704587083437544_OnT2bkuD.webp)
Seedance 2.5 Text To Video
Seedance 2.5 (Text-to-Video) generates Hollywood-grade cinematic videos from text prompts with native audio-visual synchronization, director-level camera and lighting control, and exceptional motion stability. Built on Seed's unified multimodal architecture, it leads on instruction adherence, motion quality, and visual aesthetics.
/filters:quality(82)/media/images/1786121566327260938_KpNW6yGR.webp)
Seedance 2.5 Image To Video Turbo
Seedance 2.5 (Image-to-Video Turbo) generates cinematic 720p/1080p videos from reference images —a faster, more affordable high-resolution tier with native audio-visual synchronization, director-level control, and exceptional motion stability.
/filters:quality(82)/media/images/1786122164333718452_ueCMW5CM.webp)
Seedance 2.5 Image To Video
Seedance 2.5 (Image-to-Video) generates Hollywood-grade cinematic videos from reference images and text prompts with native audio-visual synchronization, director-level camera and lighting control, and exceptional motion stability. Built on Seed's unified multimodal architecture, it preserves the input image's subject and composition while adding expressive, physically accurate motion.
/filters:quality(82)/media/images/1777710782211882265_rR0PY8hr.webp)
Seedance 2.0 Text To Video Turbo
Seedance 2.0 (Text-to-Video Turbo) generates cinematic 720p/1080p videos from text prompts —a faster, more affordable high-resolution tier with native audio-visual synchronization, director-level control, and exceptional motion stability.
/filters:quality(82)/media/images/1784114740421359958_fCL2KT3c.webp)
Seedance 2.0 Text To Video
Seedance 2.0 (Text-to-Video) generates Hollywood-grade cinematic videos from text prompts with native audio-visual synchronization, director-level camera and lighting control, and exceptional motion stability. Built on Seed's unified multimodal architecture, it leads on instruction adherence, motion quality, and visual aesthetics.
/filters:quality(82)/media/images/1777710796465656330_J8EMV5eo.webp)
Seedance 2.0 Image To Video Turbo
Seedance 2.0 (Image-to-Video Turbo) generates cinematic 720p/1080p videos from reference images —a faster, more affordable high-resolution tier with native audio-visual synchronization, director-level control, and exceptional motion stability.
/filters:quality(82)/media/images/1784114374920321015_tUkW5enw.webp)
Seedance 2.0 Image To Video
Seedance 2.0 (Image-to-Video) generates Hollywood-grade cinematic videos from reference images and text prompts with native audio-visual synchronization, director-level camera and lighting control, and exceptional motion stability. Built on Seed's unified multimodal architecture, it preserves the input image's subject and composition while adding expressive, physically accurate motion.
/filters:quality(82)/media/images/20260408113003_zrzrvv9r.webp)
Kling v3.0 Pro Text To Video
Kling 3.0 Pro delivers top-tier text-to-video generation with smooth motion, cinematic visuals, accurate prompt adherence, and native audio for ready-to-share clips. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.
/filters:quality(82)/media/images/20260408112956_u1xy8vi4.webp)
Kling v3.0 Pro Image To Video
Kling 3.0 Pro delivers top-tier image-to-video generation with smooth motion, cinematic visuals, accurate prompt adherence, and native audio for ready-to-share clips. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.
/filters:quality(82)/media/images/20260408113103_lt3p5mhp.webp)
Kling Video O3 Pro Text To Video
Kling Omni Video O3 is Kuaishou's advanced unified multi-modal video model with MVL (Multi-modal Visual Language) technology. Text-to-Video mode generates cinematic videos from text prompts with subject consistency, natural physics simulation, and precise semantic understanding. Supports audio generation. Ready-to-use REST API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408113054_tn1k4n54.webp)
Kling Video O3 Pro Image To Video
Kling Omni Video O3 Image-to-Video transforms static images into dynamic cinematic videos using MVL (Multi-modal Visual Language) technology. Maintains subject consistency while adding natural motion, physics simulation, and seamless scene dynamics. Supports audio generation. Ready-to-use REST API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1785738817804641004_fy3dmvE0.webp)
Wan 3.0 Text To Video
Wan 3.0 Text to Video generates cinematic videos from text prompts, with flexible 2-30 second duration, aspect ratio control, optional audio, and deep-thinking controls for high-quality video generation. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1785738862631835931_QQZ9i1ak.webp)
Wan 3.0 Image To Video
Wan 3.0 Image to Video animates a first-frame image into a cinematic video, with optional last-frame guidance, flexible 2-30 second duration, aspect ratio control, optional audio, and deep-thinking controls for high-quality motion and scene continuity. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1788760867264586097_7BZ9isBK.webp)
Minimax H3 Text To Video Lora
MiniMax H3 Open Weights Text to Video with custom LoRA support generates coherent videos from text prompts, with 480P / 540P / 768P output, native stereo audio, 5-15 second duration, flexible aspect ratios, and per-second billing. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1785814561183432100_tW5eoxGP.webp)
Minimax H3 Text To Video
MiniMax H3 Open Weights Text to Video generates coherent videos from text prompts, with 480P / 540P / 768P output, native stereo audio, 5-15 second duration, flexible aspect ratios, and per-second billing. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1786952854061501695_t8heoxGQ.webp)
Minimax H3 Image To Video Lora
MiniMax H3 Open Weights Image to Video animates a first-frame image, optionally with last-frame guidance, into coherent 480P / 540P / 768P videos with native stereo audio, 5-15 second duration, flexible aspect ratios, and per-second billing. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1785815150161566057_2VF6wZoP.webp)
Minimax H3 Image To Video
MiniMax H3 Open Weights Image to Video animates a first-frame image, optionally with last-frame guidance, into coherent 480P / 540P / 768P videos with native stereo audio, 5-15 second duration, flexible aspect ratios, and per-second billing. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1786523429877069688_r6rFUkDV.webp)
Seedream v5.0 Pro Layer Decomposition
Seedream V5.0 Pro Layer Decomposition separates a single image into a base image and transparent layers, enabling flexible compositing, image editing, asset extraction, and layered design workflows. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1783519376754609746_GuDMV4dU.webp)
Seedream v5.0 Pro
Seedream V5.0 Pro Text to Image by ByteDance generates high-quality images from text prompts, with aspect ratio selection, strong prompt following, and 1K / 2K output tiers for flexible image creation. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1783519346875258167_uSgmwFOX.webp)
Seedream v5.0 Pro Edit
Seedream V5.0 Pro Edit by ByteDance edits and generates images from single-image or multi-reference inputs, supporting up to 10 reference images, aspect ratio selection, and 1K / 2K output tiers. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408104619_d5wh11gv.webp)
Nano Banana Pro Text To Image Multi
Google's Nano Banana Pro (Gemini 3.0 Pro Image) is a next-generation text-to-image model capable of generating multiple high-quality images in a single run. Extremely low cost — only $0.07 per image. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408104626_u8it8kvq.webp)
Nano Banana Pro Text To Image Ultra
Google's Nano Banana Pro (Gemini 3.0 Pro Image) is a cutting-edge text-to-image model enabling high-res image generation optimized for phones. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408104631_p3ms4grb.webp)
Nano Banana Pro Text To Image
Google's Nano Banana pro (Gemini 3.0 Pro Image) is a cutting-edge text-to-image model enabling high-res 4K image generation optimized for phones. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408104610_zrhfvtpp.webp)
Nano Banana Pro Edit Multi
Google's Nano Banana Pro (Gemini 3.0 Pro Image) Edit is a next-generation image editing model capable of generating multiple high-quality edited images in a single run. Extremely low cost — only $0.07 per image. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408104636_1r1p0ttd.webp)
Nano Banana Pro Edit Ultra
Google Nano Banana Pro (Gemini 3.0 Pro Image) Edit enables image editing with highres output. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408104615_axtp5nwb.webp)
Nano Banana Pro Edit
Google Nano Banana Pro (Gemini 3.0 Pro Image) Edit enables image editing with 4K-capable output. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408104558_t6mp9lus.webp)
Nano Banana 2 Text To Image Fast
Google Nano Banana 2 Fast (Gemini 3.1 Flash Image) is the cheapest Nano Banana 2 option, starting at just $0.045 per image. Delivers fast text-to-image generation with 2K default output and 4K support. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408104604_srl763ml.webp)
Nano Banana 2 Text To Image
Google Nano Banana 2 (Gemini 3.1 Flash Image) delivers Pro-quality image generation at Flash speed with 512px to 4K resolution support. Features include improved text rendering, character consistency for up to 5 characters, and real-world knowledge integration. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408104544_zo6kw479.webp)
Nano Banana 2 Edit Fast
Google Nano Banana 2 Edit Fast (Gemini 3.1 Flash Image) is the cheapest Nano Banana 2 editing option, starting at just $0.045 per image. Enables fast image editing with 2K default output and 4K support. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408104538_77qog4ig.webp)
Nano Banana 2 Edit
Google Nano Banana 2 Edit (Gemini 3.1 Flash Image) enables advanced image editing with 4K-capable output, fast iteration, and precise instruction following. Supports text translation, localization within images, and maintains subject consistency during edits. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1788941905586203566_DzViKczP.webp)
Gpt Image 2.5 Sunburst Edit
OpenAI's GPT Image 2.5 Sunburst Edit edits one or more reference images from natural-language instructions, with five quality tiers up to 4K. Sunburst is the precision-focused GPT Image 2.5 tier that spends more time per image for extra fidelity on intricate detail. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1788932339711919494_vrbhE9lL.webp)
Gpt Image 2.5 Sunburst Text To Image
OpenAI's GPT Image 2.5 Sunburst Text-to-Image generates high-quality images from natural-language prompts, with five quality tiers up to 4K. Sunburst is the precision-focused GPT Image 2.5 tier that spends more time per image for extra fidelity on intricate detail. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1782721508918099909_oCMW3dT2.webp)
Gpt Image 2 Text To Image
OpenAI's GPT Image 2 Text-to-Image generates high-quality images from natural-language prompts. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1782721434327650196_n8i8hqAK.webp)
Gpt Image 2 Edit
OpenAI's GPT Image 2 Edit enables image editing from natural-language instructions with one or more reference images. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1785949826547283917_F8QZ8hqB.webp)
Qwen Image 3.0 Text To Image
Qwen Image 3.0 Text to Image is a high-quality image generation model that creates high-quality images from text prompts, with advanced prompt understanding, strong visual quality, and up to 2K output for creative and production workflows. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1785949851816447753_YuENW5PZ.webp)
Qwen Image 3.0 Edit
Qwen Image 3.0 Edit is a high-quality image editing model that transforms existing images with natural-language instructions, delivering advanced instruction understanding, superior visual quality, and up to 2K output for creative and production workflows. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
Examples
See WaveSpeedAI in action
Real outputs generated by the WaveSpeedAI API. Hover any video to preview, click to open the full-size viewer.
How to
How to use the WaveSpeedAI API
Four steps from signup to a finished generation. Full Python, Node.js, and cURL examples are in the API section below.
- 01
Get an API key
Sign up for a WaveSpeedAI account and copy your API key from the dashboard. New accounts come with free starter credits — enough to run the playground a few dozen times before billing kicks in.
- 02
Submit a prediction
POST your input as JSON to https://api.wavespeed.ai/api/v3/bytedance/seedance-2.5/text-to-video. The endpoint returns a prediction id immediately — generations are async so you don't hold an open connection during inference.
- 03
Poll for completion
GET https://api.wavespeed.ai/api/v3/predictions/{request_id}/result. Return outputs on completed; stop with an error on failed, cancelled, timeout, or deleted; continue polling for every other status.
- 04
Read the output URL
Once status is"completed", read the URL from data.outputs[0]. The URL points to your generated media on the WaveSpeedAI CDN — image, video, audio, or 3D file depending on the model variant you called.
Use cases
What you can build with these models
Common workflows developers and creators use these APIs for.
Move a Higgsfield API integration in an afternoon
Both APIs are asynchronous: submit JSON, get a request id, then poll or take a webhook. On WaveSpeedAI you POST to /api/v3/<model> with a Bearer API key and read the prediction result, so porting is mostly a change of base URL, auth header, and model path.
Image and video in one pipeline
Generate stills with Nano Banana Pro, Seedream 5.0, or GPT Image 2.5, then animate them with Seedance 2.5, Kling 3.0, or Wan 3.0 image-to-video, on one key and one balance. Higgsfield API's published catalog does not list Nano Banana, Seedream, or GPT Image.
Open-weights MiniMax H3 with LoRA
WaveSpeedAI hosts the open-weights MiniMax H3 with text-to-video, image-to-video, reference-to-video, video edit and extend, LoRA endpoints for custom styles and characters, and a text-to-image mode.
Long single shots and tiered video
Seedance 2.5 renders single shots up to 30 seconds with native audio; Kling 3.0 and Kling O3 come in Standard, Pro, and 4K tiers; Wan 3.0 adds a Prime tier. Pick per shot from the same integration.
Agents and batch jobs
Call every model from Python or JavaScript SDKs, plain HTTP, the CLI, or an MCP server, so coding agents and batch scripts use the same endpoints and prices as your app.
Tips
Tips for prompting these models
Practical advice for getting better outputs from these models — drawn from the patterns that work across video models in production pipelines.
- 01
Be specific about camera moves
Mention concrete cinematography vocabulary — orbit, dolly-in, push-in, pan-left, crane shot, handheld follow. Generic prompts produce static or arbitrary camera choices; named camera moves map directly to motion intent in the model's training data and dramatically improve shot quality.
- 02
Anchor character identity with reference images
If your prompt depends on a specific person, character, or product, upload a reference image alongside the prompt. Without a reference, identity drifts across frames and across shots — the same character ends up looking like a slightly different person each generation.
- 03
Describe lighting and time of day
Lighting cues like 'golden hour, soft warm directional light' or 'overcast diffused light, slate-grey sky' improve quality and consistency far more than vague quality modifiers. Lighting is one of the strongest priors the model conditions on.
- 04
Use negative prompts to suppress common failure modes
Useful negatives for video: 'frame flicker, motion blur, watermark, text artifacts, distorted hands, low resolution, jpeg compression'. Negative prompts cost nothing and noticeably reduce the rate of generations you'd otherwise re-roll.
- 05
Pick the shortest duration that captures your beat
Most prompts work best at 5-8 seconds. Longer clips amplify temporal inconsistencies (subject morphing, environment drift). If you need a 20-second sequence, generate three 6-8 second clips and edit them together — quality stays higher than one long generation.
- 06
Match aspect ratio to platform up front
9:16 for TikTok / Reels / Shorts, 16:9 for landscape feeds and YouTube, 1:1 for post grids. Models train slightly differently per aspect ratio — cropping a 16:9 to 9:16 after the fact loses both fidelity and the composition the model intended.
Pricing
WaveSpeedAI API pricing
Pricing is per-output. The final charge scales with the parameters you set in each variant's playground (resolution, duration, output count, references).
API
Call the WaveSpeedAI API
Sign up for an API key at wavespeed.ai/accesskey, then submit a prediction via REST. The playground generates ready-to-paste samples for any combination of inputs.
POSThttps://api.wavespeed.ai/api/v3/bytedance/seedance-2.5/text-to-video
# 1. Submit the prediction.
SUBMIT_RESPONSE=$(curl --silent --show-error --fail-with-body \
-X POST "https://api.wavespeed.ai/api/v3/bytedance/seedance-2.5/text-to-video" \
-H "Content-Type: application/json" \
-H "Authorization: Bearer $WAVESPEED_API_KEY" \
-d '{
"prompt": "A cinematic shot of a city at sunset, soft golden light",
"aspect_ratio": "16:9",
"resolution": "720p",
"duration": 5,
"generate_audio": true
}')
TASK=$(printf '%s' "$SUBMIT_RESPONSE" | jq 'if has("data") then .data else . end')
PREDICTION_ID=$(printf '%s' "$TASK" | jq -r '.id')
if [ -z "$PREDICTION_ID" ] || [ "$PREDICTION_ID" = "null" ]; then
printf 'Submission response did not contain a prediction id
' >&2
exit 1
fi
RESULT_URL="https://api.wavespeed.ai/api/v3/predictions/$PREDICTION_ID/result"
# 2. Poll until the prediction finishes.
while true; do
RESPONSE=$(curl --silent --show-error --fail-with-body "$RESULT_URL" \
-H "Authorization: Bearer $WAVESPEED_API_KEY")
RESULT=$(printf '%s' "$RESPONSE" | jq 'if has("data") then .data else . end')
STATUS=$(printf '%s' "$RESULT" | jq -r '.status')
case "$STATUS" in
completed) printf '%s\n' "$RESULT" | jq '.outputs'; break ;;
failed|cancelled|timeout|deleted) printf '%s\n' "$RESULT" | jq . >&2; exit 1 ;;
*) sleep 2 ;;
esac
doneconst submitUrl = "https://api.wavespeed.ai/api/v3/bytedance/seedance-2.5/text-to-video";
const apiKey = process.env.WAVESPEED_API_KEY;
if (!apiKey) throw new Error('Set WAVESPEED_API_KEY');
async function requestJson(url, options = {}) {
const response = await fetch(url, options);
if (!response.ok) throw new Error(await response.text());
return response.json();
}
// 1. Submit the prediction.
const body = await requestJson(submitUrl, {
method: "POST",
headers: {
"Authorization": `Bearer ${apiKey}`,
"Content-Type": "application/json",
},
body: JSON.stringify({
"prompt": "A cinematic shot of a city at sunset, soft golden light",
"aspect_ratio": "16:9",
"resolution": "720p",
"duration": 5,
"generate_audio": true
}),
});
const task = body.data ?? body;
const resultUrl = `https://api.wavespeed.ai/api/v3/predictions/${task.id}/result`;
// 2. Poll until the prediction finishes.
while (true) {
const resultBody = await requestJson(resultUrl, {
headers: { "Authorization": `Bearer ${apiKey}` },
});
const result = resultBody.data ?? resultBody;
if (result.status === "completed") {
console.log(result.outputs);
break;
}
if (["failed", "cancelled", "timeout", "deleted"].includes(result.status)) throw new Error(JSON.stringify(result));
await new Promise(resolve => setTimeout(resolve, 2000));
}import json
import os
import time
from urllib.request import Request, urlopen
api_key = os.environ["WAVESPEED_API_KEY"]
headers = {"Authorization": f"Bearer {api_key}", "Content-Type": "application/json"}
payload = {
"prompt": "A cinematic shot of a city at sunset, soft golden light",
"aspect_ratio": "16:9",
"resolution": "720p",
"duration": 5,
"generate_audio": True
}
def request_json(url, data=None):
request = Request(url, data=data, headers=headers, method="POST" if data else "GET")
with urlopen(request) as response:
return json.load(response)
# 1. Submit the prediction.
body = request_json("https://api.wavespeed.ai/api/v3/bytedance/seedance-2.5/text-to-video", json.dumps(payload).encode())
task = body.get("data", body)
result_url = f"https://api.wavespeed.ai/api/v3/predictions/{task['id']}/result"
# 2. Poll until the prediction finishes.
while True:
result_body = request_json(result_url)
result = result_body.get("data", result_body)
status = result.get("status")
if status == "completed":
print(result.get("outputs", []))
break
if status in {"failed", "cancelled", "timeout", "deleted"}:
raise RuntimeError(result)
time.sleep(2)Compare
WaveSpeedAI vs Higgsfield API
Where WaveSpeedAI and the Higgsfield API differ.
WaveSpeedAI vs Higgsfield API: models
Higgsfield API lists 82 image and video entries, including its own Soul and Cinema Studio models plus Seedance, Kling, Wan, MiniMax, and Qwen Image. WaveSpeedAI covers those third-party families and adds Nano Banana, Seedream, GPT Image, Veo, audio, 3D, and language models: 1,000+ in total.
WaveSpeedAI vs Higgsfield API: credits
Higgsfield API funds expire one year after purchase. WaveSpeedAI credits never expire, and there is no membership or subscription: top up when you need to and pay per request, with each endpoint's price on its model page and in the pricing table on this page. The same credits also pay for the browser studios.
WaveSpeedAI vs Higgsfield API: open models
Both list MiniMax H3. On WaveSpeedAI it is the open-weights model hosted by WaveSpeedAI, with LoRA, reference-to-video, edit, extend, and image endpoints next to the standard text-to-video and image-to-video.
FAQ
Higgsfield API Alternative — Frequently asked questions
Models, pricing, migration — common questions about moving from the Higgsfield API to WaveSpeedAI.
Is WaveSpeedAI an alternative to the Higgsfield API?
Yes. WaveSpeedAI serves Seedance 2.5 and 2.0, Kling 3.0 and O3, Wan 3.0, MiniMax H3, and Qwen Image 3.0 as REST endpoints, plus Seedream 5.0, Nano Banana 2 and Pro, and GPT Image 2.5 and 2, with one API key and pay-per-request pricing.
Which models can I call that the Higgsfield API does not list?
Higgsfield API's published model catalog does not include Nano Banana, Seedream, GPT Image, Veo, or audio models. On WaveSpeedAI, Nano Banana 2 and Pro, Seedream 5.0, GPT Image 2.5 and 2, Veo 3.1, and ElevenLabs voices sit behind the same key as the video models.
How do I switch from the Higgsfield API?
Create a key at /accesskey, then POST your input as JSON to https://api.wavespeed.ai/api/v3/<model> with an Authorization: Bearer header. You get a prediction id back; poll https://api.wavespeed.ai/api/v3/predictions/<id>/result or pass a webhook. Map each model to its WaveSpeedAI path from the endpoint list on this page.
Is MiniMax H3 on WaveSpeedAI the open-weights model?
Yes. The wavespeed-ai/minimax-h3 endpoints run the open-weights MiniMax H3 on WaveSpeedAI, including LoRA variants for custom styles and characters, reference-to-video, video edit and extend, and a text-to-image mode.
Do I need a membership or subscription, and do credits expire?
No membership and no subscription. WaveSpeedAI is pay as you go: top up credits when you need them and pay per request at the price shown on each model page and in the pricing table above. Credits never expire. Eligible new accounts get free starter credits.
How do I call these APIs?
Sign up for a WaveSpeedAI account, copy your API key from /accesskey, then POST your input as JSON to https://api.wavespeed.ai/api/v3/<model>, for example https://api.wavespeed.ai/api/v3/bytedance/seedance-2.5/text-to-video. The endpoint returns a prediction id. Poll the result endpoint or pass a webhook. Python / Node.js / cURL examples are above.
How much do these APIs cost?
The endpoints on this page start at $0.024 per run; each model's price is in the pricing table above and on its model page. The exact cost scales with the parameters you set (resolution, duration, output count, references), and the playground shows it before you submit.
Start building on WaveSpeedAI
Free starter credits on signup. One API key across 1,000+ AI models from every major provider.