Introducing WaveSpeedAI Hunyuan Video Foley on WaveSpeedAI
hunyuan tencent

Introducing WaveSpeedAI Hunyuan Video Foley on WaveSpeedAI

HunyuanVideo-Foley generates realistic Foley and ambient audio from an uploaded video using a text prompt to describe desired sounds. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

5 min read
Introducing WaveSpeedAI InfiniteTalk Video-to-Video on WaveSpeedAI
wan alibaba

Introducing WaveSpeedAI InfiniteTalk Video-to-Video on WaveSpeedAI

Audio-driven InfiniteTalk turns one video plus audio into realistic talking or singing videos with lip-sync in 480p or 720p. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

5 min read
Introducing WaveSpeedAI Think Sound on WaveSpeedAI
announcement model-release

Introducing WaveSpeedAI Think Sound on WaveSpeedAI

ThinkSound turns uploaded videos into realistic, text-guided audio. Upload a video and add a text prompt to generate lifelike sound. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

5 min read
Introducing WaveSpeedAI WAN 2.1 Synthetic To Real Ditto on WaveSpeedAI
wan alibaba

Introducing WaveSpeedAI WAN 2.1 Synthetic To Real Ditto on WaveSpeedAI

WAN 2.1 Synthetic To Real Ditto mirrors motion and facial expressions in video-to-video synthetic-to-real conversion. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

5 min read
Introducing WaveSpeedAI Qwen Image Edit LoRA on WaveSpeedAI
qwen alibaba

Introducing WaveSpeedAI Qwen Image Edit LoRA on WaveSpeedAI

Qwen-Image-Edit LoRA (20B) enables bilingual Chinese/English image-to-image editing with style preservation and semantic and appearance edits. Ready-to-use REST API, best performance, no coldstarts, affordable pricing.

5 min read
Introducing WaveSpeedAI Video Outpainter on WaveSpeedAI
outpaint image-editing

Introducing WaveSpeedAI Video Outpainter on WaveSpeedAI

WaveSpeedAI Video Outpainter expands any video beyond its original boundaries while preserving motion, identity, and scene coherence. Perfect for aspect-ratio changes, reframing, adding safe margins, or generating new visual context without cropping or losing content.

5 min read
Introducing WaveSpeedAI WAN 2.2 Video Edit on WaveSpeedAI
wan alibaba

Introducing WaveSpeedAI WAN 2.2 Video Edit on WaveSpeedAI

Wan 2.2 Video Edit lets you modify videos via text prompts (e.g., change clothing or characters). Powered by Wan 2.2, it supports 480p ($0.20/5s) and 720p ($0.40/5s), up to 120s. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

6 min read
Introducing WaveSpeedAI FLUX Dev LoRA on WaveSpeedAI
flux image-generation

Introducing WaveSpeedAI FLUX Dev LoRA on WaveSpeedAI

FLUX.1 [dev] endpoint with LoRA support for fast, high-quality image generation and simple personalization via pre-trained LoRA adapters. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

5 min read
Introducing WaveSpeedAI Jib Mix Qwen Image Text-to-Image LoRA on WaveSpeedAI
qwen alibaba

Introducing WaveSpeedAI Jib Mix Qwen Image Text-to-Image LoRA on WaveSpeedAI

Jib Mix Qwen LoRA specializes in producing more natural, attractive faces and is particularly strong at rendering Asian facial features for next-gen text-to-image generation with LoRA support. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

5 min read
Introducing WaveSpeedAI Qwen Image Text-to-Image LoRA on WaveSpeedAI
flux image-generation

Introducing WaveSpeedAI Qwen Image Text-to-Image LoRA on WaveSpeedAI

Qwen-Image LoRA is a 20B MMDiT next-gen text-to-image model with LoRA support for fast customization and refined image generation. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

5 min read
Introducing WaveSpeedAI WAN 2.2 Text-to-Image LoRA on WaveSpeedAI
wan alibaba

Introducing WaveSpeedAI WAN 2.2 Text-to-Image LoRA on WaveSpeedAI

WAN 2.2 generates super-detailed images from text prompts and supports custom LoRAs for fine-grained style and subject control. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

5 min read
Introducing MiniMax Speech 02 Hd on WaveSpeedAI
announcement model-release

Introducing MiniMax Speech 02 Hd on WaveSpeedAI

Minimax Speech 02 HD is Minimax's high-definition text-to-speech model delivering clear HD voices; pricing $0.05 per 1,000 characters. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

5 min read