#video-to-video
17 articles
/filters:quality(82)/media/images/1773962750383987480_n3hqzHRZ.webp)
Kling O3 4K Adds Video Edit and Video Reference
Kling O3 4K on WaveSpeedAI now includes video-edit and video-reference endpoints. Every Kling O3 tier, parameter, and price, plus when to pick std, pro, or 4K.
Introducing VOID Video Inpainting on WaveSpeedAI
VOID Video Inpainting — remove objects from video with mask-guided AI inpainting. Quad-mask or auto SAM-3 masks, optional Pass 2 refinement for temporal consistency. Now live on WaveSpeedAI.
Introducing AI Parkour Video on WaveSpeedAI
AI Parkour Video turns any portrait into an action-packed parkour clip. 6 style presets, image-to-video and video-to-video modes, 720p output.
/filters:quality(82)/media/images/20260408105844_p19jk8j0.webp)
Introducing Video Body Swap on WaveSpeedAI
Video Body Swap replaces the body in a target video with your face. Seamless frame-by-frame compositing with natural blending and consistent identity. REST API, $0.08/second, no cold starts.
/filters:quality(82)/media/images/20260408110527_g86tqkpw.webp)
Introducing InfiniteTalk Fast Video-to-Video Multi on WaveSpeedAI
InfiniteTalk Fast multi-character lip sync converts video and two audio tracks into realistic talking or singing videos. 50% cheaper than standard, up to 10 minutes. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408111359_7pt74rzh.webp)
Introducing InfiniteTalk Video-to-Video Multi on WaveSpeedAI
InfiniteTalk Video-to-Video Multi creates realistic multi-character lip-synced videos from video and two audio inputs. Supports 480p/720p, up to 10 minutes, with full-body coherence. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408111057_76nv36zg.webp)
Introducing WAN 2.1 Mocha on WaveSpeedAI
MoCha performs Video-To-Video character swaps using reference images, replacing a video's character without per-frame pose or depth maps. Ready-to-use REST inference API, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408111250_80jkk4zs.webp)
Introducing WAN 2.2 Fun Control on WaveSpeedAI
Wan2.2-Fun-Control uses Control Codes and multi-modal inputs to generate preset-controlled videos up to 120s at 720p; released under Apache 2.0 for commercial use. Ready-to-use REST API, no coldstarts, affordable.
/filters:quality(82)/media/images/20260408111035_3dw0csq5.webp)
Introducing WAN 2.1 Synthetic To Real Ditto on WaveSpeedAI
WAN 2.1 Synthetic To Real Ditto mirrors motion and facial expressions in video-to-video synthetic-to-real conversion. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408104110_g9dbn0mw.webp)
Introducing PixVerse LipSync on WaveSpeedAI
PixVerse LipSync converts audio into realistic lip-sync animations with advanced algorithms for precise mouth movements and timing for video avatars. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408112331_zcz6s4ks.webp)
Introducing Sync LipSync 2 Pro on WaveSpeedAI
Lipsync-2-pro creates studio-grade lip synchronization for video-to-video editing in minutes, not weeks. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408112334_a88uix94.webp)
Introducing Sync React 1 on WaveSpeedAI
Sync React-1 is a production-grade video-to-video lip-sync model. It maps any speech track to a target face, producing phoneme-accurate visemes and smooth timing while preserving identity, head pose, lighting, and background. Supports emotion and intensity control, multilingual speech, and long take
/filters:quality(82)/media/images/20260408110300_o090kwbj.webp)
Introducing Sam3 Video on WaveSpeedAI
SAM3 Video is a unified foundation model for prompt-based video segmentation. Provide text, point, box, or mask prompts and the model segments and tracks targets across frames with strong temporal consistency. Supports concept-level (“segment anything with concepts”) and multi-object masks for e
/filters:quality(82)/media/images/20260408113125_76v9lo8m.webp)
Introducing Kuaishou Kling V2.6 Pro Motion Control on WaveSpeedAI
Kling 2.6 Pro Motion Control turns reference motion clips (dance, action, gesture) into smooth, realistic animations. Upload a character image (or source video) and a motion video; the model transfers the movement while preserving identity and temporal consistency. Ready-to-use REST API with fast re
/filters:quality(82)/media/images/20260408112251_85t85z7x.webp)
Introducing Decart Lucy Restyle on WaveSpeedAI
Lucy-Restyle is a state-of-the-art text-guided video editing model that transforms videos while preserving original motion, camera angles, and temporal consistency. Edit videos with natural language prompts. Ready-to-use REST inference API, ultra-fast processing, studio-grade quality.
/filters:quality(82)/media/images/20260408104117_58e257su.webp)
Introducing PixVerse Swap on WaveSpeedAI
PixVerse Swap replaces backgrounds, people, and objects directly inside existing videos for quick scene changes and creative edits with natural-looking results. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.
/filters:quality(82)/media/images/1747837572414527813_ptrnkfc8.webp)
Wan2.1-VACE Now Live on WaveSpeedAl: All-in-One Video Creation and Editing Model
We are excited to introduce the wan-2.1-14b-vace, an all-in-one video creation and editing model developed by Alibaba, now live on WaveSpeedAI!