/filters:quality(82)/media/images/1773962750383987480_n3hqzHRZ.webp)
Edit, enhance, and extend your footage with WaveSpeedAI’s AI-powered video editing tools.
/filters:quality(82)/media/images/1773962750383987480_n3hqzHRZ.webp)
Seedance 2.5 (Video-Edit) edits an input video from a natural-language prompt. The reference video drives subject identity, composition, and motion while the model rewrites lighting, style, weather, environment, or specific elements as instructed. Built on ByteDance Seed's unified multimodal architecture for cinematic, motion-stable output. Ready-to-use REST API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1786121872456914873_7EjsCMV4.webp)
Seedance 2.5 (Video-Edit) edits an input video from a natural-language prompt. The reference video drives subject identity, composition, and motion while the model rewrites lighting, style, weather, environment, or specific elements as instructed. Built on ByteDance Seed's unified multimodal architecture for cinematic, motion-stable output. Ready-to-use REST API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1786119739052589494_EwMYENW5.webp)
Seedance 2.5 (Video-Edit Turbo) is the turbo tier for editing an input video from a natural-language prompt — faster, more affordable high-resolution output while preserving subject identity, composition, and motion. Ready-to-use REST API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1789553531699655583_WOX6R1ak.webp)
Wan 3.0 Prime Video Edit edits existing videos with text prompts and optional reference images or audio, supporting prompt-guided scene changes, visual refinements, and multimodal video editing. Inputs longer than 15 seconds are trimmed to the first 15 seconds, with output aspect ratio automatically matched to the input video. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1789553517177857861_OyM2sT76.webp)
Wan 3.0 Video Edit edits existing videos with text prompts and optional reference images or audio, supporting prompt-guided scene changes, visual refinements, and multimodal video editing. Inputs longer than 15 seconds are trimmed to the first 15 seconds, with output aspect ratio automatically matched to the input video. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1788329903545692055_lEMW5fpy.webp)
MiniMax H3 Open Weights Video-Edit edits an input video from a natural-language prompt. The input video drives subject identity, composition, and motion while the model rewrites lighting, style, weather, environment, or specific elements as instructed, with native stereo audio generated in the same pass. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1790573118566544886_jzIR0oxG.webp)
MiniMax H3 Open Weights ControlNet Union generates a new video that follows the motion and composition of a source video. Pose, depth, edges, lines, scribble or grayscale structure is extracted from the source automatically and guides the output, optionally with reference images for the subject or style, with native stereo audio generated in the same pass. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1782150883895961015_UoILIIDE.webp)
Seedance 2.0 Mini is ByteDance's faster, lower-cost video generation model for text to video and image to video. It creates cinematic multi-shot videos with AI camera control, consistent characters across scenes, 720P / 1080P output, 5-12s duration, and flexible aspect ratios. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1784117507698664813_nXmwFOW6.webp)
Seedance 2.0 Mini Video Edit is ByteDance's faster, lower-cost video editing model for prompt-guided video modification. It edits existing videos with cinematic multi-shot quality, AI camera control, consistent characters, 480P / 720P / 1080P / 4K output, 4-15s duration, and flexible aspect ratios. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1777694893579905189_ykuDNW6V.webp)
Alibaba Happy Horse 1.0 (Video Edit) performs prompt-driven video editing with multi-image reference support, supporting 720p/1080p output. Ready-to-use REST API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1782844482067816517_tMQErwk9.webp)
Gemini Omni Flash Video Edit applies natural-language edit instructions to existing videos, enabling prompt-guided changes to scenes, style, motion, and visual details while preserving the original video context. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1787905962736525570_4oY8irAJ.webp)
Gemini Omni 1.1 Flash Video Edit applies natural-language edits to existing videos, supporting output resolutions from 360P to 4K for prompt-guided video modification, scene refinements, creative edits, marketing content, and production workflows. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1784114527178681757_HyHRZ8h2.webp)
Seedance 2.0 (Video-Edit) edits an input video from a natural-language prompt. The reference video drives subject identity, composition, and motion while the model rewrites lighting, style, weather, environment, or specific elements as instructed. Built on ByteDance Seed's unified multimodal architecture for cinematic, motion-stable output. Ready-to-use REST API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1777709912318665482_YFuDNW5f.webp)
Seedance 2.0 (Video-Edit Turbo) is the turbo tier for editing an input video from a natural-language prompt — faster, more affordable high-resolution output while preserving subject identity, composition, and motion. Ready-to-use REST API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1777709897111402859_z86iqAJS.webp)
Seedance 2.0 Fast (Video-Edit Turbo) is the fastest, cheapest turbo tier for editing an input video from a natural-language prompt — high-resolution output with optimized cost and speed. Ready-to-use REST API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1784117974221900004_TrAJS1vF.webp)
Seedance 2.0 Fast (Video-Edit) edits an input video from a natural-language prompt at a faster, cheaper tier. Built on ByteDance Seed's unified multimodal architecture, it preserves subject identity, composition, and motion while rewriting lighting, style, weather, environment, or specific elements as instructed. Ready-to-use REST API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408111403_q1ar1bex.webp)
Audio-driven InfiniteTalk turns one video plus audio into realistic talking or singing videos with lip-sync in 480p or 720p. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408104408_wr1r10h3.webp)
WAN 2.7 Video Edit performs prompt-driven video editing with multi-image reference support, supporting 720p/1080p output. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1782717317688496873_qX6fpyHQ.webp)
Luma Ray 3.2 Video Edit is a fast AI video-to-video editing model that re-renders an existing source video from a text prompt while preserving the original motion and timing. Ready-to-use REST inference API for video restyling, creative edits, product videos, advertising creatives, social media clips, visual storytelling, and professional video editing workflows with simple integration, no coldstarts, and affordable pricing.
/filters:quality(82)/media/images/1788253152394476475_GrQ09irB.webp)
Kling O3 Omni 4K Video Edit edits input videos in 4K with natural-language instructions and optional image references, supporting prompt-guided video modification, visual style updates, scene refinements, motion control, and high-quality cinematic editing. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408113108_8lwvxxl5.webp)
Kling Omni Video O3 Video-Edit enables conversational video editing through natural language commands. Remove objects, change backgrounds, modify styles, adjust weather/lighting, and transform scenes with simple text instructions like 'remove pedestrians' or 'change daytime to dusk'. Ready-to-use REST API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1791277754645280429_8KT2bkuD.webp)
AI Video Captioner adds animated, word-by-word captions to any talking video: 12 caption styles, AI keyword highlights, matching emoji, optional silence and filler removal, about 100 languages with translation, plus an SRT subtitle file. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408113021_v77zgrk2.webp)
Kling Omni Video O3 Video-Edit (Standard) enables natural-language video edits: remove or replace objects, swap backgrounds, restyle scenes, change weather/lighting, and apply localized 3-10s transformations with strong temporal consistency. Built for stable production use with a ready-to-use REST API, no cold starts, and predictable pricing.
/filters:quality(82)/media/images/1790000904404251148_XPd7GS5n.webp)
AI Video Editor Auto Clip cuts long or raw footage down to a short highlight clip: it reviews one or several videos end to end, picks the moments that matter, follows your prompt for topic, style and length, and delivers a captioned clip of roughly 15 to 60 seconds. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1790000903351131363_qe7BWoNb.webp)
AI Video Editor Talking Head cleans up a talking-to-camera video: it removes filler words, false starts, repeats and dead air, keeps the speaker framed, and burns in accurate word-timed captions. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1789204884164289691_UxPZ8gY7.webp)
Video Restorer cleans up damaged footage: strip compression artifacts from low-bitrate video or bring out-of-focus footage back into sharp focus, keeping the subject, framing and motion exactly as they are. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1789373313254033883_XpajsBKU.webp)
Video Colorizer restores natural color to black-and-white, monochrome or faded footage while keeping every subject, the framing and the motion exactly as they are. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1789373573992139107_MifpyHQZ.webp)
Video Cleaner removes people, vehicles and other moving subjects from a video and rebuilds the background behind them, producing a clean plate of the same shot with the camera motion intact. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1789373960834369719_9yHR0ajt.webp)
Video Relighter re-lights outdoor footage to a chosen lighting preset and sun direction, or re-renders the same shot at night, keeping composition, camera movement and subject motion intact. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408112923_6t50m3fi.webp)
Kling Omni Video O1 Video-Edit enables conversational video editing through natural language commands. Remove objects, change backgrounds, modify styles, adjust weather/lighting, and transform scenes with simple text instructions like 'remove pedestrians' or 'change daytime to dusk'. Ready-to-use REST API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408112929_0qehcs8k.webp)
Kling Omni Video O1 Video-Edit enables conversational video editing through natural language commands. Remove objects, change backgrounds, modify styles, adjust weather/lighting, and transform scenes with simple text instructions like 'remove pedestrians' or 'change daytime to dusk'. Ready-to-use REST API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408105712_yysellho.webp)
WaveSpeedAI Video Outpainter expands any video beyond its original boundaries while preserving motion, identity, and scene coherence. Perfect for aspect-ratio changes, reframing, adding safe margins, or generating new visual context without cropping or losing content.
/filters:quality(82)/media/images/20260408112948_ocz12m2q.webp)
Kling Omni Video O1 Video-Edit (Standard) enables natural-language video edits: remove or replace objects, swap backgrounds, restyle scenes, change weather/lighting, and apply localized 3–10s transformations with strong temporal consistency. Built for stable production use with a ready-to-use REST API, no cold starts, and predictable pricing.
/filters:quality(82)/media/images/20260408110532_2m95tzbp.webp)
Audio-driven infinitetalk-fast turns one video plus audio into realistic talking or singing videos with lip-sync. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1789113620754515887_hYMV3cmw.webp)
FLUX 3 Video Edit applies prompt-guided changes to existing videos while preserving motion, timing, and framing, supporting precise video modification, visual style updates, scene refinements, creative edits, ads, and production workflows. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408104037_00c2v63i.webp)
LTX-2 Retake performs targeted retakes on any section of a video—replace visuals, audio, or both—while preserving timing and continuity with $0.1 per output video second. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1789970314573989572_WyXbKS2b.webp)
Depth Anything V3 Video turns any video into a temporally consistent depth map video with no flicker, ideal for replicating camera moves and motion with depth-controlled video generation. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1789970316193742632_cKvENW6g.webp)
Depth Anything Video estimates temporally consistent depth maps from video input, with stable depth across the whole clip and no flicker. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408111241_3bzr5rd1.webp)
Wan 2.2 Video Edit lets you modify videos via text prompts (e.g., change clothing or characters). Powered by Wan 2.2, it supports 480p ($0.20/5s) and 720p ($0.40/5s), up to 120s. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
Video Edit 컬렉션의 모든 모델을 단일 REST API로 실행하세요. 생성당 과금 — 구독 없음, 최소 요금 없음 — 99.9% 가동률 인프라에서 업계 최고의 지연 시간을 제공합니다.
모든 Video Edit 모델에 대한 호출당 가격. 가격은 각 모델 페이지에 표시되며 플랫폼 수수료는 추가되지 않습니다.
대부분의 Video Edit 이미지 모델은 2초 이내에 완료됩니다. 비디오 및 3D 모델은 셀프 호스팅 대안보다 몇 배 더 빠릅니다.
다중 리전 페일오버와 자동 재시도로 프로바이더 장애 중에도 운영 트래픽을 온라인 상태로 유지합니다.
각 모델에는 모델 페이지에 호출당 자체 가격이 표시되어 있습니다. 성공한 생성 단위로 청구되며 구독 요금이나 최소 요금은 없습니다.
이 컬렉션의 이미지 모델은 일반적으로 2초 이내에 완료됩니다. 비디오 및 3D 모델은 길이와 해상도에 따라 다르지만 보통 셀프 호스팅 실행보다 몇 배 더 빠릅니다.
조건을 충족하는 신규 계정은 신용카드 없이 Video Edit 모델을 체험할 수 있는 $1 프로모션 크레딧을 받을 수 있습니다. 모든 가입에 체험 크레딧이 보장되지는 않으므로 생성 전에 계정 잔액을 확인하세요.
표준 계정에는 넉넉한 동시 작업 제한이 있습니다. Enterprise 플랜은 맞춤형 RPM, 더 높은 동시성, 전용 용량을 제공합니다 — 자세한 내용은 영업팀에 문의하세요.