/filters:quality(82)/media/images/1773962750383987480_n3hqzHRZ.webp)
Edit, enhance, and extend your footage with WaveSpeedAI’s AI-powered video editing tools.
/filters:quality(82)/media/images/1773962750383987480_n3hqzHRZ.webp)
Seedance 2.5 (Video-Edit) edits an input video from a natural-language prompt. The reference video drives subject identity, composition, and motion while the model rewrites lighting, style, weather, environment, or specific elements as instructed. Built on ByteDance Seed's unified multimodal architecture for cinematic, motion-stable output. Ready-to-use REST API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1786121872456914873_7EjsCMV4.webp)
Seedance 2.5 (Video-Edit) edits an input video from a natural-language prompt. The reference video drives subject identity, composition, and motion while the model rewrites lighting, style, weather, environment, or specific elements as instructed. Built on ByteDance Seed's unified multimodal architecture for cinematic, motion-stable output. Ready-to-use REST API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1786119739052589494_EwMYENW5.webp)
Seedance 2.5 (Video-Edit Turbo) is the turbo tier for editing an input video from a natural-language prompt — faster, more affordable high-resolution output while preserving subject identity, composition, and motion. Ready-to-use REST API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1789553531699655583_WOX6R1ak.webp)
Wan 3.0 Prime Video Edit edits existing videos with text prompts and optional reference images or audio, supporting prompt-guided scene changes, visual refinements, and multimodal video editing. Inputs longer than 15 seconds are trimmed to the first 15 seconds, with output aspect ratio automatically matched to the input video. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1789553517177857861_OyM2sT76.webp)
Wan 3.0 Video Edit edits existing videos with text prompts and optional reference images or audio, supporting prompt-guided scene changes, visual refinements, and multimodal video editing. Inputs longer than 15 seconds are trimmed to the first 15 seconds, with output aspect ratio automatically matched to the input video. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1788329903545692055_lEMW5fpy.webp)
MiniMax H3 Open Weights Video-Edit edits an input video from a natural-language prompt. The input video drives subject identity, composition, and motion while the model rewrites lighting, style, weather, environment, or specific elements as instructed, with native stereo audio generated in the same pass. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1790573118566544886_jzIR0oxG.webp)
MiniMax H3 Open Weights ControlNet Union generates a new video that follows the motion and composition of a source video. Pose, depth, edges, lines, scribble or grayscale structure is extracted from the source automatically and guides the output, optionally with reference images for the subject or style, with native stereo audio generated in the same pass. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1782150883895961015_UoILIIDE.webp)
Seedance 2.0 Mini is ByteDance's faster, lower-cost video generation model for text to video and image to video. It creates cinematic multi-shot videos with AI camera control, consistent characters across scenes, 720P / 1080P output, 5-12s duration, and flexible aspect ratios. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1784117507698664813_nXmwFOW6.webp)
Seedance 2.0 Mini Video Edit is ByteDance's faster, lower-cost video editing model for prompt-guided video modification. It edits existing videos with cinematic multi-shot quality, AI camera control, consistent characters, 480P / 720P / 1080P / 4K output, 4-15s duration, and flexible aspect ratios. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1777694893579905189_ykuDNW6V.webp)
Alibaba Happy Horse 1.0 (Video Edit) performs prompt-driven video editing with multi-image reference support, supporting 720p/1080p output. Ready-to-use REST API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1782844482067816517_tMQErwk9.webp)
Gemini Omni Flash Video Edit applies natural-language edit instructions to existing videos, enabling prompt-guided changes to scenes, style, motion, and visual details while preserving the original video context. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1787905962736525570_4oY8irAJ.webp)
Gemini Omni 1.1 Flash Video Edit applies natural-language edits to existing videos, supporting output resolutions from 360P to 4K for prompt-guided video modification, scene refinements, creative edits, marketing content, and production workflows. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1784114527178681757_HyHRZ8h2.webp)
Seedance 2.0 (Video-Edit) edits an input video from a natural-language prompt. The reference video drives subject identity, composition, and motion while the model rewrites lighting, style, weather, environment, or specific elements as instructed. Built on ByteDance Seed's unified multimodal architecture for cinematic, motion-stable output. Ready-to-use REST API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1777709912318665482_YFuDNW5f.webp)
Seedance 2.0 (Video-Edit Turbo) is the turbo tier for editing an input video from a natural-language prompt — faster, more affordable high-resolution output while preserving subject identity, composition, and motion. Ready-to-use REST API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1777709897111402859_z86iqAJS.webp)
Seedance 2.0 Fast (Video-Edit Turbo) is the fastest, cheapest turbo tier for editing an input video from a natural-language prompt — high-resolution output with optimized cost and speed. Ready-to-use REST API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1784117974221900004_TrAJS1vF.webp)
Seedance 2.0 Fast (Video-Edit) edits an input video from a natural-language prompt at a faster, cheaper tier. Built on ByteDance Seed's unified multimodal architecture, it preserves subject identity, composition, and motion while rewriting lighting, style, weather, environment, or specific elements as instructed. Ready-to-use REST API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408111403_q1ar1bex.webp)
Audio-driven InfiniteTalk turns one video plus audio into realistic talking or singing videos with lip-sync in 480p or 720p. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408104408_wr1r10h3.webp)
WAN 2.7 Video Edit performs prompt-driven video editing with multi-image reference support, supporting 720p/1080p output. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1782717317688496873_qX6fpyHQ.webp)
Luma Ray 3.2 Video Edit is a fast AI video-to-video editing model that re-renders an existing source video from a text prompt while preserving the original motion and timing. Ready-to-use REST inference API for video restyling, creative edits, product videos, advertising creatives, social media clips, visual storytelling, and professional video editing workflows with simple integration, no coldstarts, and affordable pricing.
/filters:quality(82)/media/images/1788253152394476475_GrQ09irB.webp)
Kling O3 Omni 4K Video Edit edits input videos in 4K with natural-language instructions and optional image references, supporting prompt-guided video modification, visual style updates, scene refinements, motion control, and high-quality cinematic editing. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408113108_8lwvxxl5.webp)
Kling Omni Video O3 Video-Edit enables conversational video editing through natural language commands. Remove objects, change backgrounds, modify styles, adjust weather/lighting, and transform scenes with simple text instructions like 'remove pedestrians' or 'change daytime to dusk'. Ready-to-use REST API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1791277754645280429_8KT2bkuD.webp)
AI Video Captioner adds animated, word-by-word captions to any talking video: 12 caption styles, AI keyword highlights, matching emoji, optional silence and filler removal, about 100 languages with translation, plus an SRT subtitle file. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408113021_v77zgrk2.webp)
Kling Omni Video O3 Video-Edit (Standard) enables natural-language video edits: remove or replace objects, swap backgrounds, restyle scenes, change weather/lighting, and apply localized 3-10s transformations with strong temporal consistency. Built for stable production use with a ready-to-use REST API, no cold starts, and predictable pricing.
/filters:quality(82)/media/images/1790000904404251148_XPd7GS5n.webp)
AI Video Editor Auto Clip cuts long or raw footage down to a short highlight clip: it reviews one or several videos end to end, picks the moments that matter, follows your prompt for topic, style and length, and delivers a captioned clip of roughly 15 to 60 seconds. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1790000903351131363_qe7BWoNb.webp)
AI Video Editor Talking Head cleans up a talking-to-camera video: it removes filler words, false starts, repeats and dead air, keeps the speaker framed, and burns in accurate word-timed captions. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1789204884164289691_UxPZ8gY7.webp)
Video Restorer cleans up damaged footage: strip compression artifacts from low-bitrate video or bring out-of-focus footage back into sharp focus, keeping the subject, framing and motion exactly as they are. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1789373313254033883_XpajsBKU.webp)
Video Colorizer restores natural color to black-and-white, monochrome or faded footage while keeping every subject, the framing and the motion exactly as they are. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1789373573992139107_MifpyHQZ.webp)
Video Cleaner removes people, vehicles and other moving subjects from a video and rebuilds the background behind them, producing a clean plate of the same shot with the camera motion intact. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1789373960834369719_9yHR0ajt.webp)
Video Relighter re-lights outdoor footage to a chosen lighting preset and sun direction, or re-renders the same shot at night, keeping composition, camera movement and subject motion intact. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408112923_6t50m3fi.webp)
Kling Omni Video O1 Video-Edit enables conversational video editing through natural language commands. Remove objects, change backgrounds, modify styles, adjust weather/lighting, and transform scenes with simple text instructions like 'remove pedestrians' or 'change daytime to dusk'. Ready-to-use REST API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408112929_0qehcs8k.webp)
Kling Omni Video O1 Video-Edit enables conversational video editing through natural language commands. Remove objects, change backgrounds, modify styles, adjust weather/lighting, and transform scenes with simple text instructions like 'remove pedestrians' or 'change daytime to dusk'. Ready-to-use REST API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408105712_yysellho.webp)
WaveSpeedAI Video Outpainter expands any video beyond its original boundaries while preserving motion, identity, and scene coherence. Perfect for aspect-ratio changes, reframing, adding safe margins, or generating new visual context without cropping or losing content.
/filters:quality(82)/media/images/20260408112948_ocz12m2q.webp)
Kling Omni Video O1 Video-Edit (Standard) enables natural-language video edits: remove or replace objects, swap backgrounds, restyle scenes, change weather/lighting, and apply localized 3–10s transformations with strong temporal consistency. Built for stable production use with a ready-to-use REST API, no cold starts, and predictable pricing.
/filters:quality(82)/media/images/20260408110532_2m95tzbp.webp)
Audio-driven infinitetalk-fast turns one video plus audio into realistic talking or singing videos with lip-sync. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1789113620754515887_hYMV3cmw.webp)
FLUX 3 Video Edit applies prompt-guided changes to existing videos while preserving motion, timing, and framing, supporting precise video modification, visual style updates, scene refinements, creative edits, ads, and production workflows. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408104037_00c2v63i.webp)
LTX-2 Retake performs targeted retakes on any section of a video—replace visuals, audio, or both—while preserving timing and continuity with $0.1 per output video second. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1789970314573989572_WyXbKS2b.webp)
Depth Anything V3 Video turns any video into a temporally consistent depth map video with no flicker, ideal for replicating camera moves and motion with depth-controlled video generation. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/1789970316193742632_cKvENW6g.webp)
Depth Anything Video estimates temporally consistent depth maps from video input, with stable depth across the whole clip and no flicker. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
/filters:quality(82)/media/images/20260408111241_3bzr5rd1.webp)
Wan 2.2 Video Edit lets you modify videos via text prompts (e.g., change clothing or characters). Powered by Wan 2.2, it supports 480p ($0.20/5s) and 720p ($0.40/5s), up to 120s. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.
Jalankan model apa pun di koleksi Video Edit melalui satu REST API. Bayar per generasi — tanpa langganan, tanpa minimum — dengan latensi terdepan di infrastruktur dengan uptime 99,9%.
Harga per panggilan untuk setiap model Video Edit. Harga tercantum di halaman setiap model — tanpa biaya platform tambahan.
Sebagian besar model gambar Video Edit selesai di bawah 2 detik. Model video dan 3D beberapa kali lebih cepat daripada alternatif yang di-hosting sendiri.
Failover multi-region dan retry otomatis menjaga lalu lintas produksi tetap online — bahkan saat provider mengalami gangguan.
Setiap model memiliki harga per panggilan tersendiri yang tercantum di halaman model. Kami menagih per generasi berhasil, tanpa biaya langganan atau minimum.
Model gambar di koleksi ini biasanya selesai di bawah 2 detik. Model video dan 3D bergantung pada durasi dan resolusi, tetapi biasanya beberapa kali lebih cepat dari run yang di-hosting sendiri.
Akun baru yang memenuhi syarat dapat menerima kredit promosi $1 untuk mencoba model Video Edit tanpa kartu kredit. Kredit uji coba tidak dijamin untuk setiap pendaftaran; periksa saldo akun sebelum membuat konten.
Akun standar memiliki batas concurrent job yang murah hati. Paket Enterprise menawarkan RPM khusus, concurrency lebih tinggi, dan kapasitas khusus — hubungi sales untuk detailnya.
Telusuri katalog lengkap kami dari model AI tercanggih — gambar, video, 3D, audio, LLM, dan banyak lagi.
wavespeed.ai/models →Integrasikan AI ke dalam aplikasi Anda sendiri. API RESTful dengan pustaka klien — tanpa cold start, bayar sesuai penggunaan.
wavespeed.ai/docs →