Nano Banana 2.1 现已上线 — Google 最新模型 | 立即体验 →
Vidu Models

Vidu Models

Shengshu's Vidu offers comprehensive AI video generation solutions with multiple specialized models and precise creative control.

Shengshu's Vidu offers comprehensive AI video generation solutions with multiple specialized models and precise creative control.

所有模型

38 个模型
vidu/q3-ad
image-to-video$0.1500

vidu/q3-ad

Vidu Q3 Ad Video generates commercial ad videos from 1 to 7 reference images with prompt guidance, supporting 720P / 1080P output and synchronized audio for product ads, brand campaigns, marketing creatives, and promotional videos. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

vidu/q3/drama-clip
image-to-video$1.1200

vidu/q3/drama-clip

Vidu Q3 Drama Clip generates 8-12 second script-driven drama videos from structured assets, including characters, scenes, and tools. It is ideal for compact story scenes, storyboard shots, and focused narrative moments. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

vidu/q3/drama
image-to-video$1.1200

vidu/q3/drama

Vidu Q3 Drama generates complete script-driven drama videos from scripts and structured assets, including characters, scenes, tools, and references. It plans the narrative structure, scene pacing, and transitions to create a story-driven drama in one request, supporting up to 180 seconds. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

vidu/q3/image-to-video
image-to-video$0.3500

vidu/q3/image-to-video

Vidu Q3 Image-to-Video turns text prompts into high-quality videos with exceptional visual fidelity and diverse motion. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

vidu/q3/text-to-video
text-to-video$0.3500

vidu/q3/text-to-video

Vidu Q3 Text-to-Video turns text prompts into high-quality videos with exceptional visual fidelity and diverse motion. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

vidu/q3-turbo/image-to-video
image-to-video$0.3000

vidu/q3-turbo/image-to-video

Vidu Q3 Turbo Image-to-Video animates static images with high-quality motion and faster processing. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

vidu/q3-pro/image-to-video
image-to-video$0.2500

vidu/q3-pro/image-to-video

Vidu Q3 Pro Image-to-Video animates still images with high-quality motion via viduq3-pro (1–16s). Billing follows Vidu's published Q3-pro per-second rates by resolution. Ready-to-use REST inference API on WaveSpeed.

vidu/q3/image-to-video-pro
image-to-video$0.4500

vidu/q3/image-to-video-pro

Vidu Q3 Image-to-Video Pro generates high-resolution videos (720p/1080p/2K/4K) from images with exceptional visual fidelity and diverse motion. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

vidu/q3/reference-to-video
reference-to-video$0.3500

vidu/q3/reference-to-video

Vidu Q3 Reference-to-Video Mix generates multi-entity consistent videos from 1-4 reference images with text prompt guidance. Supports 360p to 1080p resolutions, up to 16 seconds duration, multiple aspect ratios, and optional audio generation. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

vidu/q3/start-end-to-video
image-to-video$0.3500

vidu/q3/start-end-to-video

Vidu Q3 Start End Image-to-Video turns text prompts into high-quality videos with exceptional visual fidelity and diverse motion. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

vidu/q3-turbo/start-end-to-video
image-to-video$0.3000

vidu/q3-turbo/start-end-to-video

Vidu Q3 Turbo Start-End-to-Video creates smooth transitions between two images with faster processing. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

vidu/q3-pro/start-end-to-video
image-to-video$0.2500

vidu/q3-pro/start-end-to-video

Vidu Q3 Pro Start-End-to-Video creates smooth transitions between two keyframes with viduq3-pro (1–16s). Billing follows Vidu's published Q3-pro per-second rates by resolution. Ready-to-use REST inference API on WaveSpeed.

vidu/q3-pro/text-to-video
text-to-video$0.2500

vidu/q3-pro/text-to-video

Vidu Q3 Pro Text to Video is a fast AI video generation model that creates high-quality, audio-capable videos from text prompts with support for 1–16 second outputs. Ready-to-use REST inference API for cinematic clips, advertising creatives, social media videos, product visuals, storytelling, and professional text-to-video workflows with simple integration, no coldstarts, and affordable pricing.

vidu/image-to-video-2.0
image-to-video$0.3000

vidu/image-to-video-2.0

Vidu Image to Video 2.0 converts images into smooth-transition videos with exceptional visual quality and diverse, natural motion. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

vidu/reference-to-video-2.0
reference-to-video$0.4000

vidu/reference-to-video-2.0

Vidu Reference-to-Video 2.0 turns references into videos that preserve characters, objects, and environments with Multi-Entity Consistency. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

vidu/start-end-to-video-2.0
image-to-video$0.3000

vidu/start-end-to-video-2.0

Vidu Start-End to Video 2.0 generates smooth transition videos interpolating between given start and end images for natural morphing effects. Ready-to-use REST API, best performance, no coldstarts, affordable pricing.

vidu/image-to-video
image-to-video$0.2000

vidu/image-to-video

Vidu Image-to-Video converts images into smooth-transition videos with high visual quality and diverse motion for cinematic results. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

vidu/text-to-video
text-to-video$0.4000

vidu/text-to-video

Vidu Text to Video converts text prompts into high-quality 720p videos with exceptional visual fidelity and diverse motion dynamics. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

vidu/start-end-to-video
image-to-video$0.2000

vidu/start-end-to-video

Vidu Start-End to Video converts a start and end image into a smooth transition Image-to-Video clip that morphs scenes seamlessly. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

vidu/image-to-video-q2-pro
image-to-video$0.1500

vidu/image-to-video-q2-pro

Vidu Q2 Pro turns a single still image into smooth, cinematic image-to-video with stable motion, clean edges, and consistent lighting. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

vidu/image-to-video-q2-turbo
image-to-video$0.1000

vidu/image-to-video-q2-turbo

Vidu Q2 Turbo Image-to-Video turns a single image into smooth, cinematic motion with fast, high-quality output. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

vidu/text-to-video-2.0
text-to-video$0.3000

vidu/text-to-video-2.0

Vidu Text-to-Video 2.0 converts text prompts into high-quality 720p videos with exceptional visual detail and diverse motion dynamics. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

vidu/one-click-v2/mv
audio-to-video$0.2500

vidu/one-click-v2/mv

Vidu One-Click V2 MV transforms images and audio into videos with camera movements and subtitle support. Create professional video content with dynamic shots and text overlays in one click. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

vidu/q2-pro/image-to-video-fast
image-to-video$0.0800

vidu/q2-pro/image-to-video-fast

Vidu Q2 Pro Fast Image to Video generates high-quality videos from a single image with faster generation speed. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.

vidu/q2-pro/start-end-to-video-fast
image-to-video$0.0800

vidu/q2-pro/start-end-to-video-fast

Vidu Q2 Pro Fast Start-End to Video generates smooth video transitions between start and end images with faster generation speed. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.

vidu/reference-to-image-q2
image-to-image$0.0400

vidu/reference-to-image-q2

Vidu Reference-to-Image Q2 generates high-quality images from 1–7 reference images plus a text prompt, preserving style and composition while allowing controlled changes to subjects, backgrounds, and fine details. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.

vidu/text-to-image-q2
text-to-image$0.0300

vidu/text-to-image-q2

Vidu Text-to-Image Q2 converts text prompts into high-quality images with exceptional visual detail and creative flexibility. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

vidu/text-to-video-q2
text-to-video$0.1000

vidu/text-to-video-q2

Vidu Q2 Text-to-Video turns text prompts into high-quality videos with exceptional visual fidelity and diverse motion. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

vidu/template/halloween
video-effects$0.0500

vidu/template/halloween

Vidu Halloween Templates delivers ready-made image and video templates for spooky promos and event invites with overlays. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

vidu/reference-to-video-q2
reference-to-video$0.1000

vidu/reference-to-video-q2

Vidu Q2 is an Image-to-Video and Reference-to-Video model that emphasizes subtle facial expressions and smooth push-pull camera moves for natural motion. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

vidu/q2-turbo/video-extend
video-extend$0.0500

vidu/q2-turbo/video-extend

Vidu Q2 Turbo Extend Video seamlessly extends existing videos by 1-7 seconds with consistent motion and scene continuity. Supports optional end-frame image guidance for precise control. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

vidu/q2-pro/video-extend
video-extend$0.0750

vidu/q2-pro/video-extend

Vidu Q2 Pro Extend Video seamlessly extends existing videos by 1-7 seconds with high-quality motion and scene continuity. Supports optional end-frame image guidance for precise control. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

vidu/start-end-to-video-q2-turbo
image-to-video$0.1000

vidu/start-end-to-video-q2-turbo

Vidu Q2 Turbo Start-End to Video creates smooth Image-to-Video transitions between start and end images with fast high-quality results. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

vidu/start-end-to-video-q2-pro
image-to-video$0.1500

vidu/start-end-to-video-q2-pro

Vidu Q2 Pro Start-End to Video produces smooth image-to-video transitions between start and end images for seamless morphs. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

vidu/text-to-video-q1
text-to-video$0.4000

vidu/text-to-video-q1

Vidu Text-to-Video Q1 converts text prompts into high-quality videos with exceptional visual fidelity and motion diversity. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

vidu/image-to-video-q1
image-to-video$0.4000

vidu/image-to-video-q1

Vidu Image-to-Video creates smooth transition videos from specified start and end images, producing seamless image-to-video outputs for presentations and storytelling. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

vidu/start-end-to-video-q1
image-to-video$0.4000

vidu/start-end-to-video-q1

Vidu Q1 Start-End To Video turns specified start and end images into smooth image-to-video transitions for morphs and scene fades. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

vidu/reference-to-video-q1
reference-to-video$0.4000

vidu/reference-to-video-q1

Generate videos from reference images while keeping characters, objects, and scene identity consistent using Multi-Entity Consistency. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

Vidu Models

Vidu is Shengshu Technology's advanced AI image and video generation model suite, covering text-to-video, image-to-video, reference-to-video, start-end frame generation, text-to-image, reference-to-image, music video creation, short drama generation, and commercial ad video workflows. The collection brings together Vidu Q3, Q2, Q1, 2.0, and specialized creative models, giving creators and developers flexible tools for cinematic storytelling, social media content, product videos, branded campaigns, and scalable creative production.

Built for professional video creation, Vidu focuses on strong prompt understanding, smooth motion, stable subject consistency, cinematic camera control, and robust temporal coherence. It supports both general-purpose generation and scenario-specific workflows, from animating still images and generating videos from prompts to creating short dramas, music videos, and commercial advertising content.

Core Model Capabilities

Featured Scenario Models:

Vidu includes featured scenario models for specialized production workflows. vidu/drama is designed for short-form drama, episodic storytelling, emotionally expressive scenes, and character-focused narrative content. vidu/ad is built for commercial advertising workflows, helping create polished product showcases, brand videos, promotional creatives, and conversion-focused campaign assets. vidu/one-click-v2/mv transforms images and audio into dynamic music videos with camera movement and subtitle support, making it ideal for creator content, music promotion, social clips, and fast visual-audio production.

Text-to-Video Generation:

Generate cinematic videos directly from natural-language prompts with strong prompt adherence, coherent scene structure, controllable camera movement, and natural multi-character interactions.

Image-to-Video Generation:

Animate still images into dynamic video clips while preserving subject identity, composition, lighting, and visual style. Vidu's newer Q3 and Q2 models provide stronger motion quality, structural fidelity, and cinematic realism for more polished outputs.

Reference-to-Video Generation:

Create videos guided by reference images or visual inputs, helping maintain character identity, object appearance, wardrobe consistency, visual style, and scene continuity across generated clips.

Start-End Frame Video Generation:

Generate smooth motion between a defined starting frame and ending frame, making it useful for transitions, reveals, storyboard-based motion design, product shots, and controlled narrative progression.

Image Generation:

Create cinematic key visuals, posters, thumbnails, concept frames, and reference-guided still images from text prompts or multiple reference images.

Fast and Pro Workflows:

Choose newer Q3 models for stronger overall quality, Q2 Pro models for polished production assets, Turbo and Fast variants for rapid iteration, and earlier models for lightweight drafts or cost-efficient generation.

Production-Ready Video API:

Access Vidu models through scalable APIs for automated video generation, social media production, advertising pipelines, product content, creative testing, and commercial media workflows.

Vidu AI Models on WaveSpeedAI give creators and developers fast access to a complete AI video and image generation toolkit with flexible pricing, scalable API access, and production-ready creative quality.

Vidu Models API — 价格与性能

通过单一 REST API 运行 Vidu Models 系列中的任意模型。按生成计费 — 无订阅、无最低消费 — 在 99.9% 可用性的基础设施上提供行业领先的延迟。

为什么在 WaveSpeedAI 上运行 Vidu Models

透明定价

每个 Vidu Models 模型都有按调用计价。价格在每个模型的页面上列出 — 不收取额外的平台费。

为低延迟优化

大多数 Vidu Models 图像模型在 2 秒内完成。视频和 3D 模型比自托管方案快数倍。

99.9% 可用性

多区域故障转移和自动重试可确保您的生产流量保持在线 — 即使在供应商故障期间。

常见问题

Vidu Models API 多少钱?+

每个模型在其模型页面上都列有自己的按调用价格。我们按每次成功生成计费,没有订阅费或最低消费。

Vidu Models 模型在 WaveSpeedAI 上有多快?+

本系列中的图像模型通常在 2 秒内完成。视频和 3D 模型取决于时长和分辨率,但通常比自托管运行快数倍。

不用信用卡可以试用 API 吗?+

符合条件的新账户可能获得 $1 的推广额度,用于无需信用卡试用 Vidu Models 模型。并非每次注册都会获得试用额度,请在生成前查看账户余额。

有速率限制吗?+

标准账户有充足的并发任务限制。企业版计划提供自定义 RPM、更高并发和专用容量 — 详情请联系销售。