Nano Banana 2.1 現已上線 — Google 最新模型 | 立即體驗 →
Kling O3 Models

Kling O3 Models

Kling Omni3 enables unified audio-video creation in a single step, delivering finer detail, more fluid motion, and deeper, more immersive narrative experiences.

Kling Omni3 enables unified audio-video creation in a single step, delivering finer detail, more fluid motion, and deeper, more immersive narrative experiences.

所有模型

16 個模型
kwaivgi/kling-video-o3-std/image-to-video
image-to-video$0.4200

kwaivgi/kling-video-o3-std/image-to-video

Kling Omni Video O3 (Standard) Image-to-Video transforms static images into dynamic cinematic videos using MVL (Multi-modal Visual Language) technology. Maintains subject consistency while adding natural motion, physics simulation, and seamless scene dynamics. Supports audio generation. Ready-to-use REST API, best performance, no coldstarts, affordable pricing.

kwaivgi/kling-video-o3-4k/image-to-video
image-to-video$2.1000

kwaivgi/kling-video-o3-4k/image-to-video

Kling Video O3 4K Image-to-Video transforms static images into dynamic cinematic 4K videos. Maintains subject consistency while adding natural motion, physics simulation, and seamless scene dynamics. Supports start/end frame control, multi-prompt, and optional audio generation. Ready-to-use REST API, best performance, no coldstarts, affordable pricing.

kwaivgi/kling-video-o3-pro/image-to-video
image-to-video$0.5600

kwaivgi/kling-video-o3-pro/image-to-video

Kling Omni Video O3 Image-to-Video transforms static images into dynamic cinematic videos using MVL (Multi-modal Visual Language) technology. Maintains subject consistency while adding natural motion, physics simulation, and seamless scene dynamics. Supports audio generation. Ready-to-use REST API, best performance, no coldstarts, affordable pricing.

kwaivgi/kling-video-o3-pro/reference-to-video
reference-to-video$0.5600

kwaivgi/kling-video-o3-pro/reference-to-video

Kling Omni Video O3 Reference-to-Video generates creative videos using character, prop, or scene references from multiple viewpoints. Extracts subject features and creates new video content while maintaining identity consistency across frames. Supports audio generation. Ready-to-use REST API, best performance, no cold starts, affordable pricing.

kwaivgi/kling-video-o3-4k/reference-to-video
reference-to-video$2.1000

kwaivgi/kling-video-o3-4k/reference-to-video

Kling Video O3 4K Reference-to-Video generates creative 4K videos using character, prop, or scene references from multiple viewpoints. Extracts subject features and creates new video content while maintaining identity consistency across frames. Supports multi-reference images, video guidance, and optional audio generation. Ready-to-use REST API, best performance, no cold starts, affordable pricing.

kwaivgi/kling-video-o3-4k/video-edit
video-to-video$2.3100

kwaivgi/kling-video-o3-4k/video-edit

Kling O3 Omni 4K Video Edit edits input videos in 4K with natural-language instructions and optional image references, supporting prompt-guided video modification, visual style updates, scene refinements, motion control, and high-quality cinematic editing. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

kwaivgi/kling-video-o3-4k/video-reference
video-to-video$2.3100

kwaivgi/kling-video-o3-4k/video-reference

Kling O3 Omni 4K Video Reference generates 4K AI videos guided by an input video, text prompt, and optional reference images, supporting video-to-video creation with visual reference guidance, motion consistency, camera control, and high-quality cinematic output. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

kwaivgi/kling-video-o3-std/reference-to-video
reference-to-video$0.4200

kwaivgi/kling-video-o3-std/reference-to-video

Kling Omni Video O3 (Standard) Reference-to-Video generates creative videos using character, prop, or scene references from multiple viewpoints. Extracts subject features and creates new video content while maintaining identity consistency across frames. Supports audio generation. Ready-to-use REST API, best performance, no cold starts, affordable pricing.

kwaivgi/kling-video-o3-4k/text-to-video
text-to-video$2.1000

kwaivgi/kling-video-o3-4k/text-to-video

Kling Video O3 4K generates cinematic 4K videos from text prompts with subject consistency, natural physics simulation, and precise semantic understanding. Supports multi-prompt scene transitions, element references, and optional audio generation. Ready-to-use REST API, best performance, no coldstarts, affordable pricing.

kwaivgi/kling-video-o3-pro/text-to-video
text-to-video$0.5600

kwaivgi/kling-video-o3-pro/text-to-video

Kling Omni Video O3 is Kuaishou's advanced unified multi-modal video model with MVL (Multi-modal Visual Language) technology. Text-to-Video mode generates cinematic videos from text prompts with subject consistency, natural physics simulation, and precise semantic understanding. Supports audio generation. Ready-to-use REST API, best performance, no coldstarts, affordable pricing.

kwaivgi/kling-video-o3-std/text-to-video
text-to-video$0.4200

kwaivgi/kling-video-o3-std/text-to-video

Kling Omni Video O3 (Standard) is Kuaishou's advanced unified multi-modal video model with MVL (Multi-modal Visual Language) technology. Text-to-Video mode generates cinematic videos from text prompts with subject consistency, natural physics simulation, and precise semantic understanding. Supports audio generation. Ready-to-use REST API, best performance, no coldstarts, affordable pricing.

kwaivgi/kling-video-o3-pro/video-edit
video-to-video$0.8400

kwaivgi/kling-video-o3-pro/video-edit

Kling Omni Video O3 Video-Edit enables conversational video editing through natural language commands. Remove objects, change backgrounds, modify styles, adjust weather/lighting, and transform scenes with simple text instructions like 'remove pedestrians' or 'change daytime to dusk'. Ready-to-use REST API, best performance, no coldstarts, affordable pricing.

kwaivgi/kling-video-o3-std/video-edit
video-to-video$0.6300

kwaivgi/kling-video-o3-std/video-edit

Kling Omni Video O3 Video-Edit (Standard) enables natural-language video edits: remove or replace objects, swap backgrounds, restyle scenes, change weather/lighting, and apply localized 3-10s transformations with strong temporal consistency. Built for stable production use with a ready-to-use REST API, no cold starts, and predictable pricing.

kwaivgi/kling-image-o3/edit
image-to-image$0.0280

kwaivgi/kling-image-o3/edit

Kling O3 Edit is an AI image editing model with 4K resolution and multi-image reference support, enabling high-quality transformations with multiple reference inputs. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

kwaivgi/kling-image-o3/text-to-image
text-to-image$0.0280

kwaivgi/kling-image-o3/text-to-image

Kling O3 is Kuaishou's advanced AI image generation model with support for 4K resolution, delivering ultra-high-quality visuals with exceptional detail. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

kwaivgi/kling-elements-advanced
video-tools$0.0100

kwaivgi/kling-elements-advanced

Kling Advanced Elements creates custom AI elements from reference images or videos for consistent character and object appearance across Kling video generations. Supports multi-image elements with frontal and reference images, video character elements, and optional voice binding. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

Kling O3 Models

Kling O3 is a production-ready AI video and image generation model suite for text-to-video, image-to-video, reference-to-video, video editing, text-to-image, and image editing workflows. The collection includes Pro, Standard, and 4K video models, giving creators and developers flexible options for visual quality, generation speed, cost efficiency, and high-resolution production.

Designed for sound-on video creation, Kling O3 supports HD video generation, lip-sync, audio-visual alignment, multilingual prompt workflows, and flexible short-form video outputs. It is suitable for social media clips, ads, e-commerce videos, product showcases, explainers, tutorials, character scenes, and scalable creative production pipelines.

Core Model Capabilities

Text-to-Video Generation:

Create short cinematic videos directly from natural-language prompts with coherent motion, scene structure, camera direction, and synchronized audio-visual output.

Image-to-Video Generation:

Animate still images into dynamic videos while preserving subject identity, composition, visual style, and scene continuity.

Reference-to-Video Generation:

Generate videos guided by reference inputs to preserve character identity, visual direction, framing, motion style, and creative consistency.

AI Video Editing:

Edit and transform existing videos with prompt-guided control for scene refinement, visual adjustment, subject changes, and creative video adaptation.

AI Image Generation:

Create high-quality images from text prompts for key visuals, product concepts, social media assets, advertising creatives, and visual development.

AI Image Editing:

Modify and refine images with natural-language instructions while preserving structure, subject consistency, and important visual details.

Pro and Standard Tiers:

Use Pro models for higher visual fidelity, stronger detail preservation, and polished final outputs, or Standard models for faster, more cost-efficient iteration and high-volume production.

4K Video Workflows:

Use 4K models for high-resolution text-to-video, image-to-video, reference-to-video, and video-edit workflows when premium detail, sharper output, and professional delivery quality are required.

Audio-Visual and Lip-Sync Support:

Generate sound-on videos with aligned voice, motion, and visual timing, making Kling O3 useful for talking content, ads, explainers, tutorials, and short narrative videos.

Kling O3 Models on WaveSpeedAI give creators and developers fast access to a flexible AI image and video generation toolkit with scalable APIs, flexible pricing, and production-ready creative capabilities.

Kling O3 Models API — 價格與效能

透過單一 REST API 執行 Kling O3 Models 系列中的任何模型。按生成計費 — 無訂閱、無最低消費 — 在可用率 99.9% 的基礎架構上提供業界領先的延遲。

為什麼在 WaveSpeedAI 上執行 Kling O3 Models

透明的價格

每個 Kling O3 Models 模型都採按呼叫計費。價格列在每個模型的頁面上 — 不會額外加收平台費。

為低延遲最佳化

大多數 Kling O3 Models 影像模型在 2 秒內完成。影片與 3D 模型比自架方案快數倍。

99.9% 可用率

多區域故障轉移與自動重試可在供應商故障期間 — 仍將您的生產流量保持線上。

常見問題

Kling O3 Models API 多少錢?+

每個模型在其模型頁面上都列有自己的按呼叫價格。我們按每次成功生成計費,沒有訂閱費或最低消費。

Kling O3 Models 模型在 WaveSpeedAI 上有多快?+

本系列中的影像模型通常在 2 秒內完成。影片與 3D 模型取決於長度與解析度,但通常比自架執行快數倍。

不用信用卡可以試用 API 嗎?+

符合資格的新帳戶可能獲得 $1 的推廣額度,用於無需信用卡試用 Kling O3 Models 模型。並非每次註冊都會獲得試用額度,請在生成前查看帳戶餘額。

有速率限制嗎?+

標準帳戶具有充足的並行任務限制。Enterprise 方案提供自訂 RPM、更高並行性和專屬容量 — 詳情請聯繫業務。