Seedream 5.0 Pro đã ra mắt | Thử trong Trình tạo ảnh →
Object Detection and Segmentation

Object Detection and Segmentation

Detect, identify, and segment objects in images and videos with AI models on WaveSpeed

Tuyển chọn của chúng tôi

wavespeed-ai/moondream3-preview/point
image-to-text

wavespeed-ai/moondream3-preview/point

Moondream3 Point finds objects in images and returns precise coordinate points for computer vision tasks, enabling accurate point localization. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

Tất cả mô hình

10 mô hình
wavespeed-ai/moondream3-preview/point
image-to-text

wavespeed-ai/moondream3-preview/point

Moondream3 Point finds objects in images and returns precise coordinate points for computer vision tasks, enabling accurate point localization. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

wavespeed-ai/moondream3-preview/detect
image-to-text

wavespeed-ai/moondream3-preview/detect

Moondream3 Detect: Precise object bounding boxes in images for accurate computer vision localization. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

wavespeed-ai/sam-3d-body
image-to-3d

wavespeed-ai/sam-3d-body

Advanced SAM 3D body generation model for creating detailed 3D human body models from images with optional mask-based segmentation. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

wavespeed-ai/sam-3d-objects
image-to-3d

wavespeed-ai/sam-3d-objects

Advanced SAM 3D objects generation model for creating detailed 3D object models from images with text prompts and optional mask-based segmentation. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

wavespeed-ai/sam3-video
video-to-video

wavespeed-ai/sam3-video

SAM3 Video is a unified foundation model for prompt-based video segmentation. Provide text, point, box, or mask prompts and the model segments and tracks targets across frames with strong temporal consistency. Supports concept-level (“segment anything with concepts”) and multi-object masks for editing, analytics, and VFX. Ready-to-use REST inference API with fast response, no cold starts, and affordable pricing.

wavespeed-ai/sam3-image
image-to-image

wavespeed-ai/sam3-image

SAM 3 is a unified foundation model for promptable image segmentation using text, points, or boxes to detect and segment objects. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

wavespeed-ai/sam3-video-rle
video-to-text

wavespeed-ai/sam3-video-rle

SAM 3 Video RLE is a unified foundation model for prompt-based segmentation in video. Track and segment objects across frames using text, points, or boxes, returning RLE encoded masks for efficient processing. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

wavespeed-ai/sam3-image-rle
image-to-text

wavespeed-ai/sam3-image-rle

SAM 3 RLE is a unified foundation model for promptable image segmentation using text, points, or boxes to detect and segment objects. Returns RLE (Run-Length Encoding) encoded masks for efficient storage and processing. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

bria/embed-product
image-to-image

bria/embed-product

Bria Embed Product seamlessly integrates product images into scene backgrounds with natural lighting and perspective matching. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

wavespeed-ai/void-video-inpainting/mask
video-to-video

wavespeed-ai/void-video-inpainting/mask

VOID Video Inpainting removes objects from videos using mask-guided inpainting. Supports quad-mask or auto-generated SAM-3 masks, optional Pass 2 refinement for temporal consistency, adjustable denoising steps, guidance scale, and temporal window size. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

API Object Detection and Segmentation — giá cả & hiệu năng

Chạy bất kỳ mô hình nào trong bộ sưu tập Object Detection and Segmentation qua một REST API duy nhất. Trả tiền theo lần tạo — không đăng ký gói, không mức tối thiểu — với độ trễ hàng đầu ngành trên hạ tầng có thời gian hoạt động 99,9%.

Vì sao nên chạy Object Detection and Segmentation trên WaveSpeedAI

Giá cả minh bạch

Giá theo mỗi lần gọi cho mọi mô hình Object Detection and Segmentation. Giá được liệt kê trên từng trang mô hình — không có phí nền tảng cộng thêm.

Tối ưu cho độ trễ thấp

Hầu hết các mô hình hình ảnh Object Detection and Segmentation hoàn thành trong chưa đầy 2 giây. Các mô hình video và 3D chạy nhanh hơn vài lần so với các giải pháp tự lưu trữ.

Thời gian hoạt động 99,9%

Chuyển đổi dự phòng đa khu vực và tự động thử lại giữ cho lưu lượng sản xuất của bạn luôn trực tuyến — kể cả khi nhà cung cấp gặp sự cố.

Câu hỏi thường gặp

API Object Detection and Segmentation có giá bao nhiêu?+

Mỗi mô hình có giá riêng theo mỗi lần gọi được liệt kê trên trang mô hình. Chúng tôi tính phí theo mỗi lần tạo thành công, không có phí đăng ký hay mức tối thiểu.

Các mô hình Object Detection and Segmentation chạy nhanh thế nào trên WaveSpeedAI?+

Các mô hình hình ảnh trong bộ sưu tập này thường hoàn thành trong chưa đầy 2 giây. Các mô hình video và 3D phụ thuộc vào thời lượng và độ phân giải nhưng thường nhanh hơn vài lần so với chạy tự lưu trữ.

Tôi có thể dùng thử API mà không cần thẻ tín dụng không?+

Có — mỗi tài khoản nhận $1 tín dụng miễn phí khi đăng ký, đủ để thử hầu hết các mô hình Object Detection and Segmentation mà không cần thẻ tín dụng.

Có giới hạn tốc độ không?+

Tài khoản tiêu chuẩn có giới hạn tác vụ đồng thời rộng rãi. Các gói doanh nghiệp cung cấp RPM tùy chỉnh, mức đồng thời cao hơn và dung lượng riêng — liên hệ bộ phận kinh doanh để biết chi tiết.