
Detect objects, faces, poses, text, depth, and more with powerful AI detection and analysis models on WaveSpeed

Scalable Text Content Moderator for filtering and classifying user-generated text, ideal for safety and compliance workflows. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

Scalable Text Content Moderator for filtering and classifying user-generated text, ideal for safety and compliance workflows. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

Image Content Moderator provides automated image moderation to detect and flag policy-violating or inappropriate images for automation. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

Moondream3 Point finds objects in images and returns precise coordinate points for computer vision tasks, enabling accurate point localization. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

Molmo2-4B Image Captioner: Generate detailed, accurate captions for images with customizable detail levels (low, medium, high). Open-source vision-language model with object grounding capabilities. Ready-to-use REST API, no cold starts, affordable pricing.

Molmo2-4B Video Captioner: Generate detailed, accurate captions for videos with customizable detail levels (low, medium, high). Open-source vision-language model with temporal understanding capabilities. Ready-to-use REST API, no cold starts, duration-based pricing.

Molmo2-4B Video QA: Answer questions about video content with temporal understanding. Open-source vision-language model. Ready-to-use REST API, no cold starts, duration-based pricing.

Molmo2-4B Video Understanding: Analyze videos with specialized tasks (general, summary, analysis, counting, scene description). Open-source vision-language model with temporal understanding. Ready-to-use REST API, no cold starts, duration-based pricing.

Molmo2-4B Image QA: Answer questions about images with support for multi-image comparison (1-2 images). Open-source vision-language model. Ready-to-use REST API, no cold starts, affordable pricing.

Molmo2-4B Text Content Moderator: Analyze text content for safety, appropriateness, and policy compliance. Detects hate speech, violence, sexual content, and other harmful categories. Open-source vision-language model. Ready-to-use REST API, no cold starts, affordable pricing.

Molmo2-4B Image Content Moderator: Analyze image content for safety, appropriateness, and policy compliance. Detects violence, nudity, gore, and other harmful visual content. Open-source vision-language model. Ready-to-use REST API, no cold starts, affordable pricing.

Molmo2-4B Video Content Moderator analyzes video content for safety, appropriateness, and policy compliance. Detects violence, nudity, gore, and other harmful visual content in videos using an open-source vision-language model. Ready-to-use REST inference API, best performance, no cold starts, affordable pricing.
Chạy bất kỳ mô hình nào trong bộ sưu tập Content Detection Models qua một REST API duy nhất. Trả tiền theo lần tạo — không đăng ký gói, không mức tối thiểu — với độ trễ hàng đầu ngành trên hạ tầng có thời gian hoạt động 99,9%.
Giá theo mỗi lần gọi cho mọi mô hình Content Detection Models. Giá được liệt kê trên từng trang mô hình — không có phí nền tảng cộng thêm.
Hầu hết các mô hình hình ảnh Content Detection Models hoàn thành trong chưa đầy 2 giây. Các mô hình video và 3D chạy nhanh hơn vài lần so với các giải pháp tự lưu trữ.
Chuyển đổi dự phòng đa khu vực và tự động thử lại giữ cho lưu lượng sản xuất của bạn luôn trực tuyến — kể cả khi nhà cung cấp gặp sự cố.
Mỗi mô hình có giá riêng theo mỗi lần gọi được liệt kê trên trang mô hình. Chúng tôi tính phí theo mỗi lần tạo thành công, không có phí đăng ký hay mức tối thiểu.
Các mô hình hình ảnh trong bộ sưu tập này thường hoàn thành trong chưa đầy 2 giây. Các mô hình video và 3D phụ thuộc vào thời lượng và độ phân giải nhưng thường nhanh hơn vài lần so với chạy tự lưu trữ.
Có — mỗi tài khoản nhận $1 tín dụng miễn phí khi đăng ký, đủ để thử hầu hết các mô hình Content Detection Models mà không cần thẻ tín dụng.
Tài khoản tiêu chuẩn có giới hạn tác vụ đồng thời rộng rãi. Các gói doanh nghiệp cung cấp RPM tùy chỉnh, mức đồng thời cao hơn và dung lượng riêng — liên hệ bộ phận kinh doanh để biết chi tiết.
Duyệt qua danh mục đầy đủ của chúng tôi về các mô hình AI tiên tiến — hình ảnh, video, 3D, âm thanh, LLM, v.v.
wavespeed.ai/models →Tích hợp AI vào ứng dụng của riêng bạn. RESTful API với thư viện máy khách — không cần khởi động nguội, trả tiền cho mỗi lần sử dụng.
wavespeed.ai/docs →