Nano Banana 2.1 EN LIGNE — Le dernier de Google | Essayer →
Object Detection and Segmentation

Object Detection and Segmentation

Detect, identify, and segment objects in images and videos with AI models on WaveSpeed

Notre sélection

bytedance/seedream-v5.0-pro/layer-decomposition
image-to-image

bytedance/seedream-v5.0-pro/layer-decomposition

Seedream V5.0 Pro Layer Decomposition separates a single image into a base image and transparent layers, enabling flexible compositing, image editing, asset extraction, and layered design workflows. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

Tous les modèles

18 modèles
bytedance/seedream-v5.0-pro/layer-decomposition
image-to-image

bytedance/seedream-v5.0-pro/layer-decomposition

Seedream V5.0 Pro Layer Decomposition separates a single image into a base image and transparent layers, enabling flexible compositing, image editing, asset extraction, and layered design workflows. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

bytedance/seedream-v5.0-flash/layer-decomposition
image-to-image

bytedance/seedream-v5.0-flash/layer-decomposition

Seedream V5.0 Flash Layer Decomposition separates a single image into a base image and transparent layers faster and at lower cost than Seedream V5.0 Pro, enabling flexible compositing, image editing, asset extraction, and layered design workflows at one price for 1K / 1.5K / 2K. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

wavespeed-ai/sam3-video
video-tools

wavespeed-ai/sam3-video

SAM3 Video is a unified foundation model for prompt-based video segmentation. Provide text, point, box, or mask prompts and the model segments and tracks targets across frames with strong temporal consistency. Supports concept-level (“segment anything with concepts”) and multi-object masks for editing, analytics, and VFX. Ready-to-use REST inference API with fast response, no cold starts, and affordable pricing.

wavespeed-ai/depth-anything/video
video-tools

wavespeed-ai/depth-anything/video

Depth Anything Video estimates temporally consistent depth maps from video input, with stable depth across the whole clip and no flicker. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

wavespeed-ai/moondream3-preview/point
image-to-text

wavespeed-ai/moondream3-preview/point

Moondream3 Point finds objects in images and returns precise coordinate points for computer vision tasks, enabling accurate point localization. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

wavespeed-ai/qwen-image/layered
image-to-image

wavespeed-ai/qwen-image/layered

Qwen-Image Layered is a unified image-layer decomposition model for prompt-guided compositing. Provide points, boxes, or rough masks to isolate subjects and regions, and the model splits a single image into multiple RGBA layers with clean alpha, soft edges, and correct occlusion order. Ready-to-use REST inference API with fast response, no cold starts, and affordable pricing.

wavespeed-ai/sam3-image
image-tools

wavespeed-ai/sam3-image

SAM 3 is a unified foundation model for promptable image segmentation using text, points, or boxes to detect and segment objects. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

wavespeed-ai/sam3-video-rle
video-tools

wavespeed-ai/sam3-video-rle

SAM 3 Video RLE is a unified foundation model for prompt-based segmentation in video. Track and segment objects across frames using text, points, or boxes, returning RLE encoded masks for efficient processing. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

wavespeed-ai/sam3-image-rle
image-tools

wavespeed-ai/sam3-image-rle

SAM 3 RLE is a unified foundation model for promptable image segmentation using text, points, or boxes to detect and segment objects. Returns RLE (Run-Length Encoding) encoded masks for efficient storage and processing. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

wavespeed-ai/sam-3d-objects
image-to-3d

wavespeed-ai/sam-3d-objects

Advanced SAM 3D objects generation model for creating detailed 3D object models from images with text prompts and optional mask-based segmentation. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

bria/embed-product
image-to-image

bria/embed-product

Bria Embed Product seamlessly integrates product images into scene backgrounds with natural lighting and perspective matching. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

wavespeed-ai/moondream3-preview/detect
image-to-text

wavespeed-ai/moondream3-preview/detect

Moondream3 Detect: Precise object bounding boxes in images for accurate computer vision localization. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

wavespeed-ai/sam-3d-body
image-to-3d

wavespeed-ai/sam-3d-body

Advanced SAM 3D body generation model for creating detailed 3D human body models from images with optional mask-based segmentation. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

bria/ad-delayer
image-tools

bria/ad-delayer

Bria Ad Delayer splits flat advertisement images into editable image, text, and vector layers, returning a complete structured layer document as JSON for ad editing, localization, redesign, marketing creatives, and production workflows. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

wavespeed-ai/depth-anything-v3/video
video-tools

wavespeed-ai/depth-anything-v3/video

Depth Anything V3 Video turns any video into a temporally consistent depth map video with no flicker, ideal for replicating camera moves and motion with depth-controlled video generation. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

wavespeed-ai/depth-anything-v3/image
image-tools

wavespeed-ai/depth-anything-v3/image

Depth Anything V3 Image estimates a sharp, detailed depth map from a single image, ready for depth-conditioned generation, relighting, 3D and compositing workflows. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

wavespeed-ai/depth-anything/image
image-tools

wavespeed-ai/depth-anything/image

Depth Anything Image turns a single image into a detailed grayscale depth map for depth-conditioned generation, relighting, 3D and compositing workflows. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

wavespeed-ai/void-video-inpainting/mask
video-to-video

wavespeed-ai/void-video-inpainting/mask

VOID Video Inpainting removes objects from videos using mask-guided inpainting. Supports quad-mask or auto-generated SAM-3 masks, optional Pass 2 refinement for temporal consistency, adjustable denoising steps, guidance scale, and temporal window size. Ready-to-use REST inference API, best performance, no coldstarts, affordable pricing.

API Object Detection and Segmentation — tarifs et performances

Exécutez n'importe quel modèle de la collection Object Detection and Segmentation via une seule API REST. Paiement à la génération — sans abonnement ni minimum — avec une latence à l'état de l'art sur une infrastructure à 99,9 % de disponibilité.

Pourquoi exécuter Object Detection and Segmentation sur WaveSpeedAI

Tarification transparente

Tarification à l'appel pour chaque modèle Object Detection and Segmentation. Le prix est indiqué sur la page de chaque modèle — pas de frais de plateforme en plus.

Optimisé pour une faible latence

La plupart des modèles d'image Object Detection and Segmentation s'exécutent en moins de 2 secondes. Les modèles vidéo et 3D sont plusieurs fois plus rapides que les alternatives auto-hébergées.

99,9 % de disponibilité

Bascule multi-régions et nouvelles tentatives automatiques maintiennent votre trafic de production en ligne — même en cas de panne fournisseur.

Questions fréquentes

Combien coûte l'API Object Detection and Segmentation ?+

Chaque modèle a son propre prix par appel indiqué sur sa page. Nous facturons à chaque génération réussie, sans abonnement ni minimum.

À quelle vitesse les modèles Object Detection and Segmentation fonctionnent-ils sur WaveSpeedAI ?+

Les modèles d'image de cette collection se terminent généralement en moins de 2 secondes. Les modèles vidéo et 3D dépendent de la durée et de la résolution mais sont en général plusieurs fois plus rapides que les exécutions auto-hébergées.

Puis-je essayer l'API sans carte de crédit ?+

Les nouveaux comptes éligibles peuvent recevoir 1 $ de crédits promotionnels pour essayer les modèles Object Detection and Segmentation sans carte de crédit. Ces crédits ne sont pas garantis à chaque inscription ; vérifiez votre solde avant de générer.

Y a-t-il des limites de taux ?+

Les comptes standard ont des limites de jobs concurrents généreuses. Les plans Enterprise proposent un RPM personnalisé, une concurrence plus élevée et de la capacité dédiée — contactez le service commercial pour les détails.