Segment Anything Model (SAM) 3.1 (2 minute read)

TLDR AI Models

Summary

SAM 3.1 can detect, segment, and track objects in images and video using text prompts on the Meta Model API, with specified inference costs.

SAM 3.1 can detect, segment, and track objects in images and video on the Meta Model API. Users just enter text prompts to identify, segment, and follow any object in images or video. It is served on inference built specifically for its architecture. The model costs $2.50 per 1,000 images or $0.20 per 1,000 frames of video.
Original Article

Similar Articles

SAM 3: Segment Anything with Concepts

Papers with Code Trending

SAM 3 introduces a unified model for promptable concept segmentation and tracking, achieving state-of-the-art performance with a decoupled recognition and localization architecture and a scalable data engine.