Introducing SAM 2: The next generation of Meta Segment Anything Model for videos and images
Today, we’re announcing the Meta Segment Anything Model 2 (SAM 2), the next generation of the Meta Segment Anything Model, now supporting object segmentation in videos and images. We’re releasing SAM 2 under an Apache 2.0 license, so anyone can use it to build their own experiences. We’re also sharing SA-V, the dataset we used to build SAM 2 under a CC BY 4.0 license and releasing a web-based demo experience where everyone can try a version of our model in action. Object segmentation—identifying the pixels in an image that correspond to an object of interest—is a fundamental task in the field of computer vision. The Meta Segment Anything Model (SAM) released last year introduced a foundation model for this task on images. Our latest model, SAM 2, is the first unified model for real-time, promptable object segmentation in images and videos, enabling a step-change in the video segmentation experience and seamless use across image and video applications. SAM 2 exceeds previous capabilitie
Update: Expanding access to Meta Segment Anything 2.1 on Amazon SageMaker JumpStart Products AI Research Resources About Meta Model API Try Meta AI Open Source Update: Expanding access to Meta Segment Anything 2.1 on Amazon SageMaker JumpStart February 12, 2025 • 15 minute read Updated February 12, 2025: Last July, we released Meta Segment Anything 2, a follow-up to our popular open source segmentation model, offering developers a unified model for real-time promptable object segmentation and tracking in images and videos. We’ve been blown away by the impact SAM 2 has made across the community
Explore this link on the map →related reading
- Fine-tune Segment-Anything model. Tips & caveats | by Rustem Glue | Jun, 2023 | Medium | Mediummedium.com
- The First Fully General Computer Action Model | blogsi.inc
- segment-geospatialsamgeo.gishub.org
- Video models are zero-shot learners and reasonersarxiv.org
- [2506.09985] V-JEPA 2: Self-Supervised Video Models Enable Understanding, Prediction and Planningarxiv.org
- V-JEPA: The next step toward advanced machine intelligenceai.meta.com
- Seoul World Model: Grounding World Simulation Models in a Real-World Metropolisseoul-world-model.github.io
- Replicate - Run AI with an APIreplicate.com
- [2603.14482] V-JEPA 2.1: Unlocking Dense Features in Video Self-Supervised Learningarxiv.org
- The Model That Dreams the Worldmoe-capital.com
- Flexible Diffusion Modeling of Long Videosarxiv.org
- Reintroducing Sievesieve.ai