# Meta SAM 3 Released: Video Tracking, 3D & Open Weights!

Source: https://www.youtube.com/watch?v=r4abIemObmQ
Recap page: https://rapidrecap.app/video/r4abIemObmQ
Generated: 2025-11-23T14:32:45.33+00:00

---
## Quick Overview

Meta released Segment Anything Model 3 (SAM 3), a unified model for detection, segmentation, and tracking in images and videos using text, exemplar, and visual prompts, alongside the Segment Anything Playground for interactive experimentation, which includes features like video cutouts, image cutouts, 3D scene creation, and 3D body reconstruction.

**Key Points:**
- Meta announced Segment Anything Model 3 (SAM 3), a unified model for detection, segmentation, and tracking of objects in both images and videos.
- SAM 3 supports text, exemplar, and visual prompts, building on the capabilities of Segment Anything 2.
- The release includes the Segment Anything Playground, an interactive platform where users can experiment with SAM 3 features like creating video cutouts, image cutouts, 3D scenes, and 3D bodies.
- SAM 3 capabilities will soon enable new effects in Instagram's video creation app, Edits, and will be available on Vibes on the Meta AI app and meta.ai.
- SAM 3D, a suite for 3D objects and human reconstruction from a single image, is also being shared, setting a new standard for grounded 3D reconstruction.
- The model demonstrates state-of-the-art performance across various benchmarks, including Concept Segmentation, Visual Segmentation, Counting, and Reasoning Segmentation.
- The video showcases SAM 3's ability to segment and track objects (like people, elephants, zebras, and jets) in complex scenes and video sequences using text prompts.

![Screenshot at 00:00: The initial demonstration shows SAM 3 segmenting two elephants in a savanna scene using text prompting, highlighting its ability to isolate specific objects based on simple descriptions.](https://ss.rapidrecap.app/screens/r4abIemObmQ/00-00-00.png)

**Context:** This video announces the release of Meta's next-generation visual foundation model, Segment Anything Model 3 (SAM 3), which extends the SAM series to handle object tracking in videos and introduces advanced prompting capabilities like text and visual prompts. The announcement also covers the introduction of the Segment Anything Playground, a web platform allowing users to test these new capabilities immediately across various tasks such as video cutouts and 3D reconstruction.

## Detailed Analysis

Meta officially announced Segment Anything Model 3 (SAM 3), an evolution of the SAM series that unifies detection, segmentation, and tracking across images and videos using text, exemplar, and visual prompts. This new iteration is noted to be 50% faster for positive prompts in challenging, fine-grained domains compared to previous models. Alongside SAM 3, Meta released the Segment Anything Playground, an interactive environment where users can test features like creating video cutouts (03:51), image cutouts (05:19), creating 3D scenes (06:21), and generating 3D bodies (06:56) from 2D inputs. The announcement highlights practical applications, such as integrating SAM 3 into Instagram's Edits app for new video effects and powering Facebook Marketplace's 'View in Room' feature via SAM 3D. The video provides extensive demonstrations, showing SAM 3 successfully tracking objects like people in a park (00:01), fish underwater (00:09), penguins (00:14), dogs (00:19), and jets in formation (09:40), all driven by text prompts like "people," "zebra," or "jet." Furthermore, the video details the SAM 3 data engine workflow, which relies on a hybrid human and AI verification loop for quality control, and showcases SAM 3D's ability to generate 3D models (06:44) from single images. Benchmarks confirm SAM 3 achieves state-of-the-art results across numerous segmentation tasks (02:58).

### SAM 3 Introduction

- Unified model for detection, segmentation, and tracking in images/video
- Supports text, exemplar, and visual prompts
- 50% faster for positive prompts than previous iterations.

### Segment Anything Playground Features

- Offers modules for Create video cutouts (SAM 3)
- Create image cutouts (SAM 3)
- Create 3D scenes (SAM 3D)
- Create 3D bodies (SAM 3D).

### Video Segmentation Demonstration (Penguins/Jets)

- SAM 3 uses a detector and tracker with a memory bank to maintain object identity across frames (01:41); successfully tracks multiple jets despite occlusion (10:27).

### SAM 3D Reconstruction

- Takes an image input, processes it through an Image Encoder, Geometry module, and Texture & Refinement module to output Mesh or Gaussians (02:40).

### Playground Walkthrough (Video Cutouts)

- User searches for 'soccer ball' (04:11), the model tracks it across the video (04:38), allowing users to review and refine objects (04:47) before applying effects like 'Blur' or 'Cartoon' (06:14).

### Application Integration

- SAM 3 enables new effects in Instagram's Edits app and features on Vibes and meta.ai (01:17), and SAM 3D powers Facebook Marketplace's 'View in Room' (02:50).

![Screenshot at 00:00: Initial demonstration of SAM 3 using text prompting to segment two elephants based on the prompt 'elephant'.](https://ss.rapidrecap.app/screens/r4abIemObmQ/00-00-00.png)
![Screenshot at 02:58: Benchmark table showing SAM 3 achieving state-of-the-art results across several segmentation metrics, outperforming competitors like OwlV2 and Gemini 2.5 Pro.](https://ss.rapidrecap.app/screens/r4abIemObmQ/00-02-58.png)
![Screenshot at 03:22: Demonstration of the Segment Anything Playground workflow, showing selection of media, describing the object \('dog'\), applying effects \('Blur'\), and saving the creation.](https://ss.rapidrecap.app/screens/r4abIemObmQ/00-03-22.png)
![Screenshot at 09:40: Visualization of SAM 3 tracking multiple jets in formation using the text prompt 'jet', showing each tracked object with a unique, colored label.](https://ss.rapidrecap.app/screens/r4abIemObmQ/00-09-40.png)
![Screenshot at 10:37: Demonstration within the Playground showing the 'Blur faces' template applied to a video of children on a ride, highlighting automated face detection and masking.](https://ss.rapidrecap.app/screens/r4abIemObmQ/00-10-37.png)
