Model Details [\[π Tech Report\]](https://arxiv.org/abs/2501.07256) [\[π Github\]](https://github.com/facebookresearch/EdgeTAM) [\[π€ Demo\]](https://huggingface.co/spaces/yonigozlan/EdgeTAM hf) EdgeTAM is an on device executable variant of the SAM 2 for promptable segmentation and tracking in videos. It runs 22Γ faster than SAM 2 and achieves 16 FPS on iPhone 15 Pro Max without quantization. How to use with Transformers Automatic Mask Generation with Pipeline EdgeTAM can be used for automatic mask generation to segment all objects in an image using the mask generation pipeline: Basic Image Segmentation Single Point Click You can segment objects by providing a single point click on the object you want to segment: Multiple Points for Refinement You can provide multiple points to refine the segmentation: Bounding Box Input EdgeTAM also supports bounding box inputs for segmentation: Multiple Objects Segmentation You can segment multiple objects simultaneously: Batch Inference Batched Images Process multiple images simultaneously for improved efficiency: Batched Objects per Image Segment multiple objects within each image using batch inference: Batched Images with Batched Objects andβ¦
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy