Suggestion - Integrate MobileSAM into the pipeline for lightweight and faster inference

z-x-yang / Segment-and-Track-Anything

An open-source project dedicated to tracking and segmenting any objects in videos, either automatically or interactively. The primary algorithms utilized include the Segment Anything Model (SAM) for key-frame segmentation and Associating Objects with Transformers (AOT) for efficient tracking and propagation purposes.

GNU Affero General Public License v3.0

2.77k stars 335 forks source link

Suggestion - Integrate MobileSAM into the pipeline for lightweight and faster inference #70

Closed qiaoyu1002 closed 1 year ago

qiaoyu1002 commented 1 year ago

Reference: https://github.com/ChaoningZhang/MobileSAM

Our project performs on par with the original SAM and keeps exactly the same pipeline as the original SAM except for a change on the image encode, therefore, it is easy to Integrate into any project.

MobileSAM is around 60 times smaller and around 50 times faster than original SAM, and it is around 7 times smaller and around 5 times faster than the concurrent FastSAM. The comparison of the whole pipeline is summarzed as follows:

LiNO3Dy commented 1 year ago

Thank you for the suggestion! We will implement this in future versions to make our model more lightweight and faster in processing.