正在加载视频...

视频加载失败

SAMURAI: Adapting Segment Anything Model for Zero-Shot Visual Tracking with Motion-Aware check out this SAM2 vs SAMURAI comparison! - paper: - code: - license: Apache-2.0

124,419 次观看 • 1 年前 •via X (Twitter)

11 条评论

SkalskiP 的头像
SkalskiP1 年前

- enhance the visual tracking accuracy of SAM 2 by incorporating motion information through motion modeling, to effectively handle the fast-moving and occluded objects - propose a motion-aware memory selection mechanism that reduces error in crowded scenes in contrast to the original fixed-window memory by selectively storing relevant frames decided by a mixture of motion and affinity scores

SkalskiP 的头像
SkalskiP1 年前

state-of-the-art performance on various VOT benchmarks, including GOT-10k, LaSOT-ext, and NeedForSpeed

SkalskiP 的头像
SkalskiP1 年前

can't wait to have some fun with SAMURAI as I did with SAM2

Rainmaker 的头像
Rainmaker2 年前

Can Machine Learning beat the market? Check out this post on my free Substack where I share code and commentary for an XGBoost model and a Random Forest model that both deliver powerful performances.

BensenHsu 的头像
BensenHsu1 年前

The researchers aim to enhance the visual object tracking capabilities of the Segment Anything Model 2 (SAM 2) by addressing its limitations in handling crowded scenes and managing occlusions. The proposed SAMURAI framework demonstrates significant improvements over existing methods on various visual object tracking benchmarks, such as LaSOT, LaSOT ext, and GOT-10k, without the need for additional training or fine-tuning. full paper:

Data 的头像
Data1 年前

Anyone who is against ML/AI tools should be locked in a room and forced to rotoscope this mask by hand. They will be e/acc when they are let out.

X Æ A-12 的头像
X Æ A-121 年前

Amazing work ! 😍

Carlos Alarcón 的头像
Carlos Alarcón1 年前

This is insane, great work !!

Tekholms 的头像
Tekholms1 年前

Absolutely mind blowing! IDK how you keep improving so quickly?? Any experiments with these results on live video feeds?

Brede 的头像
Brede1 年前

This is very cool! Tracking is incredibly hard. Would love to see this applied to multi-object tracking

justboulatbek 的头像
justboulatbek1 年前

I fear this kind of instruments among others are gonna be used in drones in their last mile before chasing the running target

相关视频

🚀 The Segment Anything Model (SAM) has been upgraded to SAM2, featuring an efficient image encoder for segmenting images and videos. But does SAM2 outperform SAM1 in medical image and video segmentation? We're thrilled to present our paper "Segment Anything in Medical Images and Videos: Benchmark and Deployment"! We comprehensively benchmark SAM2 across 11 medical image modalities and videos. 📄 Paper: 💻 Code: **Highlights:** 1. SAM2 doesn’t always outperform SAM1 in 2D medical images, but excels in video segmentation, making it more accurate and efficient for 3D images, such as CT and MR scans. 2. MedSAM still outperforms SAM2 on most 2D modalities, but SAM2 surpasses MedSAM for 3D image segmentation in a slice-by-slice approach. 3. Segmentation performance varies with model size; sometimes the smallest model outperforms larger ones. 4. Fine-tuning SAM2 significantly boosts its performance for medical image segmentation. While SAM2 may struggle with challenging objects that have unclear boundaries or low contrast, it excels in generating good initial segmentation masks for common medical images and videos. However, the official interface doesn’t support medical data formats and has limitations on video length. To address this, we've developed a 3D Slicer Plugin and Gradio API for efficient 3D medical image and video segmentation. We invite you to try them out and provide feedback! 🔧 Deployment: - 3D Slicer Plugin: - Gradio API: (Note: Due to GPU limitations, the online API is available for only 12 hours and may be slow. We highly recommend deploying the Gradio API with your own computing resources: A big shoutout to Jun Ma (JunMa) who recently joined our UHN AI hub (UHN AI Hub) as Machine Learning Lead, and kudos to all co-authors: Sumin Kim, Feifei Li, Mohammed Baharoon (Mohammed Baharoon), Reza Asakereh, and Hongwei Lyu! This is true teamwork! Looking forward to collaborating with the community to advance 3D medical image and video segmentation foundation models! University Health Network U of T Department of Computer Science Department of Laboratory Medicine & Pathobiology Temerty Centre for AI in Medicine (T-CAIREM) Vector Institute #MedTech #AIinHealthcare #DeepLearning #MedicalImaging #SAM2 #MedSAM #AIResearch

Bo Wang

178,579 次观看 • 2 年前