π― I'm excited to introduce SambaMOTR, a novel tracker that models long-range dependencies, tracklet interdependencies, and temporal occlusions in complex scenarios like dance performances, sports, and dynamic animal groups. πΊβ½οΈπ¦
πΎ Project Page: https://t.co/ki92ovx7K3
[1/n]
π¨ Next week I'll be at #ICLR2025 in Singapore, presenting our Spotlight paper:
Samba: Synchronized Set-of-Sequences Modeling for Multiple Object Tracking π
π Fri Apr 25, 10:00β12:30
π Hall 3 + Hall 2B, Poster #81
π Project: https://t.co/ki92ovwzUv
Swing by and say hi! π
TL;DR: We propose SambaMOTR, a new E2E multiple object tracking framework that learns to track entirely from data using Samba, our novel set-of-sequences model.
π¨ Next week I'll be at #ICLR2025 in Singapore, presenting our Spotlight paper:
Samba: Synchronized Set-of-Sequences Modeling for Multiple Object Tracking π
π Fri Apr 25, 10:00β12:30
π Hall 3 + Hall 2B, Poster #81
π Project: https://t.co/ki92ovwzUv
Swing by and say hi! π
@wightmanr Awesome! I can launch the experiments on the hybrid medium tomorrow and let you know soon, the gain should be visible already from the first epochs
@wightmanr it seems that the huge performance boost (+6 mAP on COCO) was due to ImageNet-12k pretraining. The anti-aliasing version without ImageNet-12k pretraining performs slightly worse than the conv large without both ImageNet-12k and antialising.
Interestingly, I donβt see this difference for FasterViT, where ImageNet-12k pretraining does not help.
Do you think that it would be possible to have the other MobileNet-v4 models with ImageNet-12k pretraining?
Thanks for the reply! Iβm now training starting from the equivalent conv large aa model without ImageNet-12k pretraining to double check whether itβs really aa that makes the difference. Iβll keep you posted
Iβm definitely looking for the smaller-sized models with anti-aliasing, but it would be interesting to have also the hybrid medium and large if it isnβt asking too much
@lekkima@wightmanr I think it comes from anti-aliased CNNs https://t.co/tZcDxSJfOq
it is not discussed in the MobileNet-V4 paper, it was added to the timm implementation: https://t.co/xiF65CgxLa
Hey @wightmanr, great job integrating MobileNet-V4 in timm! I'm finding that the anti-aliased conv large version works so much better as initialization for object detection compared to the standard conv large version. Any chance that you would also provide the anti-aliased weights for the other model sizes?
This work wouldn't have been possible without my collaborators: @lpiccinelli944, @derek_siyuanli, @YungHsuYang, Bernt Schiele, and Luc Van Gool.
πΎ Project Page: https://t.co/ki92ovx7K3
π Paper: https://t.co/XfG2GypN1g
At the heart of SambaMOTR is Samba, a set of synchronized state-space models. Each trackletβs memory is updated based on observations, then synchronized. This ensures that long-term dependencies and interactions among tracklets are captured to predict the next states accurately.
SambaMOTR excels in complex datasets like DanceTrack (synchronized dance), SportsMOT (interdependent player movements), and BFT (bird flock tracking). The result? Accurate, robust tracking despite occlusions and complex dynamics. π₯
SambaMOTR combines a transformer-based object detector with our novel set-of-sequences Samba model in a tracking-by-propagation framework. Samba jointly processes the trackletsβ memory and their current observations to predict future queries and update track memory.
π― I'm excited to introduce SambaMOTR, a novel tracker that models long-range dependencies, tracklet interdependencies, and temporal occlusions in complex scenarios like dance performances, sports, and dynamic animal groups. πΊβ½οΈπ¦
πΎ Project Page: https://t.co/ki92ovx7K3
[1/n]
Walker wasn't our only tracking paper at #ECCV2024! π
@derek_siyuanli and @YungHsuYang presented SLAck, an open-vocabulary tracker merging semantic, location, and appearance cues to boost association. π―
Check it out: https://t.co/TJttqAKsE6
That's a wrap from Milan! π
ππΆββοΈ Thrilled to have presented our self-supervised multiple-object tracker Walker at #ECCV2024 in my home country! πΆββοΈπ
Big thanks to everyone who stopped by for all the great discussions! π
Learn more about our paper here: https://t.co/KyMkY3MySp