REVIEW 2 cited by
MeMOTR: Long-Term Memory-Augmented Transformer for Multi-Object Tracking
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
As a video task, Multiple Object Tracking (MOT) is expected to capture temporal information of targets effectively. Unfortunately, most existing methods only explicitly exploit the object features between adjacent frames, while lacking the capacity to model long-term temporal information. In this paper, we propose MeMOTR, a long-term memory-augmented Transformer for multi-object tracking. Our method is able to make the same object's track embedding more stable and distinguishable by leveraging long-term memory injection with a customized memory-attention layer. This significantly improves the target association ability of our model. Experimental results on DanceTrack show that MeMOTR impressively surpasses the state-of-the-art method by 7.9% and 13.0% on HOTA and AssA metrics, respectively. Furthermore, our model also outperforms other Transformer-based methods on association performance on MOT17 and generalizes well on BDD100K. Code is available at https://github.com/MCG-NJU/MeMOTR.
Forward citations
Cited by 2 Pith papers
-
MapExpert: Online HD Map Construction with Simple and Efficient Sparse Map Element Expert
MapExpert uses shape-specific sparse expert networks and a learnable temporal fusion module to improve online HD map construction by about 1.4-1.8 mAP over MapTracker on nuScenes and Argoverse2.
-
A Framework for Multi-View Multiple Object Tracking using Single-View Multi-Object Trackers on Fish Data
A YOLOv8-ByteTrack pipeline plus stereo triangulation can produce 3D fish tracks for some underwater video pairs, but the claimed multi-view accuracy improvement is not demonstrated.
Discussion (0). Continue with ORCID to comment.