Sign Up to like & get
recommendations!
0
Published in 2025 at "IEEE Transactions on Circuits and Systems for Video Technology"
DOI: 10.1109/tcsvt.2025.3576341
Abstract: Referring multi-object tracking (RMOT) aims to identify specific targets based on sentence descriptions. To enhance multi-modal learning, previous works typically rely on a simple fusion module at early or late stages. However, those methods frequently…
read more here.
Keywords:
referring multi;
aware graph;
correlation aware;
alignment ... See more keywords
Sign Up to like & get
recommendations!
0
Published in 2025 at "IEEE Transactions on Multimedia"
DOI: 10.1109/tmm.2025.3557710
Abstract: Referring Multi-Object Tracking (RMOT) aims to dynamically track an arbitrary number of referred targets in a video sequence according to the language expression. Previous methods mainly focus on cross-modal fusion at the feature level with…
read more here.
Keywords:
referring multi;
visual linguistic;
alignment;
language ... See more keywords