cs.CV 2405.07364

BoQ: A Place is Worth a Bag of Learnable Queries

BoQ uses learnable global queries with cross-attention, outperforming SOTA in visual place recognition with high speed and efficiency.

Amar Ali-Bey, Brahim Chaib-draa, Philippe Giguère

2024-05-13 59
cs.CV 2405.06574

Deep video representation learning: a survey

This survey organizes video representation learning around RGB, optical flow, CNNs, RNNs, and GCNs, without reporting unified benchmark scores.

Elham Ravanbakhsh, Yongqing Liang, J. Ramanujam et al.

2024-05-11 22