Exo2EgoDVC: Dense Video Captioning of Egocentric Procedural Activities Using Web Instructional Videos
Proposes a cross-view dense video captioning framework using adversarial learning to transfer knowledge from web instructional videos to egocentric videos.
Takehiko Ohkawa, Takuma Yagi, Taichi Nishimura et al.