Home /Research /ViSeRet: A simple yet effective approach to moment retrieval via fine-grained video segmentation

OTHER

ViSeRet: A simple yet effective approach to moment retrieval via fine-grained video segmentation

Aiden Seungjoon Lee, Hanseok Oh, Minjoon Seo

Year: 2021
Access: Open access

Abstract

Video-text retrieval has many real-world applications such as media analytics, surveillance, and robotics. This paper presents the 1st place solution to the video retrieval track of the ICCV VALUE Challenge 2021. We present a simple yet effective approach to jointly tackle two video-text retrieval tasks (video retrieval and video corpus moment retrieval) by leveraging the model trained only on the video retrieval task. In addition, we create an ensemble model that achieves the new state-of-the-art performance on all four datasets (TVr, How2r, YouCook2r, and VATEXr) presented in the VALUE Challenge.

Keywords

cs.CVcs.AIcs.CL

ViSeRet: A simple yet effective approach to moment retrieval via fine-grained video segmentation

Abstract

Keywords

Related papers

A dual-loop framework for manufacturability-aware topology optimization of electric vehicle structures via wire arc additive manufacturing

Geometric digital twin: A digital and intelligent model for aero-engine assembly accuracy prediction

Revolutionizing Industries Through AI-Driven Robotics

Design and dynamic performance prediction of a novel large-aperture offset-feed deployable antenna