首页 /研究 /ViSeRet: A simple yet effective approach to moment retrieval via fine-grained video segmentation

OTHER

ViSeRet: A simple yet effective approach to moment retrieval via fine-grained video segmentation

Aiden Seungjoon Lee, Hanseok Oh, Minjoon Seo

发表年份: 2021
访问权限: 开放获取

摘要

Video-text retrieval has many real-world applications such as media analytics, surveillance, and robotics. This paper presents the 1st place solution to the video retrieval track of the ICCV VALUE Challenge 2021. We present a simple yet effective approach to jointly tackle two video-text retrieval tasks (video retrieval and video corpus moment retrieval) by leveraging the model trained only on the video retrieval task. In addition, we create an ensemble model that achieves the new state-of-the-art performance on all four datasets (TVr, How2r, YouCook2r, and VATEXr) presented in the VALUE Challenge.

关键词

cs.CVcs.AIcs.CL

ViSeRet: A simple yet effective approach to moment retrieval via fine-grained video segmentation

摘要

关键词

相关论文

一种面向线弧增材制造的电动汽车结构可制造性拓扑优化的双环框架

几何数字孪生：一种用于航空发动机装配精度预测的数字智能模型

通过人工智能驱动的机器人技术革新产业

新型大口径偏置馈电可展开天线设计与动态性能预测