Egocentric video co-summarization using transfer learning and refined random walk on a constrained graph

被引:1
|
作者
Sahu, Abhimanyu [1 ]
Chowdhury, Ananda S. [1 ]
机构
[1] Jadavpur Univ, Dept Elect & Telecommun Engn, Kolkata 700032, India
关键词
Egocentric video; Transfer learning; Constrained graph; Random walks; Label refinement;
D O I
10.1016/j.patcog.2022.109128
中图分类号
TP18 [人工智能理论];
学科分类号
081104 ; 0812 ; 0835 ; 1405 ;
摘要
In this paper, we address the problem of egocentric video co-summarization. We show how a shot level accurate summary can be obtained in a time-efficient manner using random walk on a constrained graph in transfer learned feature space with label refinement. While applying transfer learning, we propose a new loss function capturing egocentric characteristics in a pre-trained ResNet on the set of auxiliary egocentric videos. Transfer learning is used to generate i) an improved feature space and ii) a set of labels to be used as seeds for the test egocentric video. A complete weighted graph is created for a test video in the new transfer learned feature space with shots as the vertices. We derive two types of cluster label constraints in form of Must-Link (ML) and Cannot-link (CL) based on the similarity of the shots. ML constraints are used to prune the complete graph which is shown to result in substantial computational advantage, especially, for the long duration videos. We derive expressions for the number of vertices and edges for the ML-constrained graph and show that this graph remains connected. Random walk is applied to obtain labels of the unmarked shots in this new graph. CL constraints are applied to refine the cluster labels. Finally, shots closest to individual cluster centres are used to build the summary. Experiments on the short duration videos as in CoSum and TVSum datasets and long duration videos as in ADL and EPIC-Kitchens datasets clearly demonstrate the advantage of our solution over several state-of-the-art methods.(c) 2022 Elsevier Ltd. All rights reserved.
引用
收藏
页数:10
相关论文
共 7 条
  • [1] Shot Level Egocentric Video Co-summarization
    Sahu, Abhimanyu
    Chowdhury, Ananda S.
    2018 24TH INTERNATIONAL CONFERENCE ON PATTERN RECOGNITION (ICPR), 2018, : 2887 - 2892
  • [2] Scalable Video Summarization using Skeleton Graph and Random Walk
    Panda, Rameswar
    Kuanar, Sanjay K.
    Chowdhury, Ananda S.
    2014 22ND INTERNATIONAL CONFERENCE ON PATTERN RECOGNITION (ICPR), 2014, : 3481 - 3486
  • [3] Egocentric Video Summarization Based on People Interaction Using Deep Learning
    Ghafoor, Humaira A.
    Javed, Ali
    Irtaza, Aun
    Dawood, Hassan
    Dawood, Hussain
    Banjar, Ameen
    MATHEMATICAL PROBLEMS IN ENGINEERING, 2018, 2018
  • [4] Wireless Capsule Endoscopy Video Summarization using Transfer Learning and Random Forests
    Kaur, Parminder
    Kumar, Rakesh
    INTERNATIONAL JOURNAL OF ADVANCED COMPUTER SCIENCE AND APPLICATIONS, 2023, 14 (09) : 353 - 358
  • [5] Scene Classification for Sports Video Summarization Using Transfer Learning
    Rafiq, Muhammad
    Rafiq, Ghazala
    Agyeman, Rockson
    Choi, Gyu Sang
    Jin, Seong-Il
    SENSORS, 2020, 20 (06)
  • [6] On Improving the Properties of Random Walk on Graph using Q-learning
    Matsuo, Ryotaro
    Miyashita, Tomoyuki
    Suzuki, Taisei
    Ohsaki, Hiroyuki
    IEICE COMMUNICATIONS EXPRESS, 2023, 12 (01): : 36 - 41
  • [7] Representation Learning of Enhanced Graphs Using Random Walk Graph Convolutional Network
    Li, Xing
    Wei, Wei
    Zhang, Ruizhi
    Shi, Zhenyu
    Zheng, Zhiming
    Feng, Xiangnan
    ACM TRANSACTIONS ON INTELLIGENT SYSTEMS AND TECHNOLOGY, 2023, 14 (03)