Modality Compensation Network: Cross-Modal Adaptation for Action Recognition.

Song; Sijie;Liu; Jiaying;Li; Yanghao;Guo; Zongming;

doi:10.1109/TIP.2020.2967577

Modality Compensation Network: Cross-Modal Adaptation for Action Recognition.

Clicks: 284

ID: 88075

2020

Article Quality & Performance Metrics

Overall Quality Improving Quality

0.0 /100

Combines engagement data with AI-assessed academic quality

Reader Engagement Steady Performance

30.0 /100

281 views

66 readers

AI Quality Assessment

Not analyzed

Abstract

EN
- Turkish
- Spanish
- Portuguese
- Arabic
- Chinese
- French
- German
- Indonesian
- Russian
- Thai

With the prevalence of RGB-D cameras, multimodal video data have become more available for human action recognition. One main challenge for this task lies in how to effectively leverage their complementary information. In this work, we propose a Modality Compensation Network (MCN) to explore the relationships of different modalities, and boost the representations for human action recognition. We regard RGB/ optical flow videos as source modalities, skeletons as auxiliary modality. Our goal is to extract more discriminative features from source modalities, with the help of auxiliary modality. Built on deep Convolutional Neural Networks (CNN) and Long Short Term Memory (LSTM) networks, our model bridges data from source and auxiliary modalities by a modality adaptation block to achieve adaptive representation learning, that the network learns to compensate for the loss of skeletons at test time and even at training time. We explore multiple adaptation schemes to narrow the distance between source and auxiliary modal distributions from different levels, according to the alignment of source and auxiliary data in training. In addition, skeletons are only required in the training phase. Our model is able to improve the recognition performance with source data when testing. Experimental results reveal that MCN outperforms stateof- the-art approaches on four widely-used action recognition benchmarks.

Reference Key	song2020modalityieee Use this key to autocite in the manuscript while using SciMatic Manuscript Manager or Thesis Manager
Authors	Song, Sijie;Liu, Jiaying;Li, Yanghao;Guo, Zongming;
Journal	ieee transactions on image processing : a publication of the ieee signal processing society
Year	2020
DOI	10.1109/TIP.2020.2967577 Searching for DOI...
URL	https://doi.org/10.1109/TIP.2020.2967577
Keywords	Deep learning artificial intelligence laparoscopic surgery gynaecological surgery

Citations

No citations found. To add a citation, contact the admin at info@scimatic.org

Comments

Login to comment Register

No comments yet. Be the first to comment on this article.