FineAction: A Fine-Grained Video Dataset for Temporal Action Localization
Clicks: 40
ID: 282437
2021
Article Quality & Performance Metrics
Overall Quality
Not rated
Combines reader engagement with the AI quality analysis. This
article has not been analysed, so there is no overall score —
reader engagement is measured and shown alongside.
Reader Engagement
Steady Performance
11.7
/100
40 views
23 readers
AI Quality Assessment
Not analyzed
Readership in this journal
SteadyRanked #619 of 803 articles by views in arXiv
Most read
Least read
Bar heights use a square-root scale. Only the 120 most-read articles are drawn; the journal has 803 in total.
Mint this article as an NFT
Not yet mintedCreate a permanent, verifiable on-chain record of this article on the Scimatic Network. The NFT is held in your Journament account, and you can withdraw it to your own wallet at any time.
5
SUSD
one-off · no wallet required
Abstract
Temporal action localization (TAL) is an important and challenging problem in
video understanding. However, most existing TAL benchmarks are built upon the
coarse granularity of action classes, which exhibits two major limitations in
this task. First, coarse-level actions can make the localization models overfit
in high-level context information, and ignore the atomic action details in the
video. Second, the coarse action classes often lead to the ambiguous
annotations of temporal boundaries, which are inappropriate for temporal action
localization. To tackle these problems, we develop a novel large-scale and
fine-grained video dataset, coined as FineAction, for temporal action
localization. In total, FineAction contains 103K temporal instances of 106
action categories, annotated in 17K untrimmed videos. Compared to the existing
TAL datasets, our FineAction takes distinct characteristics of fine action
classes with rich diversity, dense annotations of multiple instances, and
co-occurring actions of different classes, which introduces new opportunities
and challenges for temporal action localization. To benchmark FineAction, we
systematically investigate the performance of several popular temporal
localization methods on it, and deeply analyze the influence of fine-grained
instances in temporal action localization. As a minor contribution, we present
a simple baseline approach for handling the fine-grained action detection,
which achieves an mAP of 13.17% on our FineAction. We believe that FineAction
can advance research of temporal action localization and beyond.
| Reference Key |
qiao2021fineaction
Use this key to autocite in the manuscript while using
SciMatic Manuscript Manager or Thesis Manager
|
|---|---|
| Authors | Yi Liu; Limin Wang; Yali Wang; Xiao Ma; Yu Qiao |
| Journal | arXiv |
| Year | 2021 |
| DOI |
DOI not found
|
| URL | |
| Keywords |
Citations
No citations found. To add a citation, contact the admin at info@scimatic.org
Comments
No comments yet. Be the first to comment on this article.