ReCAP: Recursive Cross Attention Network for Pseudo-Label Generation in Robotic Surgical Skill Assessment
Clicks: 106
ID: 282272
2024
Article Quality & Performance Metrics
Overall Quality
Not rated
Combines reader engagement with the AI quality analysis. This
article has not been analysed, so there is no overall score —
reader engagement is measured and shown alongside.
Reader Engagement
Emerging Content
30.0
/100
106 views
33 readers
AI Quality Assessment
Not analyzed
Readership in this journal
EmergingRanked #58 of 803 articles by views in arXiv
Most read
Least read
Bar heights use a square-root scale. Only the 120 most-read articles are drawn; the journal has 803 in total.
Mint this article as an NFT
Not yet mintedCreate a permanent, verifiable on-chain record of this article on the Scimatic Network. The NFT is held in your Journament account, and you can withdraw it to your own wallet at any time.
5
SUSD
one-off · no wallet required
Abstract
In surgical skill assessment, the Objective Structured Assessments of
Technical Skills (OSATS) and Global Rating Scale (GRS) are well-established
tools for evaluating surgeons during training. These metrics, along with
performance feedback, help surgeons improve and reach practice standards.
Recent research on the open-source JIGSAWS dataset, which includes both GRS and
OSATS labels, has focused on regressing GRS scores from kinematic data, video,
or their combination. However, we argue that regressing GRS alone is limiting,
as it aggregates OSATS scores and overlooks clinically meaningful variations
during a surgical trial. To address this, we developed a recurrent transformer
model that tracks a surgeon's performance throughout a session by mapping
hidden states to six OSATS, derived from kinematic data, using a clinically
motivated objective function. These OSATS scores are averaged to predict GRS,
allowing us to compare our model's performance against state-of-the-art (SOTA)
methods. We report Spearman's Correlation Coefficients (SCC) demonstrating that
our model outperforms SOTA using kinematic data (SCC 0.83-0.88), and matches
performance with video-based models. Our model also surpasses SOTA in most
tasks for average OSATS predictions (SCC 0.46-0.70) and specific OSATS (SCC
0.56-0.95). The generation of pseudo-labels at the segment level translates
quantitative predictions into qualitative feedback, vital for automated
surgical skill assessment pipelines. A senior surgeon validated our model's
outputs, agreeing with 77% of the weakly-supervised predictions (p=0.006).
| Reference Key |
granados2024recap
Use this key to autocite in the manuscript while using
SciMatic Manuscript Manager or Thesis Manager
|
|---|---|
| Authors | Julien Quarez; Marc Modat; Sebastien Ourselin; Jonathan Shapey; Alejandro Granados |
| Journal | arXiv |
| Year | 2024 |
| DOI |
DOI not found
|
| URL | |
| Keywords |
Citations
No citations found. To add a citation, contact the admin at info@scimatic.org
Comments
No comments yet. Be the first to comment on this article.