Predictions as Surrogates: Revisiting Surrogate Outcomes in the Age of AI
Clicks: 29
ID: 328569
2026
Article Quality & Performance Metrics
Overall Quality
Not rated
Combines reader engagement with the AI quality analysis. This
article has not been analysed, so there is no overall score —
reader engagement is measured and shown alongside.
Reader Engagement
Emerging Content
8.4
/100
29 views
28 readers
AI Quality Assessment
Not analyzed
Readership in this journal
EmergingRanked #33 of 194 articles by views in jurnal biometrika dan kependudukan
Most read
Least read
Bar heights use a square-root scale. Only the 120 most-read articles are drawn; the journal has 194 in total.
Mint this article as an NFT
Not yet mintedCreate a permanent, verifiable on-chain record of this article on the Scimatic Network. The NFT is held in your Journament account, and you can withdraw it to your own wallet at any time.
5
SUSD
one-off · no wallet required
Abstract
Abstract We establish a formal connection between the decades-old surrogate outcome model in biostatistics and economics and the emerging field of prediction-powered inference. The connection treats predictions from pre-trained models, prevalent in the age of AI, as cost-effective surrogates for expensive outcomes that are not fully observed. Building on the surrogate outcomes literature, we develop recalibrated prediction-powered inference, a more efficient approach to statistical inference than existing prediction-powered inference proposals. Our method departs from the existing proposals by using flexible machine learning techniques to learn the optimal “imputed loss” through a step we call recalibration. Importantly, the method always improves upon the estimator that relies solely on the data with available true outcomes, even when the optimal imputed loss is estimated imperfectly, and it achieves the smallest asymptotic variance among prediction-powered inference estimators if the estimate is consistent. Computationally, our optimization objective is convex whenever the loss function that defines the target parameter is convex. We further analyze the benefits of recalibration, both theoretically and numerically, in several common scenarios where machine learning predictions systematically deviate from the outcome of interest. We demonstrate significant gains in effective sample size over exist-ing prediction-powered inference proposals via three applications leveraging state-of-the-art AI models.
| Reference Key |
openalex_W4406549393
Use this key to autocite in the manuscript while using
SciMatic Manuscript Manager or Thesis Manager
|
|---|---|
| Authors | Wenlong Ji, Lihua Lei, Tijana Zrnic |
| Journal | jurnal biometrika dan kependudukan |
| Year | 2026 |
| DOI |
10.1093/biomet/asag053
|
| URL | |
| Keywords | Keywords not found |
Citations
No citations found. To add a citation, contact the admin at info@scimatic.org
Comments
No comments yet. Be the first to comment on this article.