BiomarkersIn silicoPreprint672 individualsCohort studyMulticentre

Default-filled outcome labels in a deployed cognitive-screening programme: an operator-level audit and the construction of twenty-four language-model arms

medRxiv

Abstract

BackgroundRoutine service databases are attractive sources of training labels for clinical prediction models, but the processes that write those labels are rarely audited before the labels are used.

The paper

Jun Ji, Zhigang Sun, Xiaofang Ying,
Show 10 more authorsJinyu Hao, Zhiqiang Fu, Dongmei Shi, Xiaoming Kong, Yijie Xu, Xiaojuan Zhang, Xiaoli Du, Zhiyan Zhang, Xinhong Liu, Ping Lin,
Huali Wang

Qingdao University · Peking University

medRxiv · 2 Sep 2026 · CC BY · Preprint, not peer-reviewed

doi.org/10.64898/2026.08.28.26361585