How many examples does an EHRSHOT result really use?
Separate EHRSHOT’s patients, prediction labels, pretraining population and positive/negative few-shot sampling.
Original analysis / Hospital Bench
EHRSHOT tests prediction from longitudinal coded records. MIMIC-CDM tests diagnostic decisions from selected retrospective cases. This independent analysis puts their inputs, denominators and information settings side by side. Explore what a prediction label represents, why a selected disease cohort cannot establish general emergency-care performance, and where published evidence stops. We report aggregate metadata and historical study results, with no patient material, new model runs or implied affiliation with the benchmark authors.
Separate EHRSHOT’s patients, prediction labels, pretraining population and positive/negative few-shot sampling.
Interpret interactive and full-information diagnostic results in the restricted four-condition MIMIC-CDM cohort.
Choose between structured-record prediction and diagnostic simulation by matching the benchmark’s inputs, targets and evidence limits.