Project Co-Investigator & Visiting Scholar
Reliability & Harm Evaluation of Multimodal Clinical AI
MIT Laboratory for Computational Physiology · MIT Critical Data
I evaluate ensembles of diagnostic-imaging models and quantify how errors accumulate at the image level through a consensus-based harm score. I then link those patterns to ICU care phenotypes using MIMIC-CXR and MIMIC-IV. Calibration, uncertainty, threshold sensitivity, and subgroup performance are part of the main analysis.