WonJin Yoon

Postdoctoral Research Fellow
Harvard Medical School & Boston Children’s Hospital
Harvard Medical SchoolBoston Children’s Hospital
WonJin Yoon
Publications† joint first authors.
Google Scholar profile
Type
32 publications

Peer-reviewed archival papers

Journal articles and archival conference proceedings reviewed as complete manuscripts.

2026

Monitoring Changes in Clinical Trial Primary Outcomes Using Large Language Models

I Bulovic, S Wunnava, W Yoon, AG Dunn, T Miller, FT Bourgeois
JAMA Network Open 9(9):e2634214 · Research Letter
2026

Substituting Radiology Reports for Chest Radiographs does not preserve prediction behavior in Post-Discharge Mortality Prediction

C Kim, W Yoon, H Lee, J Lee, M Afshar, J Kang, T Miller
Accepted for publication in JAMIA Open
2026

End-to-end extraction of temporal information from psychiatric discharge summaries

S Thomas, G Dinh, W Yoon, B Ren, G Savova, MH Hall, TA Miller
AMIA Joint Summits on Translational Science Proceedings, 458–467
2026

A Dataset of Psychiatric Hospital Notes with Temporal Information Annotations

TA Miller, G Dinh, D Harris, W Yoon, S Thomas, B Ren, MH Hall, G Savova
LREC 2026, 7479–7484
2025

Aspect-Oriented Summarization for Psychiatric Short-Term Readmission Prediction

W Yoon, B Ren, S Thomas, C Kim, G Savova, MH Hall, T Miller
EMNLP 2025, 28037–28054
2025

Using tournaments to calculate AUROC for zero-shot classification with LLMs

W Yoon, I Bulovic, TA Miller
Findings of EMNLP 2025, 23583–23591
2025

LCD benchmark: long clinical document benchmark on mortality prediction for language models

W Yoon, S Chen, Y Gao, Z Zhao, D Dligach, DS Bitterman, M Afshar, T Miller
Journal of the American Medical Informatics Association 32(2), 285–295
2025

Cross-Site Predictions of Readmission After Psychiatric Hospitalization With Mood or Psychotic Disorders: Retrospective Study

B Ren, W Yoon, S Thomas, G Savova, T Miller, MH Hall
JMIR Mental Health 12, e71630
2025

Overview of the 2025 shared task on chemotherapy treatment timeline extraction

J Yao, H Hochheiser, W Yoon, E Goldner, G Savova
ClinicalNLP 2025, 1–10
2024

Overview of the 2024 shared task on chemotherapy treatment timeline extraction

J Yao, H Hochheiser, W Yoon, E Goldner, G Savova
ClinicalNLP 2024, 557–569
2023

Biomedical relation extraction with knowledge base-refined weak supervision

W Yoon, S Yi, R Jackson, H Kim, S Kim, J Kang
Database 2023, baad054
2022

Biomedical NER for the Enterprise with Distillated BERN2 and the Kazu Framework

W Yoon†, R Jackson†, E Ford, V Poroshin, J Kang
EMNLP 2022 Industry Track, 619–626
Research collaboration with AstraZeneca.
2022

Data-centric and model-centric approaches for biomedical question answering

W Yoon, J Yoo, S Seo, M Sung, M Jeong, G Kim, J Kang
CLEF 2022, LNCS 13390, 204–216
Nominated as the Best Paper of CLEF 2021 Labs and accepted to the Best of 2021 Labs track.
2022

Pandemics are catalysts of scientific novelty: Evidence from COVID-19

M Liu, Y Bu, C Chen, J Xu, D Li, Y Leng, RB Freeman, ET Meyer, W Yoon, M Sung, M Jeong, J Lee, J Kang, C Min, M Song, Y Zhai, Y Ding
Journal of the Association for Information Science and Technology 73(8), 1065–1078
Selected as a cover paper of JASIST volume 73, issue 8.
2022

Sequence tagging for biomedical extractive question answering

W Yoon, R Jackson, A Lagerberg, J Kang
Bioinformatics 38(15), 3794–3801
Research collaboration with AstraZeneca.
2022

Full-text chemical identification with improved generalizability and tagging consistency

H Kim, M Sung, W Yoon, S Park, J Kang
Database 2022, baac074
2020

Answering questions on COVID-19 in real-time

J Lee, SS Yi, M Jeong, M Sung, W Yoon, Y Choi, M Ko, J Kang
NLP-COVID19 at EMNLP 2020
2020

BioBERT: a pre-trained biomedical language representation model for biomedical text mining

J Lee†, W Yoon†, S Kim, D Kim, S Kim, CH So, J Kang
Bioinformatics 36(4), 1234–1240
One of the Best Papers in the NLP section of the 2020 IMIA Yearbook.
2020

Pre-trained Language Model for Biomedical Question Answering

W Yoon, J Lee, D Kim, M Jeong, J Kang
ECML PKDD 2019, CCIS 1168, 727–740
2019

A neural named entity recognition and multi-type normalization tool for biomedical text mining

D Kim, J Lee, CH So, H Jeon, M Jeong, Y Choi, W Yoon, M Sung, J Kang
IEEE Access 7, 73729–73740
2019

Collabonet: collaboration of deep neural networks for biomedical named entity recognition

W Yoon†, CH So†, J Lee, J Kang
BMC Bioinformatics 20(Suppl 10):249

Peer-reviewed working notes and shared-task papers

2023

Exploring Approaches to Answer Biomedical Questions: From Pre-processing to GPT-4

H Kim, H Hwang, C Lee, M Seo, W Yoon, J Kang
CLEF 2023 Working Notes, BioASQ Task 11b, 132–144
2022

KU_ED at SocialDisNER: Extracting Disease Mentions in Tweets Written in Spanish

A Lain, W Yoon, H Kim, J Kang, I Simpson
SMM4H 2022 Workshop & Shared Task, 78–80
2020

Transferability of Natural Language Inference to Biomedical Question Answering

M Jeong†, M Sung†, G Kim, D Kim, W Yoon, J Yoo, J Kang
CLEF 2020 Working Notes, BioASQ Task 8b

Under review, preprints, and ongoing work

2026

CLSGen: A Dual-Head Fine-Tuning Framework for Joint Probabilistic Classification and Verbalized Explanation

W Yoon†, K Zhu†, I Bulovic, A Sehy, Y Gao, D Dligach, M Afshar, TA Miller
arXiv:2604.11801
2026

Comparing prognostic performance and reasoning between large language models and physicians

M Gjertsen†, W Yoon†, M Afshar, B Temte, B Leding, S Halliday, K Bradley, J Kim, J Mitchell, AK Sanders, E Croxford, J Caskey, M Churpek, A Mayampurath, Y Gao, T Miller, JM Kruser
medRxiv 2026.04.17.26350898
2026

Automated assessment of outcome reporting bias and safety signals across clinical trials using a domain-adapted language model

X Zhang, Q Hu, A Pan, I Bulovic, W Yoon, S Wunnava, T Miller, J Kim, F Bourgeois, A Dunn
Under review
2025

Medical hallucinations in foundation models and their impact on healthcare

Y Kim, H Jeong, S Chen, S Li, C Park, M Lu, K Alhamoud, J Mun, C Grau, M Jung, R Gameiro, L Fan, E Park, T Lin, J Yoon, W Yoon, M Sap, Y Tsvetkov, P Liang, X Xu, X Liu, C Park, H Lee, H Park, D McDuff, S Tulebaev, C Breazeal
Under review · arXiv:2503.05777

Selected conference abstracts

2026

Multimodal Fusion of Structured EHR Data and Clinical Notes for 30-Day Mortality Prediction

MT Saban, S Tootooni, W Yoon, T Miller, D Dligach
Accepted poster, AMIA 2026 Annual Symposium, November 7–11
2026

Comparing prognostic performance and reasoning between physicians and large language models

M Gjertsen†, W Yoon†, M Afshar, E Croxford, JR Caskey, Y Gao, TA Miller, JM Kruser
Podium presentation, AAHPM Annual Assembly, March 7 · JPSM 71(6), e903–e904
2025

Using Language Models for Psychiatric Short-Term Readmission Prediction

W Yoon
Oral presentation, Technology in Psychiatry Summit, December 12
2025

Detection and Monitoring of Potential Outcome Reporting Bias Using Large Language Models: Application to FDA-Regulated Drug Trials

I Bulovic, S Wunnava, W Yoon, A Dunn, T Miller, F Bourgeois
Oral presentation, Peer Review Congress, September 3–5