Publications
Conference papers, journal articles, preprints, and manuscripts.
Conference papers · Journal articles · Preprints · Manuscripts
* Equal contribution. + Co-corresponding authors.
Conference papers
2026
MedQA-MM: Shortcuts Behind Medical Visual Reasoning
B. Wang, Y. Zhang, J. Yu, C. Ong, J. Huang, Z. Li, Z. Zhang, A. Cohan, H. Yu, and Z. Yao.
EMNLP 2026 · Preprint
MediSketch: Test-Time Scaling for Multimodal Patient Education
Z. Tang, B. Wang, Z. Zhang, W. Liu, P. Xiang, J. Huang, C. Ong, H. Yu, and Z. Yao.
Findings of EMNLP 2026
LLM-Based Multi-Agent Systems for Clinical Workflows: A Survey of AI Hospitals
Zonghai Yao and Hong Yu.
ACL 2026 · Paper
Exploiting Tree Structure for Credit Assignment in Reinforcement Learning with Large Language Models
Hieu Tran*, Zonghai Yao*, and Hong Yu.
Findings of ACL 2026 · Paper
Medical thinking with multiple images
Z. Yao*, B. Wang*, Y. Zhang, J. Wang, I. Xia, Z. Tang, S. Han, F. Ouyang, Z. Yang, A. Cohan, and H. Yu.
ICLR 2026 · Paper
MedQA-CS: Objective Structured Clinical Examination (OSCE)-Style Benchmark for Evaluating LLM Clinical Skills
Z. Yao, Z. Zhang, C. Tang, X. Bian, Y. Zhao, Z. Yang, J. Wang, H. Zhou, W. Jang, F. Ouyang, and H. Yu.
EACL 2026 (oral) · Paper
Knowing When to Abstain: Medical LLMs Under Clinical Uncertainty
S. Machcha, S. Yerra, S. Gupta, A. Sahoo, S. Sultana, H. Yu, and Z. Yao.
EACL 2026 (oral) · Paper
ChatCLIDS: Simulating Persuasive AI Dialogues to Promote Closed-Loop Insulin Adoption in Type 1 Diabetes Care
Z. Yao*, T. Chafekar*, J. Wang, S. Han, F. Ouyang, J. Qian, L. Li, and H. Yu.
AAAI 2026, AI for Social Impact · Paper
PRIME: Planning and Retrieval-Integrated Memory for Enhanced Reasoning
H. Tran, Z. Yao, N. L. Tran, Z. Yang, F. Ouyang, S. Han, R. Rahimi, and H. Yu.
AAAI 2026 · Paper
2025
DischargeSim: A Simulation Benchmark for Educational Doctor-Patient Communication at Discharge
Zonghai Yao*, Michael Sun*, Won Seok Jang, Sunjae Kwon, Soie Kwon, and Hong Yu.
EMNLP 2025 · Paper
From Scores to Steps: Diagnosing and Improving LLM Performance in Evidence-Based Medical Calculations
B. Wang*, I. Xia*, Y. Zhang, J. Wang, F. Ouyang, S. Han, A. Cohan, H. Yu+, and Z. Yao+.
EMNLP 2025 (oral) · Paper
Chatbot To Help Patients Understand Their Health
W. S. Jang*, H. Tran*, M. Mistry, S. Gandluri, Y. Zhang, S. Sultana, S. Kwon, Z. Yao+, and H. Yu+.
Findings of EMNLP 2025 · Paper
RARE: Retrieval-Augmented Reasoning Enhancement for Large Language Models
H. Tran, Z. Yao, Z. Yang, J. Wang, Y. Zhang, S. Han, F. Ouyang, and H. Yu.
ACL 2025 · Paper
MCQG-SRefine: Multiple Choice Question Generation and Evaluation with Iterative Self-Critique, Correction, and Comparison Feedback
Z. Yao*, A. Parashar*, H. Zhou, W. S. Jang, F. Ouyang, Z. Yang, and H. Yu.
NAACL 2025 (oral) · Paper
2024
SYNFAC-EDIT: Synthetic Imitation Edit Feedback for Factual Alignment in Clinical Summarization
P. Mishra*, Z. Yao*, P. Vashisht, F. Ouyang, B. Wang, V. D. Mody, and H. Yu.
EMNLP 2024 · Paper
README: Bridging Medical Jargon and Lay Understanding for Patient Education through Data-Centric NLP
Z. Yao, N. Kantu, G. Wei, H. Tran, Z. Duan, S. Kwon, Z. Yang, and H. Yu.
Findings of EMNLP 2024 · Paper
Large Language Models are In-context Teachers for Knowledge Reasoning
Jiachen Zhao, Zonghai Yao, Zhichao Yang, and Hong Yu.
Findings of EMNLP 2024 · Paper
NoteChat: A Dataset of Synthetic Patient-Physician Conversations Conditioned on Clinical Notes
J. Wang*, Z. Yao*, Z. Yang, H. Zhou, R. Li, X. Wang, Y. Xu, and H. Yu.
Findings of ACL 2024 · Paper
2023
Improving Summarization with Human Edits
Zonghai Yao, Benjamin J. Schloss, and Sai P. Selvaraj.
EMNLP 2023 · Paper
Revisiting the Architectures like Pointer Networks to Efficiently Improve the Next Word Distribution, Summarization Factuality, and Beyond
H. Chang*, Z. Yao*, A. Gon, H. Yu, and A. McCallum.
Findings of ACL 2023 · Paper
Multi-label Few-shot ICD Coding as Autoregressive Generation with Prompt
Zhichao Yang, Sunjae Kwon, Zonghai Yao, and Hong Yu.
AAAI 2023 · Paper
Context Variance Evaluation of Pretrained Language Models for Prompt-based Biomedical Knowledge Probing
Z. Yao, Y. Cao, Z. Yang, and H. Yu.
AMIA 2023 (oral) · Paper
2022
MedJEx: A Medical Jargon Extraction Model with Wiki’s Hyperlink Span and Contextualized Masked Language Model Score
S. Kwon, Z. Yao, H. Jordan, D. Levy, B. Corner, and H. Yu.
EMNLP 2022 · Paper
Extracting Biomedical Factual Knowledge Using Pretrained Language Model and Electronic Health Record Context
Z. Yao, Y. Cao, Z. Yang, V. Deshpande, and H. Yu.
AMIA 2022 (oral) · Paper
2021
Improving Formality Style Transfer with Context-Aware Rule Injection
Zonghai Yao and Yu Hong.
ACL 2021 · Paper
2020
Zero-shot Entity Linking with Efficient Long Range Sequence Modeling
Zonghai Yao, Liangliang Cao, and Huapu Pan.
Findings of EMNLP 2020 · Paper
Journal articles
2026
Toward Reviewable Medical Evidence Synthesis for Care Delivery
Zonghai Yao and Hong Yu.
npj Digital Medicine, 2026 (Perspective) · Paper
Patient Journey Evaluation for Consumer AI Health Assistants
Zonghai Yao and Hong Yu.
npj Digital Medicine, 2026 (Perspective) · Paper
SynthEHR-eviction: Enhancing Eviction SDoH Detection with LLM-Augmented Synthetic EHR Data
Z. Yao*, Y. Zhao*, A. Mitra, D. Levy, E. Druhl, J. Tsai, and H. Yu.
npj Digital Medicine, 2026 · Paper
MultiViewDx: Evidence-Linked Multi-View Clinical Diagnosis
J. Wang, Z. Yao, Y. Ting, E. Chen, H. Tran, H. Yu, W. Huang, and T. Chen.
Transactions of the Association for Computational Linguistics (TACL), 2026 · Preprint
Posttraumatic Stress Disorder, Health-Related Social Needs, and Cognitive Outcomes in US Veterans
F. Ouyang, T. Pogoda, J. Reisman, R. Li, R. Pradhan, Y. Moradi, J. Mez, W. Liu, Y. Zhang, S. Sultana, J. Qian, Z. Yao, A. Mitra, and H. Yu.
JAMA Network Open, 2026 · Paper
Socioeconomic, Demographic and Geographic Disparities in Accessibility to Food Pantries in the United States
Y. Zhang, M. Lee, J. Gibbons, H. Chen, Y. Wang, Z. Yao, O. Bennett, F. Ouyang, D. Levy, K. Tucker, and H. Yu.
Scientific Reports, 2026 · Paper
Enhancing Large Language Models for Identifying and Prioritizing Important Medical Jargons From Electronic Health Record Notes Using Data Augmentation: Comparative Study
W. Jang*, S. Sultana*, Z. Yao*, H. Tran, Z. Yang, S. Kwon, and H. Yu.
JMIR AI, 2026 · Paper
2025
Unveiling GPT-4V’s Hidden Challenges Behind High Accuracy on USMLE Questions: Observational Study
Z. Yang*, Z. Yao*, M. Tasmin, P. Vashisht, W. Jang, F. Ouyang, B. Wang, D. McManus, D. Berlowitz, and H. Yu.
Journal of Medical Internet Research (JMIR), 2025 · Paper
Association between PTSD and health-related social needs in US Veterans
F. Ouyang, W. Hu, J. Reisman, T. Pogoda, K. Carlson, W. Liu, Y. Moradi, Y. Zhang, S. Han, S. Sultana, W. Jang, Z. Yao, A. Mitra, and H. Yu.
Journal of Affective Disorders, 2025 · Paper
Development of a Surveillance System to Identify Incidence of Evictions Among Patients in Veterans Affairs Medical Centers Across the United States
J. Tsai, S. S. Rajan, Z. Yao, J. Reisman, W. Liu, and H. Yu.
Journal of Community Health, 2025 · Paper
2024
BioInstruct: Instruction Tuning of Large Language Models for Biomedical Natural Language Processing
Hieu Tran, Zhichao Yang, Zonghai Yao, and Hong Yu.
Journal of the American Medical Informatics Association (JAMIA), 2024 · Paper
2023
PaniniQA: Enhancing Patient Education Through Interactive Question Answering
P. Cai*, Z. Yao*, F. Liu, D. Wang, M. Reilly, H. Zhou, L. Li, Y. Cao, A. Kapoor, A. Bajracharya, D. Berlowitz, and H. Yu.
TACL, 2023 · Paper
Automated Identification of Eviction Status from Electronic Health Record Notes
Z. Yao, J. Tsai, W. Liu, D. Levy, E. Druhl, J. Reisman, and H. Yu.
JAMIA, 2023 · Paper
2021
Named Entity Location Prediction Combining Twitter and Web
Y. Liu, W. Shen, Z. Yao, J. Wang, Z. Yang, and X. Yuan.
IEEE Transactions on Knowledge and Data Engineering (TKDE), 2021 · Paper
Preprints
2026
TARSE: Test-Time Adaptation via Retrieval of Skills and Experience for Reasoning Agents
J. Wang, Z. Yao, H. Zeng, Z. Yang, H. Zamani, and H. Yu.
Preprint, 2026 · Paper
Rethinking Patient Education as Multi-turn Multi-modal Interaction
Zonghai Yao*, Zhipeng Tang*, Chengtao Lin, Xiong Luo, Benlu Wang, Juncheng Huang, Chin Siang Ong, and Hong Yu.
Preprint, 2026 · Paper
2025
ChatThero: A Language Agent for Recovery Support
J. Wang*, Z. Yao*, Z. Yang, J. Qian, L. Li, and H. Yu.
Preprint, 2025 · Earlier preprint
MedReadCtrl: Personalizing medical text generation with readability-controlled instruction learning
Hieu Tran*, Zonghai Yao*, Won Seok Jang, Sharmin Sultana, Allen Chang, Yuan Zhang, and Hong Yu.
Preprint, 2025 · Paper
2024
JMLR: Joint Medical LLM and Retrieval Training for Enhancing Reasoning and Professional Question Answering Capability
Junda Wang, Zhichao Yang, Zonghai Yao, and Hong Yu.
Preprint, 2024 · Paper
Manuscripts
Multi-Agent Deep Research for Healthcare Predictions
J. Wang, M. Ghaffari, T. Chafekar, Z. Zhang, Y. Zhang, H. Yu, C. Morato, and Z. Yao.
Manuscript
When Evidence Changes: Temporal Grounding across Agent Autonomy Levels
J. Wang, Z. Yao, H. Zeng, M. Ghaffari, Z. Zhang, C. Morato, H. Zamani, and H. Yu.
Manuscript
Evaluating Evidence Use in Medical Deep Research for Diagnostic Decision Support
X. Lu, B. Wang, J. Wang, C. Ong, J. Huang, Y. Zhang, Z. Tang, Z. Zhang, H. Yu, and Z. Yao.
Manuscript
Measure First, Then Allocate: Calibrated Rollout Budgets for MCP Agents in Reinforcement Learning
H. Tran, Z. Yao, S. Kwon, and H. Yu.
Manuscript
ResidualForge: Toward Autonomous Improvement of Tool-Using Agent Teams
J. Wang, Z. Yao, M. Ghaffari, H. Yu, and C. Morato.
Manuscript
Learning to Consult Specialist for Medical Visual Reasoning
H. Tran, Z. Yao, J. Wang, B. Wang, J. Huang, C. Ong, and H. Yu.
Manuscript