YW
Yonghui Wu
cs.CLcs.LGeess.AScs.SDNatural Language ProcessingElectronic Health Recordscs.AInatural language processingmachine learningDeep Learning
On Valency
published · living versionsW_39gz77fp·v1 · currentpublished
Exploring the Limits of Language Modeling
with Rafal Jozefowicz, Oriol Vinyals, Mike Schuster, Noam Shazeer
1 version
Preprints & journals
233 papers in the corpus · 2003–2026A large language model for electronic health records.36572766 · Yang, Xi, Chen, Aokun, PourNejatian, Nima et al.2026 · 880 citationsNPJ digital medicine. 2022;5(1):194
Less can be better: decomposing clinical data modalities in large language model-based healthcare applications.42700364 · Peng, Cheng, Lyu, Mengxian, Chen, Ziyi et al.2026 · 0 citationsJournal of the American Medical Informatics Association : JAMIA. 2026
Structural requirements for intelligent clinical digital twins in feedback-driven care.42686992 · Wang, Yanfei, Smith, Glenn E, Lipori, Gloria P et al.2026 · 0 citationsnpj health systems. 2026;3(1)
A study of large language models for patient information extraction: Model architecture, fine-tuning strategy, and multi-task instruction tuning.41905520 · Peng, Cheng, Dong, Xinyu, Lyu, Mengxian et al.2026 · 1 citationJournal of biomedical informatics. 2026;178:105034
Developing and Validating a Machine Learning Algorithm to Predict the Risk of Incident Opioid Use Disorder Among OneFlorida+ Patients: Prognostic Modeling Study.41813236 · Faysal, Jabed Al, Lo-Ciganic, Weihsuan, Gellad, Walid F et al.2026 · 0 citationsJournal of medical Internet research. 2026;28:e79482
Adjusting Covariate Misclassification in Electronic Health Records-Based Machine Learning Prediction Models.41726448 · Yang, Shuang, Wu, Yonghui, Liu, Mei et al.2026 · 0 citationsAMIA ... Annual Symposium proceedings. AMIA Symposium. 2024;2024:1444-1453
A topic modeling analysis of stigma dimensions, social, and related behavioral circumstances in clinical notes among patients with HIV.41518824 · Chen, Ziyi, Liu, Yiyang, Prosperi, Mattia et al.2026 · 1 citationInternational journal of medical informatics. 2026;209:106269
Narrative Feature or Structured Feature? A Study of Large Language Models to Identify Cancer Patients at Risk of Heart Failure2403.11425v4 · Ziyi Chen, Mengyuan Zhang, Mustafa Mohammed Ahmed et al.2024 · 3 citationsarXiv
Feasibility of identifying factors related to Alzheimer's disease and related dementia in real-world data.42040178 · Huang, Yu, Li, Qian, Chen, Aokun et al.2026 · 0 citationsJAMIA open. 2026;9(2):ooag060
Medical foundation large language models for comprehensive text analysis and beyond.40044845 · Xie, Qianqian, Chen, Qingyu, Chen, Aokun et al.2026 · 66 citationsNPJ digital medicine. 2025;8(1):141
CLAMP - a toolkit for efficiently building customized clinical natural language processing pipelines.29186491 · Soysal, Ergin, Wang, Jingqi, Jiang, Min et al.2026 · 388 citationsJournal of the American Medical Informatics Association : JAMIA. 2018;25(3):331-336
Determining Thyroid Biopsy Appropriateness: A Retrospective Study Integrating Clinical Context and Ultrasound Features.41133765 · Habib, Saad, Perez Rodriguez Garcia, Gilberto, Zhang, Zhongyue et al.2026 · 1 citationThe Journal of clinical endocrinology and metabolism. 2026;111(4):e1014-e1022
Clinical Prediction Models for Hospital-Induced Delirium Using Structured and Unstructured Electronic Health Record Data: Protocol for a Development and Validation Study.37943599 · Ser, Sarah E, Shear, Kristen, Snigurska, Urszula A et al.2026 · 6 citationsJMIR research protocols. 2023;12:e48521
Me-LLaMA: Medical Foundation Large Language Models for Comprehensive Text Analysis and Beyond.39764122 · Xie, Qianqian, Chen, Qingyu, Chen, Aokun et al.2026 · 2 citationsResearch square. 2024
Multi-scale target-aware representation learning for fundus image enhancement.41223755 · Wu, Haofan, Huang, Yin, Wu, Yuqing et al.2026 · 1 citationNeural networks : the official journal of the International Neural Network Society. 2026;195:108273
Seedance 1.5 pro: A Native Audio-Visual Joint Generation Foundation Model2512.13507v3 · Team Seedance, Heyi Chen, Siyan Chen et al.2025 · 0 citationsarXiv
Protocol for a Single-Arm Pilot Clinical Trial: Developing and Evaluating a Machine Learning Opioid Prediction & Risk-Stratification E-Platform (DEMONSTRATE).41375825 · Hong, Je-Won J, Wilson, Debbie L, Nguyen, Khoa et al.2025 · 3 citationsJournal of clinical medicine. 2025;14(23)
Understanding the preference of online health information seeking among college students using the best-worst scaling method.41268393 · Wang, Dan, Jiang, Wang, Chen, Haihong et al.2025 · 1 citationFrontiers in public health. 2025;13:1670106
Multi-Scale Target-Aware Representation Learning for Fundus Image Enhancement2505.01831v2 · Haofan Wu, Yin Huang, Yuqing Wu et al.2025 · 1 citationarXiv
Real-world analysis of cardiovascular adverse events and risk factors after immune checkpoint inhibitor therapy.41250252 · Wang, Yanfei, Niu, Shu, Doonan, Bently P et al.2025 · 7 citationsCardio-oncology (London, England). 2025;11(1):107
Scaling up biomedical vision-language models: Fine-tuning, instruction tuning, and multi-modal learning.41138953 · Peng, Cheng, Zhang, Kai, Lyu, Mengxian et al.2025 · 7 citationsJournal of biomedical informatics. 2025;171:104946
AI uses medical records to accurately predict onset of disease 20 years into the future.41028177 · Wu, Yonghui2025 · 1 citationNature. 2025;647(8088):44-45
A Trustworthy Industrial Fault Diagnosis Architecture Integrating Probabilistic Models and Large Language Models2510.03815v1 · Yue wu2025 · 0 citationsarXiv
Development and validation of natural language processing algorithms in the national ENACT network.40979101 · Wang, Yanshan, Hilsman, Jordan, Li, Chenyu et al.2025 · 1 citationJournal of clinical and translational science. 2025;9(1):e199
Determination of pediatric asthma severity using retrospective electronic health record prescription data: an analysis before and after the 2020 national asthma education and prevention program guideline update.40459987 · Pan, Jinqian, Xu, Jie, Fedele, David et al.2025 · 0 citationsThe Journal of asthma : official journal of the Association for the Care of Asthma. 2025;62(10):1729-1739
A Topic Modeling Analysis of Stigma Dimensions, Social, and Related Behavioral Circumstances in Clinical Notes Among Patients with HIV2506.09279v2 · Ziyi Chen, Yiyang Liu, Mattia Prosperi et al.2025 · 1 citationarXiv
Development and Validation of a Machine Learning-Based Screening Algorithm to Predict High-Risk Hepatitis C Infection.40874186 · Jang, Suk-Chan, Lo-Ciganic, Wei-Hsuan, Hernandez-Con, Pilar et al.2025 · 3 citationsOpen forum infectious diseases. 2025;12(8):ofaf496
A Study of Large Language Models for Patient Information Extraction: Model Architecture, Fine-Tuning Strategy, and Multi-task Instruction Tuning2509.04753v1 · Cheng Peng, Xinyu Dong, Mengxian Lyu et al.2025 · 1 citationarXiv
Patient and nodule characteristics associated with adherence to lung cancer screening in a large integrated healthcare system.40783586 · Yang, Shuang, Liang, Muxuan, Mehta, Hiren J et al.2025 · 2 citationsScientific reports. 2025;15(1):29172
Me-LLaMA: Foundation Large Language Models for Medical Applications.38826372 · Xie, Qianqian, Chen, Qingyu, Chen, Aokun et al.2025 · 64 citationsResearch square. 2024
Identifying Opioid Overdose and Opioid Use Disorder and Related Information from Clinical Narratives Using Large Language Models.40502234 · Paredes, Daniel, Talankar, Sankalp, Peng, Cheng et al.2025 · 0 citationsAMIA Joint Summits on Translational Science proceedings. AMIA Joint Summits on Translational Science. 2025;2025:414-421
LabQAR: A Manually Curated Dataset for Question Answering on Laboratory Test Reference Ranges and Interpretation10.1101/2025.06.03.25328882v1 · Bhasuran, B., Jin, Q., Deville, A. et al.2025 · 0 citationsmedRxiv
Measurement of Semantic Textual Similarity in Clinical Texts: Comparison of Transformer-Based Models.33226350 · Yang, Xi, He, Xing, Zhang, Hansi et al.2025 · 60 citationsJMIR medical informatics. 2020;8(11):e19735
Extracting Family History of Patients From Clinical Narratives: Exploring an End-to-End Solution With Deep Learning Models.33320104 · Yang, Xi, Zhang, Hansi, He, Xing et al.2025 · 28 citationsJMIR medical informatics. 2020;8(12):e22982
Procedural complications associated with invasive diagnostic procedures after lung cancer screening with low-dose computed tomography.35124410 · Yang, Shuang, Shih, Ya-Chen Tina, Huo, Jinhai et al.2025 · 8 citationsLung cancer (Amsterdam, Netherlands). 2022;165:141-144
Scaling Up Biomedical Vision-Language Models: Fine-Tuning, Instruction Tuning, and Multi-Modal Learning2505.17436v1 · Cheng Peng, Kai Zhang, Mengxian Lyu et al.2025 · 7 citationsarXiv
COMPAC: COMputable Phenotype for Asthma in Children.40386436 · Fishe, Jennifer, Pan, Jinqian, Fedele, David et al.2025 · 1 citationResearch square. 2025
Narrative Feature or Structured Feature? A Study of Large Language Models to Identify Cancer Patients at Risk of Heart Failure.40417538 · Chen, Ziyi, Zhang, Mengyuan, Ahmed, Mustafa Mohammed et al.2025 · 3 citationsAMIA ... Annual Symposium proceedings. AMIA Symposium. 2024;2024:242-251
Gemini: A Family of Highly Capable Multimodal Models2312.11805v5 · Gemini Team Google: Rohan Anil, Sebastian Borgeaud, Jean-Baptiste Alayrac et al.2023 · 838 citationsarXiv
Toward a Computable Phenotype for Determining Eligibility of Lung Cancer Screening Using Electronic Health Records.39818952 · Yang, Shuang, Huang, Yu, Lou, Xiwei et al.2025 · 2 citationsJCO clinical cancer informatics. 2025;9:e2400139
Thyroid Ultrasound Appropriateness Identification Through Natural Language Processing of Electronic Health Records.38501072 · Jacome, Cristian Soto, Torres, Danny Segura, Fan, Jungwei W et al.2025 · 9 citationsMayo Clinic proceedings. Digital health. 2024;2(1):67-74
AutoRADP: An Interpretable Deep Learning Framework to Predict Rapid Progression for Alzheimer's Disease and Related Dementias Using Electronic Health Records10.1101/2025.04.06.25325337v1 · Yang, Q., Meng, W., Zhuang, P. et al.2025 · 0 citationsmedRxiv
Development and Validation of Natural Language Processing Algorithms in the ENACT National Electronic Health Record Research Network10.1101/2025.01.24.25321096v1 · Wang, Y., Hilsman, J., Li, C. et al.2025 · 4 citationsmedRxiv
Lung cancer screening adherence among people living with and without HIV: An analysis of an integrated health system in Florida, United States (2012-2021).37546581 · Islam, Jessica Y, Yang, Shuang, Schabath, Matthew et al.2025 · 9 citationsPreventive medicine reports. 2023;35:102334
Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context2403.05530v5 · Gemini Team Google: Petko Georgiev, Ving Ian Lei, Ryan Burnell et al.2024 · 296 citationsarXiv
Developing an automated algorithm for identification of children and adolescents with diabetes using electronic health records from the OneFlorida+ clinical research network.39344840 · Li, Piaopiao, Spector, Eliot, Alkhuzam, Khalid et al.2024 · 2 citationsDiabetes, obesity & metabolism. 2025;27(1):102-110
Robust Detection of LLM-Generated Text: A Comparative Analysis2411.06248v1 · Yongye Su, Yuqing Wu2024 · 0 citationsarXiv
Me LLaMA: Foundation Large Language Models for Medical Applications2402.12749v5 · Qianqian Xie, Qingyu Chen, Aokun Chen et al.2024 · 64 citationsarXiv
Design and development of a machine-learning-driven opioid overdose risk prediction tool integrated in electronic health records in primary care settings.39420438 · Nguyen, Khoa, Wilson, Debbie L, Diiulio, Julie et al.2024 · 12 citationsBioelectronic medicine. 2024;10(1):24
Generative AI's aggregated knowledge versus web-based curated knowledge2410.12091v1 · Ted Selker, Yunzi Wu2024 · 0 citationsarXiv
Career total: 319 works. 233 are in this corpus.Showing the 50 most recent.
Profile built from the corpus for this byline.
Author records are still filling in while the Hub is in alpha. If this is your page, you'll be able to claim it soon. Spot a mistake? Tell us.