cs.CLcs.LGeess.AScs.SDNatural Language ProcessingElectronic Health Recordscs.AInatural language processingmachine learningDeep Learning

On Valency

published · living versions
W_39gz77fp·v1 · currentpublished
Exploring the Limits of Language Modeling
with Rafal Jozefowicz, Oriol Vinyals, Mike Schuster, Noam Shazeer
1 version

Preprints & journals

233 papers in the corpus · 2003–2026
A large language model for electronic health records.36572766 · Yang, Xi, Chen, Aokun, PourNejatian, Nima et al.2026 · 880 citationsNPJ digital medicine. 2022;5(1):194
Less can be better: decomposing clinical data modalities in large language model-based healthcare applications.42700364 · Peng, Cheng, Lyu, Mengxian, Chen, Ziyi et al.2026 · 0 citationsJournal of the American Medical Informatics Association : JAMIA. 2026
Structural requirements for intelligent clinical digital twins in feedback-driven care.42686992 · Wang, Yanfei, Smith, Glenn E, Lipori, Gloria P et al.2026 · 0 citationsnpj health systems. 2026;3(1)
A study of large language models for patient information extraction: Model architecture, fine-tuning strategy, and multi-task instruction tuning.41905520 · Peng, Cheng, Dong, Xinyu, Lyu, Mengxian et al.2026 · 1 citationJournal of biomedical informatics. 2026;178:105034
Developing and Validating a Machine Learning Algorithm to Predict the Risk of Incident Opioid Use Disorder Among OneFlorida+ Patients: Prognostic Modeling Study.41813236 · Faysal, Jabed Al, Lo-Ciganic, Weihsuan, Gellad, Walid F et al.2026 · 0 citationsJournal of medical Internet research. 2026;28:e79482
Adjusting Covariate Misclassification in Electronic Health Records-Based Machine Learning Prediction Models.41726448 · Yang, Shuang, Wu, Yonghui, Liu, Mei et al.2026 · 0 citationsAMIA ... Annual Symposium proceedings. AMIA Symposium. 2024;2024:1444-1453
A topic modeling analysis of stigma dimensions, social, and related behavioral circumstances in clinical notes among patients with HIV.41518824 · Chen, Ziyi, Liu, Yiyang, Prosperi, Mattia et al.2026 · 1 citationInternational journal of medical informatics. 2026;209:106269
Feasibility of identifying factors related to Alzheimer's disease and related dementia in real-world data.42040178 · Huang, Yu, Li, Qian, Chen, Aokun et al.2026 · 0 citationsJAMIA open. 2026;9(2):ooag060
Medical foundation large language models for comprehensive text analysis and beyond.40044845 · Xie, Qianqian, Chen, Qingyu, Chen, Aokun et al.2026 · 66 citationsNPJ digital medicine. 2025;8(1):141
CLAMP - a toolkit for efficiently building customized clinical natural language processing pipelines.29186491 · Soysal, Ergin, Wang, Jingqi, Jiang, Min et al.2026 · 388 citationsJournal of the American Medical Informatics Association : JAMIA. 2018;25(3):331-336
Determining Thyroid Biopsy Appropriateness: A Retrospective Study Integrating Clinical Context and Ultrasound Features.41133765 · Habib, Saad, Perez Rodriguez Garcia, Gilberto, Zhang, Zhongyue et al.2026 · 1 citationThe Journal of clinical endocrinology and metabolism. 2026;111(4):e1014-e1022
Clinical Prediction Models for Hospital-Induced Delirium Using Structured and Unstructured Electronic Health Record Data: Protocol for a Development and Validation Study.37943599 · Ser, Sarah E, Shear, Kristen, Snigurska, Urszula A et al.2026 · 6 citationsJMIR research protocols. 2023;12:e48521
Me-LLaMA: Medical Foundation Large Language Models for Comprehensive Text Analysis and Beyond.39764122 · Xie, Qianqian, Chen, Qingyu, Chen, Aokun et al.2026 · 2 citationsResearch square. 2024
Multi-scale target-aware representation learning for fundus image enhancement.41223755 · Wu, Haofan, Huang, Yin, Wu, Yuqing et al.2026 · 1 citationNeural networks : the official journal of the International Neural Network Society. 2026;195:108273
Seedance 1.5 pro: A Native Audio-Visual Joint Generation Foundation Model2512.13507v3 · Team Seedance, Heyi Chen, Siyan Chen et al.2025 · 0 citationsarXiv
Protocol for a Single-Arm Pilot Clinical Trial: Developing and Evaluating a Machine Learning Opioid Prediction & Risk-Stratification E-Platform (DEMONSTRATE).41375825 · Hong, Je-Won J, Wilson, Debbie L, Nguyen, Khoa et al.2025 · 3 citationsJournal of clinical medicine. 2025;14(23)
Understanding the preference of online health information seeking among college students using the best-worst scaling method.41268393 · Wang, Dan, Jiang, Wang, Chen, Haihong et al.2025 · 1 citationFrontiers in public health. 2025;13:1670106
Multi-Scale Target-Aware Representation Learning for Fundus Image Enhancement2505.01831v2 · Haofan Wu, Yin Huang, Yuqing Wu et al.2025 · 1 citationarXiv
Real-world analysis of cardiovascular adverse events and risk factors after immune checkpoint inhibitor therapy.41250252 · Wang, Yanfei, Niu, Shu, Doonan, Bently P et al.2025 · 7 citationsCardio-oncology (London, England). 2025;11(1):107
Scaling up biomedical vision-language models: Fine-tuning, instruction tuning, and multi-modal learning.41138953 · Peng, Cheng, Zhang, Kai, Lyu, Mengxian et al.2025 · 7 citationsJournal of biomedical informatics. 2025;171:104946
Development and validation of natural language processing algorithms in the national ENACT network.40979101 · Wang, Yanshan, Hilsman, Jordan, Li, Chenyu et al.2025 · 1 citationJournal of clinical and translational science. 2025;9(1):e199
Determination of pediatric asthma severity using retrospective electronic health record prescription data: an analysis before and after the 2020 national asthma education and prevention program guideline update.40459987 · Pan, Jinqian, Xu, Jie, Fedele, David et al.2025 · 0 citationsThe Journal of asthma : official journal of the Association for the Care of Asthma. 2025;62(10):1729-1739
Development and Validation of a Machine Learning-Based Screening Algorithm to Predict High-Risk Hepatitis C Infection.40874186 · Jang, Suk-Chan, Lo-Ciganic, Wei-Hsuan, Hernandez-Con, Pilar et al.2025 · 3 citationsOpen forum infectious diseases. 2025;12(8):ofaf496
Patient and nodule characteristics associated with adherence to lung cancer screening in a large integrated healthcare system.40783586 · Yang, Shuang, Liang, Muxuan, Mehta, Hiren J et al.2025 · 2 citationsScientific reports. 2025;15(1):29172
Me-LLaMA: Foundation Large Language Models for Medical Applications.38826372 · Xie, Qianqian, Chen, Qingyu, Chen, Aokun et al.2025 · 64 citationsResearch square. 2024
Identifying Opioid Overdose and Opioid Use Disorder and Related Information from Clinical Narratives Using Large Language Models.40502234 · Paredes, Daniel, Talankar, Sankalp, Peng, Cheng et al.2025 · 0 citationsAMIA Joint Summits on Translational Science proceedings. AMIA Joint Summits on Translational Science. 2025;2025:414-421
Measurement of Semantic Textual Similarity in Clinical Texts: Comparison of Transformer-Based Models.33226350 · Yang, Xi, He, Xing, Zhang, Hansi et al.2025 · 60 citationsJMIR medical informatics. 2020;8(11):e19735
Extracting Family History of Patients From Clinical Narratives: Exploring an End-to-End Solution With Deep Learning Models.33320104 · Yang, Xi, Zhang, Hansi, He, Xing et al.2025 · 28 citationsJMIR medical informatics. 2020;8(12):e22982
Procedural complications associated with invasive diagnostic procedures after lung cancer screening with low-dose computed tomography.35124410 · Yang, Shuang, Shih, Ya-Chen Tina, Huo, Jinhai et al.2025 · 8 citationsLung cancer (Amsterdam, Netherlands). 2022;165:141-144
COMPAC: COMputable Phenotype for Asthma in Children.40386436 · Fishe, Jennifer, Pan, Jinqian, Fedele, David et al.2025 · 1 citationResearch square. 2025
Narrative Feature or Structured Feature? A Study of Large Language Models to Identify Cancer Patients at Risk of Heart Failure.40417538 · Chen, Ziyi, Zhang, Mengyuan, Ahmed, Mustafa Mohammed et al.2025 · 3 citationsAMIA ... Annual Symposium proceedings. AMIA Symposium. 2024;2024:242-251
Gemini: A Family of Highly Capable Multimodal Models2312.11805v5 · Gemini Team Google: Rohan Anil, Sebastian Borgeaud, Jean-Baptiste Alayrac et al.2023 · 838 citationsarXiv
Toward a Computable Phenotype for Determining Eligibility of Lung Cancer Screening Using Electronic Health Records.39818952 · Yang, Shuang, Huang, Yu, Lou, Xiwei et al.2025 · 2 citationsJCO clinical cancer informatics. 2025;9:e2400139
Thyroid Ultrasound Appropriateness Identification Through Natural Language Processing of Electronic Health Records.38501072 · Jacome, Cristian Soto, Torres, Danny Segura, Fan, Jungwei W et al.2025 · 9 citationsMayo Clinic proceedings. Digital health. 2024;2(1):67-74
Lung cancer screening adherence among people living with and without HIV: An analysis of an integrated health system in Florida, United States (2012-2021).37546581 · Islam, Jessica Y, Yang, Shuang, Schabath, Matthew et al.2025 · 9 citationsPreventive medicine reports. 2023;35:102334
Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context2403.05530v5 · Gemini Team Google: Petko Georgiev, Ving Ian Lei, Ryan Burnell et al.2024 · 296 citationsarXiv
Developing an automated algorithm for identification of children and adolescents with diabetes using electronic health records from the OneFlorida+ clinical research network.39344840 · Li, Piaopiao, Spector, Eliot, Alkhuzam, Khalid et al.2024 · 2 citationsDiabetes, obesity & metabolism. 2025;27(1):102-110
Me LLaMA: Foundation Large Language Models for Medical Applications2402.12749v5 · Qianqian Xie, Qingyu Chen, Aokun Chen et al.2024 · 64 citationsarXiv
Design and development of a machine-learning-driven opioid overdose risk prediction tool integrated in electronic health records in primary care settings.39420438 · Nguyen, Khoa, Wilson, Debbie L, Diiulio, Julie et al.2024 · 12 citationsBioelectronic medicine. 2024;10(1):24
Career total: 319 works. 233 are in this corpus.Showing the 50 most recent.

Profile built from the corpus for this byline.

Author records are still filling in while the Hub is in alpha. If this is your page, you'll be able to claim it soon. Spot a mistake? Tell us.