2026
article
OpenAlex
Timothy Keyes, Alison Callahan, Abby Pandya, Nerissa Ambers et autres
Postdeployment monitoring of artificial intelligence (AI) systems in health care is essential to ensure their safety, quality, and sustained benefit - and to support governance decisions about which systems to update, modify, or decommission. Motivated by these needs, the authors developed a …
us
(code pays fourni par la source)
Accès ouvert
2026
preprint
OpenAlex
Nigam H. Shah, Nerissa Ambers, Abby Pandya, Timothy Keyes et autres
While large language models (LLMs) can support clinical documentation needs, standalone tools struggle with "workflow friction" from manual data entry. We developed ChatEHR, a system that enables the use of LLMs with the entire patient timeline spanning several years. ChatEHR enables automations …
Accès ouvert
2026
preprint
OpenAlex
Nigam H. Shah, Nerissa Ambers, Abby Pandya, Timothy Keyes et autres
While large language models (LLMs) can support clinical documentation needs, standalone tools struggle with "workflow friction" from manual data entry. We developed ChatEHR, a system that enables the use of LLMs with the entire patient timeline spanning several years. ChatEHR enables automations …
us
(code pays fourni par la source)
Accès ouvert
2026
article
OpenAlex
Suhana Bedi, Hejie Cui, Miguel Fuentes, Alyssa Unell et autres
us
(code pays fourni par la source)
2026
preprint
OpenAlex
Timothy Keyes, Alison Callahan, Abby Pandya, Nerissa Ambers et autres
assesses whether a deployed system continues to have value in the form of benefit to clinicians, staff, and patients. Drawing on examples of deployed AI systems at our academic medical center, we provide practical guidance for creating monitoring plans based on these …
Accès ouvert
2025
preprint
OpenAlex
Timothy Keyes, Alison Callahan, Abby Pandya, Nerissa Ambers et autres
Post-deployment monitoring of artificial intelligence (AI) systems in health care is essential to ensure their safety, quality, and sustained benefit-and to support governance decisions about which systems to update, modify, or decommission. Motivated by these needs, we developed a framework for monitoring …
Accès ouvert
2025
preprint
OpenAlex
Suhana Bedi, Hejie Cui, Miguel Fuentes, Alyssa Unell et autres
While large language models (LLMs) achieve near-perfect scores on medical licensing exams, these evaluations inadequately reflect the complexity and diversity of real-world clinical practice. We introduce MedHELM, an extensible evaluation framework for assessing LLM performance for medical tasks with three key contributions. …
Accès ouvert
2025
article
OpenAlex
Crystal Chang, Hodan Farah, Haiwen Gui, Shawheen J. Rezaei et autres
Red teaming, the practice of adversarially exposing unexpected or undesired model behaviors, is critical towards improving equity and accuracy of large language models, but non-model creator-affiliated red teaming is scant in healthcare. We convened teams of clinicians, medical and engineering students, and …
us, ca, de
(code pays fourni par la source)
Accès ouvert
2025
article
OpenAlex
Obbina Abani, Amr E. Abbas, Fatima Abbas, Jaffar Abbas et autres
Background: Low dose corticosteroids (e.g., 6 mg dexamethasone) have been shown to reduce mortality for hypoxic COVID-19 patients. We have previously reported that higher dose corticosteroids cause harm in patients with clinical hypoxia but not receiving ventilatory support (the combination of non-invasive …
Accès ouvert
2025
article
OpenAlex
April S. Liang, Juan M. Banda, Thomas Savage, Abby Pandya et autres
This study evaluates the feasibility of using GPT-4 to automate precharting for specialty referrals, focusing on new patients referred to an otolaryngology clinic for nasal congestion. We describe the design decisions and strategies tested in creating this precharting utility, including methods for …
us
(code pays fourni par la source)
Accès ouvert
2024
article
OpenAlex
Ramya Tekumalla, Juan M. Banda
Electronic phenotyping involves a detailed analysis of both structured and unstructured data, employing rule-based methods, machine learning, natural language processing, and hybrid approaches. Currently, the development of accurate phenotype definitions demands extensive literature reviews and clinical experts, rendering the process time-consuming and …
us
(code pays fourni par la source)
2024
article
OpenAlex
Alison Callahan, Duncan C. McElfresh, Juan M. Banda, Gabrielle Bunney et autres
us
(code pays fourni par la source)