Aller au contenu principal
Accès ouvert déclaré 2026 article

Large Language Models for Optimizing Patient Recruitment Decisions in Voice Data Generation Projects

0Citations signalées — pas une note de qualité
5Institutions déclarées
3Pays d’affiliation déclarés

Résumé fourni par la source

Objectives: Past studies have shown that many clinical machine learning models have performance limitations due to imbalances in the training data. For voice data generation projects, the origin of the problem may lie in the recruiting methods used during data collection efforts. This study introduces a generative AI pipeline for "dataset decision support", recommending recruitment decisions based on high-dimensional insights. Methods: The publicly available GOSSIS-1-eICU dataset was filtered to create patient populations that were relevant to voice data generation projects. Lab results and vital signs from the electronic health record were also used to train a neural network for prediction of disease type. Prediction uncertainty estimates were included in the dataset as approximate indicators of health complexity. To select the best recruitment choice for addressing imbalances, an open-source large language model (LLM) was then instructed to assess dataset statistics and the characteristics of possible participants. Simulations were run in which the system constructed datasets of 250 patients. Results: -value < 0.05). Variables included race, age, BMI, sex, disease type, oxygenation status, co-morbidities, post-operative status, Glasgow Coma Scale verbal response score, the Acute Physiology Score III, prediction uncertainty, vital signs, and lab results. Conclusion: LLMs may provide useful, explainable recommendations when presented with dataset distribution statistics and candidate profiles. In the future, this simulated scenario may be extended to align with conditions in emergency departments or other high-volume settings. Level of Evidence: 3.

Ce résumé expose les affirmations des auteurs. BNTIC ne l’interprète pas comme une validation indépendante des résultats.

Contrôle bibliographique ouvert

DOI retrouvé dans Crossref DOI retrouvé ; titre concordant.

Titre Crossref
Large Language Models for Optimizing Patient Recruitment Decisions in Voice Data Generation Projects
Date Crossref
01/08/2026
Éditeur
Wiley
Type
journal-article

Ce recoupement confirme des métadonnées liées au DOI. Il ne confirme ni la méthode ni les conclusions de l’étude et ne compte pas comme une seconde source scientifique indépendante.

Institutions déclarées

Une affiliation ne permet pas de déduire la nationalité d’un auteur.

Sujets associés

Artificial Intelligence in Healthcare and EducationMachine Learning in HealthcareExplainable Artificial Intelligence (XAI)

BNTIC News n’est pas le producteur de ces données. Recherche à la demande dans Crossref et Europe PMC, sans clé ; OpenAlex reste optionnel. Aucun service payant requis, aucune réponse conservée. Sources et limites.