Efficient Protein Engineering via Integrated Language Models and Bayesian Optimization
Rattachement africain : us. Niveau de preuve : code pays fourni par la source.
Le résumé fourni par la source
Abstract This study investigates the application of advanced predictive models to reduce the cost and effort associated with protein engineering campaigns. We explore the use of protein language models (PLMs), a variant of large language models (LLMs), to predict functional performance from protein sequences. A common challenge in this domain is the scarcity of functional data. To address this, we examine zero-shot and few-shot learning methods. Another challenge is efficiently searching the vast fitness landscape for superior protein variants. We evaluate search methods, such as Bayesian optimization, to tackle this problem. The proposed methods are evaluated against a benchmark of 34 protein datasets containing sequences and their quantified functional values. Our findings demonstrate the potential of these advanced predictive models to streamline and accelerate the protein engineering process.
Ce résumé expose les affirmations des auteurs. BNTIC ne l’interprète pas comme une validation indépendante des résultats.
Le contrôle bibliographique ouvert
DOI retrouvé dans Crossref DOI retrouvé ; titre concordant.
- Titre Crossref
- Efficient Protein Engineering via Integrated Language Models and Bayesian Optimization
- Date Crossref
- 02/10/2025
- Éditeur
- openRxiv
- Type
- posted-content
Ce recoupement confirme des métadonnées liées au DOI. Il ne confirme ni la méthode ni les conclusions de l’étude, et il ne compte pas comme une seconde source scientifique indépendante.
Où se fait cette recherche
-
WinnMed pays non établi dans la noticeOrganisation à but non lucratif
-
Mayo Clinic pays non établi dans la noticeÉtablissement de santé
-
LGC pays non établi dans la noticeInstitution
WinnMed, Mayo Clinic et LGC.
Une affiliation ne permet pas de déduire la nationalité d’un auteur.