PlasmidGPT: a generative framework for plasmid design and annotation
Voice is AI-generated
Connected to paperThis paper is a preprint and has not been certified by peer review
PlasmidGPT: a generative framework for plasmid design and annotation
SHAO, B.
AbstractWe introduce PlasmidGPT, a generative language model pretrained on 153k engineered plasmid sequences from Addgene. PlasmidGPT generates de novo sequences that share similar characteristics with engineered plasmids but show low sequence identity to the training data. We demonstrate its ability to generate plasmids in a controlled manner based on the input sequence or specific design constraint. Moreover, our model learns informative embeddings of both engineered and natural plasmids, allowing for efficient prediction of a wide range of sequence-related attributes.