All Competitors

Every biological foundation model, evaluated and ranked by the bio.rodeo team

Showing 97120 of 125 filtered models

  • RiNALMo

    169135
    LBCB SciFebruary 29, 2024foundation_modellanguage_modelstructure_prediction

    RNA language model with 650M parameters pretrained on 36 million non-coding RNA sequences, generalizing structure prediction to unseen RNA families.

    RNA
    72Openness
  • ProLLaMA

    20795304
    PKU-YuanGroupFebruary 26, 2024foundation_modellanguage_modelllama+1

    Protein large language model adapted from LLaMA-2 that unifies sequence generation and superfamily classification in one 7B-parameter framework.

    Protein
    95Openness
  • scMulan

    626
    Tsinghua UniversityJanuary 29, 2024foundation_modelgenerativelanguage_model+1

    Generative language model for single-cell transcriptomics with 368M parameters, unifying cell type annotation, batch integration, and cell generation.

    Single-cell
    48Openness
  • RNA-MSM

    711091.3K
    Peking University +1 otherJanuary 11, 2024foundation_modellanguage_modelstructure_prediction

    RNA language model trained on multiple sequence alignments of Rfam families, predicting secondary structure and solvent accessibility from homology.

    RNA
    61Openness
  • IgLM

    192130
    GrayLabNovember 15, 2023antibodyfoundation_modelimmunology+1

    Generative language model trained on 558 million antibody sequences for infilling-based design of CDR loops and full-length immunoglobulin sequences.

    Protein
    12Openness
  • ProGen2

    705
    SalesforceOctober 30, 2023foundation_modelgenerativelanguage_model+1

    Protein language models from 151M to 6.4B parameters, trained on over a billion sequences for sequence generation and zero-shot fitness prediction.

    Protein
    55Openness
  • BioT5

    127191
    Renmin University of ChinaOctober 11, 2023drug_discoveryfoundation_modellanguage_model+1

    Encoder-decoder framework unifying molecules, proteins, and natural language with SELFIES notation for cross-modal drug discovery tasks.

    Language modelSmall moleculeProtein
    74Openness
  • MasterAI EAMAugust 1, 2023foundation_modellanguage_model

    Open large language models for natural science, fine-tuned on physics, chemistry, and materials science literature with automated instruction tuning.

    Language model
    24Openness
  • ProstT5

    31817.2K
    RostlabJuly 25, 2023foundation_modelinverse_foldinglanguage_model+5

    Bilingual protein language model that translates bidirectionally between amino acid sequences and the 3Di structural alphabet for inverse folding.

    Protein
    76Openness
  • TULIP

    1343
    Ecole Normale SuperieureJuly 19, 2023antibodydrug_discoverylanguage_model+3

    Unsupervised transformer language model for TCR-epitope binding prediction that generalizes to unseen epitopes without needing negative examples.

    Protein
    60Openness
  • DNABERT-2

    507456170.4K
    MAGICS LabJune 26, 2023cross_speciesdnafoundation_model+2

    Multi-species genomic foundation model swapping k-mer tokenization for byte pair encoding, matching Nucleotide Transformer with 21x fewer parameters.

    DNA & Gene
    64Openness
  • LLaVA-Med

    2.2K1.9K12.3K
    Microsoft ResearchJune 1, 2023histologyimage_captioninginstruction_tuning+7

    Biomedical vision-language assistant for question answering on radiology and pathology images, adapted from LLaVA on PubMed Central captions.

    PathologyLanguage model
    28Openness
  • tGPT

    1762159
    Tianjin Medical University Cancer Institute and HospitalApril 20, 2023cell_type_annotationfoundation_modellanguage_model+4

    Single-cell foundation model pre-trained on 22 million transcriptomes, using rank-based gene encoding for clustering and trajectory inference.

    Single-cell
    50Openness
  • Technical University of MunichJanuary 27, 2023bertdnafoundation_model+6

    Masked DNA language model trained on over 800 vertebrate genomes and conditioned on species identity to learn conserved regulatory sequence features.

    DNA & Gene
    76Openness
  • Ankh

    249733.2K
    Technical University of MunichJanuary 16, 2023efficient_inferenceembeddingsfoundation_model+1

    Parameter-efficient protein language model that matches larger models such as ESM-2 on protein prediction tasks using under 10% of the parameters.

    Protein
    24Openness
  • Biomed AIJanuary 1, 2023bertfoundation_modellanguage_model+1

    RNA language model pre-trained on 2M+ pre-mRNA sequences from 72 vertebrate species for splice-site prediction and variant effect analysis.

    RNA
    77Openness
  • ReprogBERT

    2439
    IBMJanuary 1, 2023antibodyfoundation_modellanguage_model

    Antibody CDR design model that reprograms a frozen English BERT for sequence infilling, avoiding training a dedicated protein language model.

    Protein
    56Openness
  • Galactica

    2.7K1.1K623
    Meta AINovember 16, 2022foundation_modellanguage_modelmultimodal

    Scientific large language model trained on 48 million papers, textbooks, and reference works to store, combine, and reason about scientific knowledge.

    Language model
    46Openness
  • BioGPT

    4.5K1.5K101.4K
    Microsoft Research Asia +1 otherOctober 19, 2022biomedical_literaturegenerativegpt+6

    Generative transformer pretrained on PubMed abstracts for biomedical text generation and mining, including relation extraction and question answering.

    Language model
    66Openness
  • GenSLM

    142142
    Argonne National LaboratoryOctober 12, 2022foundation_modelgene_expressiongenomics+4

    Genome-scale language model trained on prokaryotic genes and SARS-CoV-2 genomes to model viral evolution and flag emerging variants of concern.

    DNA & Gene
    56Openness
  • MoLFormer-XL

    406595209.9K
    IBM ResearchOctober 3, 2022drug_discoveryfoundation_modellanguage_model+4

    Large-scale chemical language model trained on 1.1 billion SMILES strings using linear attention transformers for molecular property prediction.

    Small molecule
    86Openness
  • RNA-FM

    386257
    ml4bio +3 othersAugust 6, 2022foundation_modellanguage_modelstructure_prediction

    RNA foundation model pretrained on 23.7 million non-coding RNA sequences, producing embeddings for structure prediction, annotation, and RNA design.

    RNA
    62Openness
  • ESM-2 & ESMFold

    4.2K5.1K1.5M
    Meta AIJuly 20, 2022foundation_modellanguage_modelstructure_prediction

    Meta AI's family of protein language models scaled to 15B parameters, paired with ESMFold for fast, alignment-free atomic-level structure prediction.

    Protein
    83Openness
  • AntiBERTa

    65160
    AlchemabMay 18, 2022antibodyfoundation_modelimmunology+1

    BERT-based antibody language model pretrained on 57M B cell receptor sequences for paratope prediction and convergent antibody discovery.

    Protein
    60Openness