All Competitors

Every biological foundation model, evaluated and ranked by the bio.rodeo team

Showing 529552 of 552 filtered models

  • Shenzhen Research Institute of Big Data +2 othersSeptember 15, 2022chest_x_rayfoundation_modelimage_text_retrieval+7

    Medical vision-language pretraining framework that injects structured medical knowledge into radiology image-text learning for VQA and retrieval.

    ImagingLanguage model
    29Openness
  • scBERT

    359622
    Tencent AI LabSeptember 1, 2022cell_type_annotationfoundation_model

    Pretrained transformer for cell type annotation of scRNA-seq data. Trained on 1.1M cells; outperforms supervised methods on cross-dataset transfer.

    Single-cell
    46Openness
  • RNA-FM

    386257
    ml4bio +3 othersAugust 6, 2022foundation_modellanguage_modelstructure_prediction

    RNA foundation model pretrained on 23.7 million non-coding RNA sequences, producing embeddings for structure prediction, annotation, and RNA design.

    RNA
    62Openness
  • MoDNA

    27
    University of Texas at ArlingtonAugust 1, 2022dnafoundation_modelgenomics+2

    Motif-oriented DNA pre-training framework that adds motif prediction to an ELECTRA generator-discriminator setup for motif-aware genomic embeddings.

    DNA & Gene
    11Openness
  • ProtGPT2

    8698.2K
    University of BayreuthJuly 27, 2022foundation_modelgenerativeprotein_design

    Autoregressive protein language model based on GPT-2 that generates de novo protein sequences sampling unexplored regions of protein space.

    Protein
    54Openness
  • ESM-2 & ESMFold

    4.2K5.1K1.5M
    Meta AIJuly 20, 2022foundation_modellanguage_modelstructure_prediction

    Meta AI's family of protein language models scaled to 15B parameters, paired with ESMFold for fast, alignment-free atomic-level structure prediction.

    Protein
    83Openness
  • Casanovo

    1944
    Noble LabJuly 17, 2022foundation_modelmass_spectrometryproteomics

    Transformer model for de novo peptide sequencing that reads amino acid sequences directly from tandem mass spectra, with no protein sequence database.

    Protein
    91Openness
  • CARP

    259
    Microsoft ResearchMay 19, 2022embeddingsfoundation_modelvariant_effect_prediction

    Protein language model family built on CNNs rather than transformers, matching transformer quality while scaling linearly with sequence length.

    Protein
    81Openness
  • AntiBERTa

    65160
    AlchemabMay 18, 2022antibodyfoundation_modelimmunology+1

    BERT-based antibody language model pretrained on 57M B cell receptor sequences for paratope prediction and convergent antibody discovery.

    Protein
    60Openness
  • RNABERT

    561345.1K
    Keio UniversityFebruary 22, 2022foundation_modellanguage_modelstructure_prediction

    RNA language model that learns base-level embeddings capturing sequence context and secondary structure, enabling fast structural alignment.

    RNA
    34Openness
  • OntoProtein

    152140210
    Zhejiang UniversityJanuary 28, 2022foundation_modelgene_ontologyknowledge_graph+1

    Protein language model that fuses Gene Ontology knowledge graphs with masked language modeling, improving protein function and interaction prediction.

    Protein
    63Openness
  • ProteinBERT

    579981
    Hebrew University of JerusalemJanuary 13, 2022foundation_modelgene_ontologylanguage_model+1

    Protein language model pretrained on UniRef90 with masked language modeling and Gene Ontology annotation prediction, at 16 million parameters.

    Protein
    86Openness
  • AbLang

    167217
    Oxford Protein Informatics Group (OPIG)January 1, 2022antibodyfoundation_modelimmunology+2

    Antibody-specific language model trained on the OAS database for restoring missing residues and generating high-quality sequence representations.

    Protein
    62Openness
  • GeneBERT

    27
    Carnegie Mellon UniversityOctober 11, 2021bertchromatinfoundation_model+6

    Multi-modal self-supervised transformer for regulatory genomics, pre-trained on DNA sequence together with transcription factor binding matrices.

    DNA & Gene
    18Openness
  • AlphaFold-Multimer

    14.8K3.2K
    Google DeepMindOctober 4, 2021foundation_modelproteomicsstructure_prediction+1

    Protein complex structure prediction model extending AlphaFold 2 with paired MSA processing and ipTM scoring for multi-chain, multimeric assemblies.

    Protein
    59Openness
  • Enformer

    15.1K1.2K
    Google DeepMindOctober 4, 2021dnafoundation_modelgene_expression+2

    Transformer that predicts gene expression and epigenomic signals from 200kb of DNA sequence, capturing distal enhancers up to 100kb from a promoter.

    DNA & Gene
    84Openness
  • ProtTrans

    1.3K1.4K
    RostlabAugust 1, 2021embeddingsfoundation_modelself_supervised+1

    Suite of six protein language models, including ProtBERT and ProtT5, trained on up to 393 billion amino acids without multiple sequence alignments.

    Protein
    71Openness
  • AlphaFold 2

    14.8K37.7K
    Google DeepMindJuly 15, 2021foundation_modelstructure_prediction

    Protein structure prediction model that folds amino acid sequences into 3D structures with atomic accuracy, scoring a median GDT of 92.4 at CASP14.

    Protein
    61Openness
  • ESM-1v

    4.2K902
    Meta AIJuly 9, 2021foundation_modellanguage_modelvariant_effect_prediction

    Protein language model for zero-shot variant effect prediction, scoring mutations by log-odds from evolutionary sequences with no MSA or assay data.

    Protein
    72Openness
  • ESM-1b

    4.2K4.6K
    Meta AIApril 5, 2021embeddingsfoundation_modelvariant_effect_prediction

    Transformer protein language model trained on 250 million protein sequences that learns structural and functional representations without supervision.

    Protein
    71Openness
  • DNABERT

    769813.7K
    Northwestern UniversityFebruary 4, 2021dnafoundation_modelgenomics+2

    Bidirectional transformer for DNA using k-mer tokenization, fine-tunable for promoter, splice site, and transcription factor binding prediction.

    DNA & Gene
    61Openness
  • Big Bird

    6333K341.8K
    Google ResearchJuly 28, 2020dnafoundation_modelgenomics+4

    Sparse attention transformer that extends BERT to 8x longer sequences via random, local, and global attention, with genomic sequence applications.

    DNA & Gene
    49Openness
  • Arizona State University +1 otherAugust 19, 2019classificationcnnct+7

    Self-supervised 3D pretrained models for CT and MRI that learn anatomical representations from unlabeled volumes and transfer to segmentation tasks.

    Imaging
    20Openness
  • UniRep

    3661.1K
    Church LabJanuary 1, 2019embeddingsfoundation_model

    Protein language model using a multiplicative LSTM over 24 million UniRef50 sequences to produce fixed-length embeddings for protein engineering.

    Protein
    49Openness