All Competitors
Every biological foundation model, evaluated and ranked by the bio.rodeo team
Showing 505–528 of 552 filtered models
GENA-LM
230—1.4KFamily of transformer-based DNA language models using BPE tokenization and BigBird sparse attention to reach context lengths up to 36,000 base pairs.
DNA & Gene60OpennessHeartBEiT
25118—Vision transformer for electrocardiograms that reads the printed 12-lead ECG as an image, enabling data-efficient diagnosis from few labeled examples.
Biosignals30OpennessGeneformer
—1.1K4.8KSingle-cell foundation model pretrained on about 30 million human transcriptomes, using rank-value encoding for context-aware gene network inference.
Single-cell96OpennessPathAsst
136104—Multimodal pathology assistant that answers questions about histology and cytology images, pairing the PathCLIP vision encoder with a Vicuna-13B LLM.
PathologyLanguage model17OpennessZero-shot antibody affinity maturation using ESM pseudolikelihood scoring. Improves binding up to 160-fold with no antigen-specific training data.
Protein42OpennesstGPT
1762159Tianjin Medical University Cancer Institute and HospitalApril 20, 2023cell_type_annotationfoundation_modellanguage_model+4Single-cell foundation model pre-trained on 22 million transcriptomes, using rank-based gene encoding for clustering and trajectory inference.
Single-cell50OpennessSTU-Net
372159—Scalable and transferable U-Net family (14M–1.4B parameters) for 3D medical image segmentation, supervised-pretrained on TotalSegmentator.
Imaging82OpennessESM-GearNet
11555—Joint sequence-structure protein representation framework that fuses ESM-2 language model embeddings with GearNet geometric graph neural networks.
Protein30OpennessBiomedCLIP
128665878.5KBiomedical vision-language model trained contrastively on 15M PubMed Central figure-caption pairs for zero-shot classification, retrieval, and VQA.
Imaging61OpennessPTUnifier
7853—Chinese University of Hong Kong, Shenzhen +2 othersFebruary 17, 2023chest_x_rayfoundation_modelimage_text_retrieval+8Medical vision-language pretraining unifying fusion-encoder and dual-encoder designs, handling image-only, text-only, and paired inputs in one model.
PathologyLanguage model56OpennessSpecies-Aware DNA LM
29538.1KMasked DNA language model trained on over 800 vertebrate genomes and conditioned on species identity to learn conserved regulatory sequence features.
DNA & Gene76OpennessSpecies-Aware DNA Language Model
18538.1KMasked DNA language model trained on 800+ species with explicit species conditioning, separating conserved regulatory motifs from background bias.
DNA & Gene92OpennessAnkh
249733.2KParameter-efficient protein language model that matches larger models such as ESM-2 on protein prediction tasks using under 10% of the parameters.
Protein24OpennessNucleotide Transformer
9012262.1KDNA foundation models from 500M to 2.5B parameters, trained on 3,200+ human genomes and 850 species for variant effect prediction.
DNA & Gene29OpennessscMoFormer
2720—Transformer framework for single-cell multi-omics that predicts cross-modality relationships using heterogeneous graphs of cells, genes, and proteins.
Single-cellProtein59OpennessProtST
1051686Multi-modal protein language model trained on sequences paired with biomedical text, enabling zero-shot function prediction and text-based retrieval.
Protein89OpennessSpliceBERT
561—RNA language model pre-trained on 2M+ pre-mRNA sequences from 72 vertebrate species for splice-site prediction and variant effect analysis.
RNA77OpennessReprogBERT
2439—Antibody CDR design model that reprograms a frozen English BERT for sequence infilling, avoiding training a dedicated protein language model.
Protein56OpennessRoentGen
88146—Text-conditioned latent diffusion model that generates synthetic chest X-rays from free-form radiology prompts by adapting Stable Diffusion.
ImagingLanguage model20OpennessGenSLM
142142—Genome-scale language model trained on prokaryotic genes and SARS-CoV-2 genomes to model viral evolution and flag emerging variants of concern.
DNA & Gene56OpennessMoLFormer-XL
406595209.9KLarge-scale chemical language model trained on 1.1 billion SMILES strings using linear attention transformers for molecular property prediction.
Small molecule86OpennessMicrobial Gene NLP
2950—Word2vec-based language model trained on 360 million microbial genes that predicts gene function from genomic context without sequence homology.
DNA & Gene87Openness