All Competitors

Every biological foundation model, evaluated and ranked by the bio.rodeo team

Showing 193209 of 209 filtered models

  • Med-Flamingo

    452618
    Stanford University +2 othersJuly 27, 2023few_shotin_context_learningmedical_imaging+6

    Multimodal medical vision-language model for few-shot visual question answering, learning new imaging tasks from in-context examples at inference.

    PathologyLanguage model
    18Openness
  • Med-PaLM M

    552
    Google Research +1 otherJuly 26, 2023generativegenomicshistology+9

    Google's generalist multimodal biomedical AI that encodes clinical text, medical images, and genomics with a single set of weights across 14 tasks.

    ImagingLanguage model
    25Openness
  • LLaVA-Med

    2.2K1.9K12.3K
    Microsoft ResearchJune 1, 2023histologyimage_captioninginstruction_tuning+7

    Biomedical vision-language assistant for question answering on radiology and pathology images, adapted from LLaVA on PubMed Central captions.

    PathologyLanguage model
    28Openness
  • PathAsst

    136104
    Westlake University +3 othersMay 24, 2023cytologyfoundation_modelhistology+7

    Multimodal pathology assistant that answers questions about histology and cytology images, pairing the PathCLIP vision encoder with a Vicuna-13B LLM.

    PathologyLanguage model
    17Openness
  • MedBLIP

    5789
    Shanghai Jiao Tong UniversityMay 18, 2023brain_mricomputer_aided_diagnosisimage_classification+7

    Vision-language framework for 3D medical image diagnosis and visual question answering, bridging frozen image encoders and LLMs, shown on brain MRI.

    ImagingLanguage model
    35Openness
  • MedVInT

    236367
    Shanghai Jiao Tong University +1 otherMay 17, 2023generativehistologymedical_image_understanding+6

    Generative medical visual question answering model that pairs a vision encoder with a language model, trained on the 227k-pair PMC-VQA dataset.

    PathologyLanguage model
    83Openness
  • PMC-CLIP

    241
    Shanghai Jiao Tong UniversityMarch 13, 2023cnncontrastive_learninghistology+7

    Biomedical vision-language model trained contrastively on 1.6M figure-caption pairs mined from PubMed Central open-access articles.

    PathologyImaging
    63Openness
  • BiomedCLIP

    128665878.5K
    Microsoft ResearchMarch 1, 2023contrastive_learningfoundation_modelimage_analysis+4

    Biomedical vision-language model trained contrastively on 15M PubMed Central figure-caption pairs for zero-shot classification, retrieval, and VQA.

    Imaging
    61Openness
  • PTUnifier

    7853
    Chinese University of Hong Kong, Shenzhen +2 othersFebruary 17, 2023chest_x_rayfoundation_modelimage_text_retrieval+8

    Medical vision-language pretraining unifying fusion-encoder and dual-encoder designs, handling image-only, text-only, and paired inputs in one model.

    PathologyLanguage model
    56Openness
  • ProtST

    1051686
    DeepGraphLearningJanuary 1, 2023contrastive_learningfoundation_modelmultimodal

    Multi-modal protein language model trained on sequences paired with biomedical text, enabling zero-shot function prediction and text-based retrieval.

    Protein
    89Openness
  • RoentGen

    88146
    Stanford UniversityNovember 23, 2022chest_radiographydata_augmentationfoundation_model+5

    Text-conditioned latent diffusion model that generates synthetic chest X-rays from free-form radiology prompts by adapting Stable Diffusion.

    ImagingLanguage model
    20Openness
  • Galactica

    2.7K1.1K623
    Meta AINovember 16, 2022foundation_modellanguage_modelmultimodal

    Scientific large language model trained on 48 million papers, textbooks, and reference works to store, combine, and reason about scientific knowledge.

    Language model
    46Openness
  • CheXzero

    234527
    Stanford UniversitySeptember 15, 2022chest_radiographycontrastive_learningimage_classification+7

    Self-supervised vision-language model for zero-shot detection of chest X-ray pathologies, trained on image-report pairs without explicit labels.

    ImagingPathology
    70Openness
  • M3AE

    134192
    Shenzhen Research Institute of Big Data +2 othersSeptember 15, 2022autoencoderimage_text_retrievalmultimodal+5

    Self-supervised medical vision-and-language pretraining via multi-modal masked autoencoders that reconstruct masked image patches and text tokens.

    PathologyLanguage model
    29Openness
  • Shenzhen Research Institute of Big Data +2 othersSeptember 15, 2022chest_x_rayfoundation_modelimage_text_retrieval+7

    Medical vision-language pretraining framework that injects structured medical knowledge into radiology image-text learning for VQA and retrieval.

    ImagingLanguage model
    29Openness
  • PubMedCLIP

    1833116.7K
    Hasso Plattner InstituteDecember 27, 2021cnncontrastive_learninghistology+8

    Medical-domain CLIP fine-tuned on radiology image-caption pairs from ROCO, serving as a drop-in visual encoder for medical visual question answering.

    PathologyLanguage model
    75Openness
  • GeneBERT

    27
    Carnegie Mellon UniversityOctober 11, 2021bertchromatinfoundation_model+6

    Multi-modal self-supervised transformer for regulatory genomics, pre-trained on DNA sequence together with transcription factor binding matrices.

    DNA & Gene
    18Openness