All Competitors

Every biological foundation model, evaluated and ranked by the bio.rodeo team

Showing 124 of 37 filtered models

  • Uni-Hema

    1
    Information Technology University of the Punjab +1 otherNovember 18, 2025classificationcnnfoundation_model+8

    Digital hematopathology foundation model unifying blood-cell detection, classification, segmentation, and visual question answering.

    Pathology
    8Openness
  • MetaboFM

    2
    Georgia Institute of TechnologyOctober 23, 2025classificationfoundation_modelmass_spectrometry_imaging+6

    Vision Transformer foundation model for spatial metabolomics, pretrained on ~4,000 curated METASPACE mass spectrometry imaging datasets.

    MetabolomicsSpatial omicsImaging
    10Openness
  • MIMO

    1227
    Peking University +1 otherOctober 11, 2025histologyinstruction_tuningmultimodal+6

    Medical vision-language model that takes visual prompts on an image and returns answers grounded in pixel-level segmentation masks.

    ImagingLanguage model
    11Openness
  • Chongqing University of TechnologyJuly 1, 2025histologyinstruction_tuningmedical_image_understanding+5

    Biomedical vision-language assistant for medical visual question answering, pairing Phi-2 with a vision encoder in a 4.2B-parameter model.

    ImagingLanguage model
    15Openness
  • EyeCLIP

    8765
    The Hong Kong Polytechnic University +5 othersJune 21, 2025clipcontrastive_learningcross_modal_retrieval+11

    CLIP-based vision-language foundation model for eye imaging, enabling zero-shot disease detection and cross-modal retrieval across 11 modalities.

    ImagingLanguage model
    15Openness
  • QoQ-Med

    5238465
    MITMay 31, 2025ecgfoundation_modelhistology+7

    Multimodal clinical foundation model reasoning jointly over 2D and 3D medical images, ECG time-series, and text reports across nine clinical domains.

    ImagingBiosignalsLanguage model
    74Openness
  • UniBiomed

    7110119
    Hong Kong University of Science and Technology +2 othersApril 30, 2025foundation_modelhistologymultimodal+6

    Universal foundation model that jointly generates diagnostic text and segments the corresponding targets across ten biomedical imaging modalities.

    ImagingLanguage model
    64Openness
  • GMAI-VL-R1

    1929
    Shanghai AI Laboratory +6 othersApril 2, 2025histologylanguage_modelmedical_image_diagnosis+7

    General medical vision-language model trained with reinforcement learning to reason step by step over medical images for diagnosis and visual QA.

    ImagingLanguage model
    17Openness
  • MedVLM-R1

    321911.2K
    Technical University of Munich +2 othersFebruary 26, 2025medical_reasoningmultimodalradiology+4

    2B-parameter medical vision-language model that uses reinforcement learning to show interpretable reasoning for radiology visual question answering.

    ImagingLanguage model
    83Openness
  • HealthGPT

    1.6K11538
    Zhejiang University +4 othersFebruary 14, 2025histologyimage_reconstructionmedical_image_generation+7

    Medical vision-language model that unifies image comprehension and generation in one autoregressive transformer via heterogeneous LoRA adapters.

    PathologyImaging
    68Openness
  • EndoChat

    514423
    Chinese University of Hong Kong +5 othersJanuary 20, 2025endoscopygrounded_dialogueinstruction_tuning+6

    Grounded multimodal language model for endoscopic surgery, supporting visual dialogue, region-based question answering, and bounding-box grounding.

    ImagingLanguage model
    22Openness
  • MUSK

    240283
    Stanford University +1 otherJanuary 8, 2025cross_modal_retrievalfoundation_modelhistology+9

    Vision-language foundation model for precision oncology, pretrained on 50M pathology images and 1B text tokens via unified masked modeling.

    PathologyLanguage model
    12Openness
  • MedPLIB

    1343910
    Baidu +3 othersDecember 12, 2024histologymedical_image_groundingmixture_of_experts+8

    Biomedical multimodal LLM that answers questions about medical images and returns pixel-level segmentation masks, using a mixture-of-experts design.

    ImagingLanguage model
    80Openness
  • BiMediX2

    742120
    Mohamed bin Zayed University of Artificial IntelligenceDecember 10, 2024histologyinstruction_tuninglanguage_model+7

    Bilingual Arabic-English medical multimodal model built on Llama 3.1 for radiology, CT, and histology image understanding and question answering.

    Language modelImagingPathology
    11Openness
  • MedRegA

    462913
    Hong Kong University of Science and Technology +1 otherOctober 24, 2024histologyimage_classificationinstruction_tuning+8

    Region-aware bilingual medical multimodal LLM that handles image- and region-level vision-language tasks across eight imaging modalities.

    PathologyLanguage model
    65Openness
  • BiomedGPT

    7084062
    Lehigh University +9 othersAugust 7, 2024foundation_modelhistologyimage_captioning+7

    Open-source, lightweight generalist vision-language foundation model for diverse biomedical imaging and text tasks.

    Language modelImagingPathology
    33Openness
  • LLaVA-Tri

    412944
    UC Santa Cruz +3 othersAugust 6, 2024foundation_modelhistologymultimodal+5

    Medical vision-language model trained on the MedTrinity-25M dataset, answering questions and generating text about radiology and histology images.

    Language modelPathology
    30Openness
  • PathChat

    478
    Mahmood Lab +4 othersJuly 10, 2024cancerdiagnosisfoundation_model+7

    Multimodal vision-language copilot for pathology that answers open-ended questions about histology images and reasons about differential diagnoses.

    PathologyLanguage model
    35Openness
  • Shenzhen Research Institute of Big Data +1 otherJune 27, 2024histologyinstruction_tuningmedical_image_understanding+5

    Open medical multimodal LLMs (7B and 34B) for visual question answering over radiology, pathology, and endoscopy images, trained on PubMedVision.

    PathologyLanguage model
    52Openness
  • MAIRA-2

    1534.4K
    Microsoft ResearchJune 6, 2024chest_x_rayinstruction_tuningmultimodal+6

    Microsoft Research multimodal LLM for grounded chest X-ray report generation, localizing each described finding with bounding boxes on the image.

    ImagingLanguage model
    35Openness
  • EyeFound

    43
    The Hong Kong Polytechnic University +3 othersMay 18, 2024disease_diagnosisfoundation_modelmasked_autoencoder+7

    Ophthalmic imaging foundation model pretrained on 2.78M images across 11 modalities for diagnosis, prognosis, and visual question answering.

    Imaging
    4Openness
  • MedDr

    1002846
    Hong Kong University of Science and TechnologyApril 23, 2024foundation_modelhistologyinstruction_tuning+7

    Generalist medical vision-language foundation model with 40B parameters, spanning radiology, pathology, dermatology, retinography, and endoscopy.

    ImagingLanguage model
    69Openness
  • Med-MoE

    15889
    Zhejiang University +2 othersApril 16, 2024histologyimage_classificationinstruction_tuning+5

    Lightweight mixture-of-experts medical vision-language model routing visual question answering and image classification to domain-specific experts.

    ImagingLanguage modelPathology
    81Openness
  • M3D

    454172924
    Beijing Academy of Artificial IntelligenceMarch 31, 2024ctimage_text_retrievalinstruction_tuning+9

    Multimodal large language model for 3D medical imaging that handles report generation, visual question answering, and segmentation on CT volumes.

    ImagingLanguage model
    77Openness