All Competitors
Every biological foundation model, evaluated and ranked by the bio.rodeo team
Showing 1–24 of 37 filtered models
Uni-Hema
—1—Information Technology University of the Punjab +1 otherNovember 18, 2025classificationcnnfoundation_model+8Digital hematopathology foundation model unifying blood-cell detection, classification, segmentation, and visual question answering.
Pathology8OpennessMetaboFM
—2—Georgia Institute of TechnologyOctober 23, 2025classificationfoundation_modelmass_spectrometry_imaging+6Vision Transformer foundation model for spatial metabolomics, pretrained on ~4,000 curated METASPACE mass spectrometry imaging datasets.
MetabolomicsSpatial omicsImaging10OpennessMIMO
1227—Medical vision-language model that takes visual prompts on an image and returns answers grounded in pixel-level segmentation masks.
ImagingLanguage model11OpennessSigPhi-Med
594Chongqing University of TechnologyJuly 1, 2025histologyinstruction_tuningmedical_image_understanding+5Biomedical vision-language assistant for medical visual question answering, pairing Phi-2 with a vision encoder in a 4.2B-parameter model.
ImagingLanguage model15OpennessEyeCLIP
8765—The Hong Kong Polytechnic University +5 othersJune 21, 2025clipcontrastive_learningcross_modal_retrieval+11CLIP-based vision-language foundation model for eye imaging, enabling zero-shot disease detection and cross-modal retrieval across 11 modalities.
ImagingLanguage model15OpennessUniBiomed
7110119Hong Kong University of Science and Technology +2 othersApril 30, 2025foundation_modelhistologymultimodal+6Universal foundation model that jointly generates diagnostic text and segments the corresponding targets across ten biomedical imaging modalities.
ImagingLanguage model64OpennessGMAI-VL-R1
1929—General medical vision-language model trained with reinforcement learning to reason step by step over medical images for diagnosis and visual QA.
ImagingLanguage model17OpennessMedVLM-R1
321911.2K2B-parameter medical vision-language model that uses reinforcement learning to show interpretable reasoning for radiology visual question answering.
ImagingLanguage model83OpennessHealthGPT
1.6K11538Zhejiang University +4 othersFebruary 14, 2025histologyimage_reconstructionmedical_image_generation+7Medical vision-language model that unifies image comprehension and generation in one autoregressive transformer via heterogeneous LoRA adapters.
PathologyImaging68OpennessEndoChat
514423Chinese University of Hong Kong +5 othersJanuary 20, 2025endoscopygrounded_dialogueinstruction_tuning+6Grounded multimodal language model for endoscopic surgery, supporting visual dialogue, region-based question answering, and bounding-box grounding.
ImagingLanguage model22OpennessMUSK
240283—Vision-language foundation model for precision oncology, pretrained on 50M pathology images and 1B text tokens via unified masked modeling.
PathologyLanguage model12OpennessBiMediX2
742120Mohamed bin Zayed University of Artificial IntelligenceDecember 10, 2024histologyinstruction_tuninglanguage_model+7Bilingual Arabic-English medical multimodal model built on Llama 3.1 for radiology, CT, and histology image understanding and question answering.
Language modelImagingPathology11OpennessMedRegA
462913Hong Kong University of Science and Technology +1 otherOctober 24, 2024histologyimage_classificationinstruction_tuning+8Region-aware bilingual medical multimodal LLM that handles image- and region-level vision-language tasks across eight imaging modalities.
PathologyLanguage model65OpennessBiomedGPT
7084062Open-source, lightweight generalist vision-language foundation model for diverse biomedical imaging and text tasks.
Language modelImagingPathology33OpennessLLaVA-Tri
412944Medical vision-language model trained on the MedTrinity-25M dataset, answering questions and generating text about radiology and histology images.
Language modelPathology30OpennessPathChat
—478—Multimodal vision-language copilot for pathology that answers open-ended questions about histology images and reasons about differential diagnoses.
PathologyLanguage model35OpennessHuatuoGPT-Vision
3992042.8KShenzhen Research Institute of Big Data +1 otherJune 27, 2024histologyinstruction_tuningmedical_image_understanding+5Open medical multimodal LLMs (7B and 34B) for visual question answering over radiology, pathology, and endoscopy images, trained on PubMedVision.
PathologyLanguage model52OpennessMAIRA-2
—1534.4KMicrosoft Research multimodal LLM for grounded chest X-ray report generation, localizing each described finding with bounding boxes on the image.
ImagingLanguage model35OpennessEyeFound
—43—The Hong Kong Polytechnic University +3 othersMay 18, 2024disease_diagnosisfoundation_modelmasked_autoencoder+7Ophthalmic imaging foundation model pretrained on 2.78M images across 11 modalities for diagnosis, prognosis, and visual question answering.
Imaging4OpennessMedDr
1002846Hong Kong University of Science and TechnologyApril 23, 2024foundation_modelhistologyinstruction_tuning+7Generalist medical vision-language foundation model with 40B parameters, spanning radiology, pathology, dermatology, retinography, and endoscopy.
ImagingLanguage model69OpennessMed-MoE
15889—Lightweight mixture-of-experts medical vision-language model routing visual question answering and image classification to domain-specific experts.
ImagingLanguage modelPathology81OpennessM3D
454172924Multimodal large language model for 3D medical imaging that handles report generation, visual question answering, and segmentation on CT volumes.
ImagingLanguage model77Openness