All Competitors
Every biological foundation model, evaluated and ranked by the bio.rodeo team
Showing 25–48 of 70 filtered models
gCIS
109—CT segmentation foundation model that uses task prompts to segment 83 anatomical structures and lesions across whole-body scans in a single network.
Imaging14OpennessMaCo
1276—Chest X-ray foundation model that pairs masked image modeling with image-report contrastive alignment for zero-shot diagnosis and phrase grounding.
Imaging74OpennessBiomedGPT
7084062Open-source, lightweight generalist vision-language foundation model for diverse biomedical imaging and text tasks.
Language modelImagingPathology33OpennessLLaVA-Tri
412944Medical vision-language model trained on the MedTrinity-25M dataset, answering questions and generating text about radiology and histology images.
Language modelPathology30OpennessCXR Foundation
19982199Chest X-ray embedding model built on ELIXR, producing image and image-text embeddings for data-efficient and zero-shot radiograph classification.
ImagingMedical SAM 2
933257—SAM2-based foundation model that segments 2D and 3D medical images by treating volumes and image sets as video object tracking.
Imaging74OpennessScribblePrompt
22066—Interactive foundation model for biomedical image segmentation, prompted with scribbles, clicks, and bounding boxes to segment unseen structures.
Imaging63OpennessHuatuoGPT-Vision
3992042.8KShenzhen Research Institute of Big Data +1 otherJune 27, 2024histologyinstruction_tuningmedical_image_understanding+5Open medical multimodal LLMs (7B and 34B) for visual question answering over radiology, pathology, and endoscopy images, trained on PubMedVision.
PathologyLanguage model52OpennessMAIRA-2
—1534.4KMicrosoft Research multimodal LLM for grounded chest X-ray report generation, localizing each described finding with bounding boxes on the image.
ImagingLanguage model35OpennessMed-Gemini
—410—Family of medical multimodal models built on Gemini, adding uncertainty-guided web search, custom modality encoders, and long-context EHR reasoning.
Language modelImaging8OpennessMedDr
1002846Hong Kong University of Science and TechnologyApril 23, 2024foundation_modelhistologyinstruction_tuning+7Generalist medical vision-language foundation model with 40B parameters, spanning radiology, pathology, dermatology, retinography, and endoscopy.
ImagingLanguage model69OpennessMed-MoE
15889—Lightweight mixture-of-experts medical vision-language model routing visual question answering and image classification to domain-specific experts.
ImagingLanguage modelPathology81OpennessM3D
454172924Multimodal large language model for 3D medical imaging that handles report generation, visual question answering, and segmentation on CT volumes.
ImagingLanguage model77OpennessVoCo
230113—Hong Kong University of Science and TechnologyFebruary 27, 2024contrastive_learningctfoundation_model+5Self-supervised pretraining framework for 3D medical image encoders that learns anatomy by predicting where a sub-volume sits within a CT scan.
Imaging69OpennessCheXagent
23075981Instruction-tuned vision-language foundation model for chest X-ray interpretation, with 8 billion parameters spanning eight clinical task types.
ImagingLanguage model32OpennessMedSAM
4.4K1.5K1.8KPromptable foundation model for universal medical image segmentation, fine-tuned from SAM on 1.57M image-mask pairs across 10 imaging modalities.
Imaging82OpennessT3D
—16—Vision-language pretraining for 3D CT volumes, aligning scans with their radiology reports for zero-shot classification, retrieval, and segmentation.
ImagingLanguage model12OpennessMAIRA-1
—92—Radiology-specific multimodal LLM that generates the findings section of a chest X-ray report from a frontal image, pairing RAD-DINO with Vicuna-7B.
ImagingLanguage model6OpennessSegVol
386122747Promptable 3D foundation model for volumetric CT segmentation, covering over 200 anatomical categories through point, box, and free-text prompts.
Imaging100OpennessQilin-Med-VL
65825Chinese medical vision-language model pairing a Vision Transformer with an LLM to caption medical images and answer clinical questions in Chinese.
Language modelPathology44OpennessSAM-Med3D
944182—Fully 3D promptable segmentation foundation model for volumetric CT and MR, encoding whole volumes so anatomy can be segmented from one prompt point.
Imaging96OpennessCXR-LLaVA
5439180Seoul National University +1 otherOctober 22, 2023abnormality_classificationchest_x_rayinstruction_tuning+6Chest X-ray vision-language model that generates free-text radiology reports, pairing a CXR-specific image encoder with a 7B LLaMA-2 language model.
ImagingLanguage model27OpennessCXR-CLIP
123138—Large-scale chest X-ray vision-language pretraining model that learns image-report alignment for zero-shot and few-shot radiograph classification.
Imaging18OpennessCLIP-Driven Universal Model
677356—Abdominal CT segmentation model driven by CLIP text embeddings, covering 25 organs and 6 tumor types with zero-shot extension to new categories.
Imaging26Openness