All Competitors
Every biological foundation model, evaluated and ranked by the bio.rodeo team
Showing 73–96 of 98 filtered models
MedDr
983659Hong Kong University of Science and TechnologyApril 23, 2024foundation_modelhistologyinstruction_tuning+7Generalist medical vision-language foundation model with 40B parameters, spanning radiology, pathology, dermatology, retinography, and endoscopy.
ImagingLanguage model69OpennessMed-MoE
158549—Lightweight mixture-of-experts medical vision-language model routing visual question answering and image classification to domain-specific experts.
ImagingLanguage modelPathology81OpennessM3D
45138861Multimodal large language model for 3D medical imaging that handles report generation, visual question answering, and segmentation on CT volumes.
ImagingLanguage model77OpennessCheXagent
229751KInstruction-tuned vision-language foundation model for chest X-ray interpretation, with 8 billion parameters spanning eight clinical task types.
ImagingLanguage model32OpennessT3D
—120—Vision-language pretraining for 3D CT volumes, aligning scans with their radiology reports for zero-shot classification, retrieval, and segmentation.
ImagingLanguage model12OpennessMAIRA-1
—16—Radiology-specific multimodal LLM that generates the findings section of a chest X-ray report from a frontal image, pairing RAD-DINO with Vicuna-7B.
ImagingLanguage model6OpennessQilin-Med-VL
65389Chinese medical vision-language model pairing a Vision Transformer with an LLM to caption medical images and answer clinical questions in Chinese.
Language modelPathology44OpennessCXR-LLaVA
5439187Seoul National University +1 otherOctober 22, 2023abnormality_classificationchest_x_rayinstruction_tuning+6Chest X-ray vision-language model that generates free-text radiology reports, pairing a CXR-specific image encoder with a 7B LLaMA-2 language model.
ImagingLanguage model27OpennessBioT5
127—204Encoder-decoder framework unifying molecules, proteins, and natural language with SELFIES notation for cross-modal drug discovery tasks.
Language modelSmall moleculeProtein74OpennessRadFM
559929—Radiology foundation model that reads interleaved 2D and 3D scans with text for diagnosis, visual question answering, and report generation.
ImagingLanguage model84OpennessDARWIN Series
24953—Open large language models for natural science, fine-tuned on physics, chemistry, and materials science literature with automated instruction tuning.
Language model24OpennessMed-Flamingo
451405—Multimodal medical vision-language model for few-shot visual question answering, learning new imaging tasks from in-context examples at inference.
PathologyLanguage model18OpennessMed-PaLM M
—138—Google's generalist multimodal biomedical AI that encodes clinical text, medical images, and genomics with a single set of weights across 14 tasks.
ImagingLanguage model25OpennessLLaVA-Med
2.2K7311.9KBiomedical vision-language assistant for question answering on radiology and pathology images, adapted from LLaVA on PubMed Central captions.
PathologyLanguage model28OpennessPathAsst
136472—Multimodal pathology assistant that answers questions about histology and cytology images, pairing the PathCLIP vision encoder with a Vicuna-13B LLM.
PathologyLanguage model17OpennessMedBLIP
57613—Vision-language framework for 3D medical image diagnosis and visual question answering, bridging frozen image encoders and LLMs, shown on brain MRI.
ImagingLanguage model35OpennessMedVInT
236190—Generative medical visual question answering model that pairs a vision encoder with a language model, trained on the 227k-pair PMC-VQA dataset.
PathologyLanguage model83OpennessPTUnifier
78310—Chinese University of Hong Kong, Shenzhen +2 othersFebruary 17, 2023chest_x_rayfoundation_modelimage_text_retrieval+8Medical vision-language pretraining unifying fusion-encoder and dual-encoder designs, handling image-only, text-only, and paired inputs in one model.
PathologyLanguage model56OpennessRoentGen
88——Text-conditioned latent diffusion model that generates synthetic chest X-rays from free-form radiology prompts by adapting Stable Diffusion.
ImagingLanguage model20OpennessBioGPT
4.5K1.5K97.1KGenerative transformer pretrained on PubMed abstracts for biomedical text generation and mining, including relation extraction and question answering.
Language model66Openness- Shenzhen Research Institute of Big Data +2 othersSeptember 15, 2022chest_x_rayfoundation_modelimage_text_retrieval+7
Medical vision-language pretraining framework that injects structured medical knowledge into radiology image-text learning for VQA and retrieval.
ImagingLanguage model29Openness M3AE
134171—Shenzhen Research Institute of Big Data +2 othersSeptember 15, 2022autoencoderimage_text_retrievalmultimodal+5Self-supervised medical vision-and-language pretraining via multi-modal masked autoencoders that reconstruct masked image patches and text tokens.
PathologyLanguage model29Openness