A comprehensive research university in Guangzhou, China, working across the humanities, natural sciences, engineering, and medicine.
Hyperbolic protein language model for alignment-free phylogenetic inference, turning ESM2-650M embeddings into distance matrices for tree placement.
Pengcheng Laboratory / Sun Yat-sen University / Tsinghua University
Released March 13, 2026
Molecular reasoning model built on DeepSeek-7B, using chain-of-thought and reinforcement learning for property prediction, generation, and reactions.
Generative virtual-cell model predicting whole-transcriptome responses to unseen compounds and genetic perturbations, from cell lines to organoids.
Designs synthesizable PROTAC degraders from reaction templates and purchasable building blocks, with reinforcement learning tuning the generator.
Hong Kong University of Science and Technology / Sun Yat-sen University / Southern Medical University / Chinese University of Hong Kong
Released November 1, 2025
Histopathology foundation model extracting general-purpose features from H&E patches by distilling the UNI, Phikon, and CONCH pathology encoders.
Tencent AI for Life Science Lab / Chinese Academy of Sciences / University of Chinese Academy of Sciences / Fudan University / Sun Yat-sen University / Shanghai Jiao Tong University / Guangzhou Medical University
Released September 16, 2025
Generative antibody and nanobody design model that co-designs CDR sequences and antigen-bound structures for de novo design and affinity maturation.
Sun Yat-sen University / University of Science and Technology of China
Released September 11, 2025
Discrete graph diffusion model for multi-property molecular generation, composing per-property score guidance over arbitrary condition subsets.
Blind protein-ligand docking model adding Ollivier-Ricci curvature descriptors and degree-aware message passing, predicting poses in 0.09 seconds.
Hong Kong University of Science and Technology / Southern Medical University / Sun Yat-sen University / Chinese University of Hong Kong / The University of Hong Kong / Chinese PLA General Hospital
Released August 10, 2025
Multi-sequence MRI foundation model pretrained on 336,476 volumetric scans, ranking first on 41 of 44 downstream clinical benchmarks.
Shanghai Jiao Tong University / Lingang Laboratory / Sun Yat-sen University / Fudan University / Shanghai AI Laboratory / Shanghai Institute of Materia Medica / MIT / Ningxia Medical University
Released August 4, 2025
Structure-based virtual screening model that jointly predicts protein-ligand complex structures and binding fitness from sequence and SMILES.
Shanghai AI Laboratory / Sun Yat-sen University / Chinese University of Hong Kong / Karlsruhe Institute of Technology
Released June 29, 2025
EEG foundation model with cross-scale spatiotemporal tokenization and sparse structured attention, evaluated on 11 decoding tasks across 16 datasets.
Hong Kong University of Science and Technology / Sun Yat-sen University / Macau University of Science and Technology / Jinan University / Zhejiang University School of Medicine / Chinese University of Hong Kong / Harvard University
Released June 24, 2025
Genome-anchored histopathology embeddings that predict molecular biomarkers, subtypes, and survival from whole-slide images alone at inference.
Multimodal viral foundation model over nucleotide and protein sequence, built for virus discovery, function annotation, and antibody design.
Pocket-conditioned 3D diffusion model for scaffold decoration, guided by evolutionary residue conservation and a protein-ligand interaction prior.
University of Macau / Beijing University of Technology / University of Florida / Macao Polytechnic University / Sun Yat-sen University / Wake Forest University School of Medicine
Released February 26, 2025
Cell Painting image encoder that turns whole-slide multi-channel microscopy into morphological profiles in one pass, with no cell segmentation step.
Hong Kong University of Science and Technology / Chinese University of Hong Kong / Tencent AI Lab / Sun Yat-sen University / Peking University Shenzhen Hospital / Shenzhen Institutes of Advanced Technology, CAS / Harvard University
Released February 12, 2025
Cervical cytology screening system pretrained on 127,471 whole-slide images from 48 centers, with test-time adaptation for new clinical sites.
Shanghai Jiao Tong University / Hong Kong University of Science and Technology / Hainan University / Sun Yat-sen University / McGill University / Mila / MIT
Released December 16, 2024
Geometric foundation model matching enzymes to the reactions they catalyze, trained on 1.5 million structure-informed enzyme-reaction pairs.
Peking University / Macau University of Science and Technology / Wenzhou Medical University / Sun Yat-sen University / University of Oxford / New York University
Released December 11, 2024
Text-guided medical image synthesis across OCT, fundus, X-ray, CT and MRI. Synthetic data lifts downstream clinical tasks by 12-17%.
Hong Kong University of Science and Technology / Sun Yat-sen University
Released October 24, 2024
Region-aware bilingual medical multimodal LLM that handles image- and region-level vision-language tasks across eight imaging modalities.
Single-cell foundation model with 800M parameters trained on ~100 million human cells, for annotation, perturbation prediction, and gene analysis.
The Hong Kong Polytechnic University / Sun Yat-sen University / National University of Singapore / EPFL
Released May 18, 2024
Ophthalmic imaging foundation model pretrained on 2.78M images across 11 modalities for diagnosis, prognosis, and visual question answering.
Alibaba Cloud / Sun Yat-sen University / University of Sydney / Fudan University / Zhejiang University / Chinese Academy of Medical Sciences / Peking Union Medical College / City University of Hong Kong
Released May 10, 2024
Unified DNA, RNA, and protein foundation model with 1.8B parameters, pretrained across 169,861 species to learn the central dogma from sequence.
Chinese University of Hong Kong, Shenzhen / Sun Yat-sen University / Shenzhen Research Institute of Big Data
Released February 17, 2023
Medical vision-language pretraining unifying fusion-encoder and dual-encoder designs, handling image-only, text-only, and paired inputs in one model.
Shenzhen Research Institute of Big Data / Chinese University of Hong Kong, Shenzhen / Sun Yat-sen University
Released September 15, 2022
Self-supervised medical vision-and-language pretraining via multi-modal masked autoencoders that reconstruct masked image patches and text tokens.
Shenzhen Research Institute of Big Data / Chinese University of Hong Kong, Shenzhen / Sun Yat-sen University
Released September 15, 2022
Medical vision-language pretraining framework that injects structured medical knowledge into radiology image-text learning for VQA and retrieval.