wisdomik/QuiltNet-B-32

https://huggingface.co/wisdomik/QuiltNet-B-32
Staleby wisdomik9.1K29updated 2 years ago

QuiltNet-B-32 is a CLIP ViT-B/32 vision-language foundation model trained on the Quilt-1M dataset curated from representative histopathology videos. It can perform various vision-language processing (VLP) tasks such as cross-modal retrieval, image classification, and visual question answering.

Sourced from

  • HuggingFacewisdomik/QuiltNet-B-32

Related resources

Vision-language model for dermatology, pretrained with MAGEN (Multi-Agent data GENeration) and O-MAKE (Ontology-based Multi-Aspect Knowledge-Enhanced pretraining).

Active592 weeks ago

BioCLIP is a foundation model for the tree of life, built using CLIP architecture as a vision model for general organismal biology. It is trained on TreeOfLife-10M, our specially-created dataset covering over 450K taxa--the most biologically diverse ML-ready dataset available to date.

Idle29.1K6 months ago

BiomedCLIP is a biomedical vision-language foundation model that is pretrained on PMC-15M, a dataset of 15 million figure-caption pairs extracted from biomedical research articles in PubMed Central, using contrastive learning.

Idle452.2K1 year ago

PathForge is a modular benchmarking framework for multiple instance learning in computational pathology. It supports whole slide image feature extraction, HDF5 artifact generation, tile overviews, benchmarking, pipeline optimization, classification, regression, survival and retrieval tasks, and support for model inference and visualization.

Active22 weeks ago
Shell
MIT