Find open-source science resources
A directory of tools, AI models, datasets, and research resources for biotech, bioinformatics, and other scientific fields. Aggregated from curated GitHub awesome-lists, HuggingFace, bio.tools, Bioconductor, and more.
Filters
Health
Domain(1)
Language(1)
License
Source
Type
149 of 6,573 resources
Showing 101–149
🚀 Meerkat-8B is a new instruction-tuned medical AI system of the Meerkat model family. The model was based on the Meta's Llama-3-8B-Instruct model and fine-tuned using our new synthetic dataset consisting of high-quality chain-of-thought reasoning paths sourced from 18 medical textbooks, along…
dmis-lab/meerkat-7b-v1.0
by dmis-lab🚀 Meerkat-7B-v1.0 is an instruction-tuned medical AI system that surpasses the passing threshold of 60% for the United States Medical Licensing Examination (USMLE) for the first time among all 7B-parameter models. The model was trained using our new synthetic dataset consisting of high-quality…
This is https://huggingface.co/kingabzpro/Qwen-3-32B-Medical-Reasoning applied to https://huggingface.co/Qwen/Qwen3-32B Original model card created by @kingabzpro
ibm-research/GP-MoLFormer-Uniq
by ibm-researchGP-MoLFormer is a class of models pretrained on SMILES string representations of 0.65-1.1B molecules from ZINC and PubChem. This repository is for the model pretrained on all the unique molecules from both datasets.
QIAIUNCC/EYE-Llama_gqa
by QIAIUNCC## Model Description EYE-Llama_gqa is a large language model specifically designed for ophthalmic question-answering (QA). It is built upon the Llama 2 architecture and fine-tuned on a the EYE-lit and EYE-QA+ dataset.
This project fine-tunes the meta-llama/Llama-4-Scout-17B-16E-Instruct model using a medical reasoning dataset (FreedomIntelligence/medical-o1-reasoning-SFT) with 4-bit quantization for memory-efficient training.
WaltonFuture/Diabetica-7B
by WaltonFutureDiabetica: Adapting Large Language Model to Enhance Multiple Medical Tasks in Diabetes Care and Management
Henrychur/MMedS-Llama-3-8B
by Henrychur# MMedS-Llama3 💻Github Repo 🖨️arXiv Paper
### Welcome to Nidum! At Nidum, we believe in pushing the boundaries of innovation by providing advanced and unrestricted AI models for every application. Dive into our world of possibilities and experience the freedom of Nidum-Llama-3.2-3B-Uncensored, tailored to meet diverse needs with…
togethercomputer/evo-1-131k-base
by togethercomputerWe identified and fixed an issue related to a wrong permutation of some projections, which affects generation quality. To use the new model revision, please load as follows:
peteparker456/medical_diagnosis_llama2
by peteparker456This model aims to be a base template for new models. It has been generated using this raw template.
> [!IMPORTANT] > Better using New version of ChemLLM! > AI4Chem/ChemLLM-7B-Chat-1.5-DPO or AI4Chem/ChemLLM-7B-Chat-1.5-SFT
yerevann/chemma-2b
by yerevannChemma-2B is a continually pretrained gemma-2b model for organic molecules. It is pretrained on 40B tokens covering 110M+ molecules from PubChem as well as their chemical properties (molecular weight, synthetic accessibility score, drug-likeness etc.) and similarities (Tanimoto distance between…
RaphaelMourad/Mistral-DNA-v1-138M-bacteria
by RaphaelMouradThe Mistral-DNA-v1-138M-bacteria Large Language Model (LLM) is a pretrained generative DNA text model with 17.31M parameters x 8 experts = 138.5M parameters. It is derived from Mistral-7B-v0.1 model, which was simplified for DNA: the number of layers and the hidden size were reduced.
tmberooney/medllama-merged
by tmberooneyModel Card for "medllama" ---------------------------
# Medical-Llama3-v2 Fine-Tuned Llama3 for Medical Q&A This repository provides a fine-tuned version of the powerful Llama3 8B model, specifically designed to answer medical questions in an informative way. It leverages the rich knowledge contained in the AI Medical Chatbot dataset…
This is an official model checkpoint for Asclepius-Mistral-7B-v0.3 (arxiv). This model is an enhanced version of Asclepius-7B, by replacing the base model with Mistral-7B-v0.3 and increasing the max sequence length to 8192.
This is an official model checkpoint for Asclepius-Llama3-8B (arxiv). This model is an enhanced version of Asclepius-7B, by replacing the base model with Llama-3 and increasing the max sequence length to 8192.
Henrychur/MMed-Llama-3-8B
by Henrychur# MMedLM 💻Github Repo 🖨️arXiv Paper
johnsnowlabs/JSL-MedLlama-3-8B-v2.0
by johnsnowlabs# JSL-MedLlama-3-8B-v2.0
Reference: R. Luu and M.J. Buehler, "BioinspiredLLM: Conversational Large Language Model for the Mechanics of Biological and Bio-Inspired Materials," Adv. Science, 2023, DOI: https://doi.org/10.1002/advs.202306724
clinicalnlplab/finetuned-Llama-2-13b-hf-PubmedQA
by clinicalnlplabMedical mT5: An Open-Source Multilingual Text-to-Text LLM for the Medical Domain
# ChemLLM-7B-Chat-1.5-DPO: LLM for Chemistry and Molecule Science ChemLLM-7B-Chat-1.5-DPO, The First Open-source Large Language Model for Chemistry and Molecule Science, Build based on InternLM-2 with ❤
Using llama.cpp release b2440 for quantization.
Goekdeniz-Guelmez/Hyperion-2.0-Mistral-7B-GGUF
by Goekdeniz-Guelmez!image/png
Using llama.cpp commit fa97464 for quantization.
!image/png
# Mr-Grammatology-clinical-problems-Mistral-7B-0.5 !image/png
This is a merge of pre-trained language models created using mergekit.
BioMistral/BioMistral-7B-SLERP
by BioMistralThis is a merge of pre-trained language models created using mergekit.
BioMistral/BioMistral-7B-DARE
by BioMistralThis is a merge of pre-trained language models created using mergekit.
knowledgator/SMILES2IUPAC-canonical-base
by knowledgatorSMILES2IUPAC-canonical-base was designed to accurately translate SMILES chemical names to IUPAC standards.
muzammil-eds/tinyllama-2.5T-Clinical-v2
by muzammil-eds# TinyLlama-1.1B
starmpcc/Asclepius-13B
by starmpccThis is official model checkpoint for Asclepius-13B (arxiv). This model is the first publicly shareable clinical LLM, trained with synthetic data.
TachyHealth/Thealth_Mixtral-8x7B
by TachyHealthepfl-llm/meditron-70b
by epfl-llmDetails coming soon
# Meditron 70B - GGUF - Model creator: EPFL LLM Team - Original model: Meditron 70B
MentaLLaMA-chat-7B is part of the MentaLLaMA project, the first open-source large language model (LLM) series for interpretable mental health analysis with instruction-following capability. This model is finetuned based on the Meta LLaMA2-chat-7B foundation model and the full IMHI instruction…
This model is a fine-tuned model based on the Llama 2_7b architecture. It has been specifically trained on a dataset comprising USMLE (United States Medical Licensing Examination) questions and answers, as well as conversations between doctors and patients.
microsoft/BioGPT-Large
by microsoftPre-trained language models have attracted increasing attention in the biomedical domain, inspired by their great success in the general natural language domain. Among the two main branches of pre-trained language models in the general language domain, i.e.
# ChemGPT 1.2B ChemGPT is based on the GPT-Neo model and was introduced in the paper Neural Scaling of Deep Chemical Models.
# ChemGPT 19M ChemGPT is based on the GPT-Neo model and was introduced in the paper Neural Scaling of Deep Chemical Models.
# ChemGPT 4.7M ChemGPT is based on the GPT-Neo model and was introduced in the paper Neural Scaling of Deep Chemical Models.