Find open-source science resources

A directory of tools, AI models, datasets, and research resources for biotech, bioinformatics, and other scientific fields. Aggregated from curated GitHub awesome-lists, HuggingFace, bio.tools, Bioconductor, and more.

2,558 of 7,078 resources

Showing 2,251–2,300

TreeHub is a taxonomical database for trees. TreeHub contains extracted phylogenetic data and integrates relevant species information for each identifier.

The tRNA Gene DataBase Curated by Experts "tRNADB-CE" was constructed by analyzing 927 complete and 1301 draft genomes of Bacteria and Archaea, 171 complete virus genomes, 121 complete chloroplast genomes, 12 complete eukaryote (Plant and Fungi) genomes and approximately 230 million DNA sequence entries that originated from environmental metagenomic clones.

Centralized repository and distribution site for variety of Tetrahymena strains and species. Maintains diverse array of wild type, mutant, and genetically engineered strains of T. thermophila, the most commonly used laboratory species, and variety of other species derived from both laboratory maintained stocks and wild isolates. All stocks are stored in liquid nitrogen to maintain genetic integrity and prevent senescence. In addition to providing worldwide access to strains currently in collection, TSC continually upgrades collection by accepting deposition of newly developed laboratory strains and well characterized wild isolates collected from clearly defined natural sites. [from RRID]

Identifiers correspond to curated associations between diseases and tRNA-derived small RNAs (tsRNAs) in the tsRNADisease database.

An identifier for institutions in the United Kingdom, used in GRID and ROR.

The UCSC Genome Browser is an on-line, and downloadable, genome browser hosted by the University of California, Santa Cruz (UCSC).[2][3][4] It is an interactive website offering access to genome sequence data from a variety of vertebrate and invertebrate species and major model organisms, integrated with a large collection of aligned annotations.

A category for fields appearing in the data of the UK BioBank

A field appearing in the data of the UK BioBank, like 'time spent outdoors in the summer'

identifier for an educational organization issued by the UK Register of Learning Providers

An identifier for an atom; the smallest unit of naming in a source, viz, a specific string with specific code values and identifiers from a specific source. As such, they can be thought of as representing a single meaning with a source Atoms are the units of terminology that come from sources and form the building blocks of the concepts in the Metathesaurus.

The UNESCO Thesaurus is a controlled and structured list of terms used in subject analysis and retrieval of documents and publications in the fields of education, culture, natural sciences, social and human sciences, communication and information. Continuously enriched and updated, its multidisciplinary terminology reflects the evolution of UNESCO's programmes and activities. [from homepage]

UniCarb-DB stores structural and mass spectrometry (MS) glycomic data, and has now grown to be one of the largest experimental glycomic MS databases. Identifiers correspond to individual glycan structures and corresponding MS spectra.

Association-Rule-Based Annotator (ARBA), a multiclass, self-training annotation system for automatic classification and annotation of UniProtKB proteins. This replaces the previous rule-based SAAS system.

The human diseases in which proteins are involved are described in UniProtKB entries with a controlled vocabulary.

identifier for a scientific journal, in the UniProt database

UniProtKB entries are tagged with keywords that can be used to retrieve particular subsets of entries.

The subcellular locations in which a protein is found are described in UniProtKB entries with a controlled vocabulary, which includes also membrane topology and orientation terms.

The UniProt Knowledgebase (UniProtKB) is a comprehensive resource for protein sequence and functional information with extensive cross-references to more than 120 external databases. Besides amino acid sequence and a description, it also provides taxonomic data and citation information. This entry represents UniProt mnemonics which combine an alphanumeric representation the protein name and a species identification code representing the biological source of the protein.

UniProt provides proteome sets of proteins whose genomes have been completely sequenced.

The post-translational modifications used in the UniProt knowledgebase (Swiss-Prot and TrEMBL). The definition of the post-translational modifications usage as well as other information is provided in the following format

The cross-references section of UniProtKB entries displays explicit and implicit links to databases such as nucleotide sequence databases, model organism databases and genomics and proteomics resources.