Find open-source science resources

A directory of tools, AI models, datasets, and research resources for biotech, bioinformatics, and other scientific fields. Aggregated from curated GitHub awesome-lists, HuggingFace, bio.tools, Bioconductor, and more.

2,419 of 6,573 resources

Showing 2,0512,100

The SciCrunch Registry holds metadata records that describe digital resources, e.g., software, databases, projects and also services. Most of these are produced as a result of government funding and are available to the scientific community. Resources are manually curated to make sure the information is accurate. We also use a web crawler to find literature mentions for the resources.

With the growing number of available genomes, the need for an environment to support effective comparative analysis increases. The original SEED Project was started in 2003 by the [Fellowship for Interpretation of Genomes (FIG)](http://thefig.info/) as a largely unfunded open source effort. Argonne National Laboratory and the University of Chicago joined the project, and now much of the activity occurs at those two institutions (as well as the University of Illinois at Urbana-Champaign, Hope college, San Diego State University, the Burnham Institute and a number of other institutions). The cooperative effort focuses on the development of the comparative genomics environment called the SEED and, more importantly, on the development of curated genomic data. This prefix provides identifiers for molecular roles that describe the function of one or more proteins in microbes and plants.

A vocabulary about species to support the environmental research community in Arizona and New Mexico

Identifier of an author or reviewer in the Semion database (now defunct)

The Scientific Event Ontology (SEO) is a reference ontology for modeling data about scientific events such as conferences, symposioums, and workshops.

SEON provides a well-grounded network of software engineering reference ontologies, and mechanisms for deriving and incorporating new integrated domain ontologies into the network. [from homepage]

The Structure-Function Linkage Database (SFLD) is a hierarchical classification of enzymes that relates specific sequence-structure features to specific chemical capabilities. It was developed by the Babbitt Laboratory in collaboration with the UCSF Resource for Biocomputing, Visualization, and Informatics. As of April 2019, the database is in static format, and will not be updated. [from https://interpro-documentation.readthedocs.io/en/latest/sfld.html, accessed March 18th, 2026]

A language for validating RDF graphs against a set of conditions

This shapes graph can be used to validate SHACL shapes graphs against a subset of the syntax rules.

Service Categories (KCDB). This service shows the related service categories for the Key Comparisons DataBase (KCDB) by quantity.

Sigma Aldrich is a life sciences supply vendor.

Identifiers for relationships between proteins and complexes, along with their type and provenance

SILVA is a resource of aligned ribosomal RNA (rRNA) gene sequences for Bacteria, Archaea, and Eukaryotes. It assigns internal, auto-incremented integer identifiers to taxa. However, these identifiers are not persistent nor unique - if the label for a taxon is changed, it assigned a new identifier and the old identifier is removed.

Extends the SIOC Core Ontology (Semantically-Interlinked Online Communities) by defining basic information on permissions and access rights.

Semantically-Interlinked Online Communities, or SIOC, is an attempt to link online community sites, to use Semantic Web technologies to describe the information that communities have about their structure and contents, and to find related information and new connections between content items and other community objects. SIOC is based around the use of machine-readable information provided by these sites.

Extends the SIOC Core Ontology (Semantically-Interlinked Online Communities) by defining basic information on community-related web services.

Extends the SIOC Core Ontology (Semantically-Interlinked Online Communities) by defining subclasses and subproperties of SIOC terms.

A modern method of records management and an automated cross-referenced subject index for accurate and comprehensive information retrieval developed by the US FDA's Bureau of Foods

SKIP is aiming to promote the exchange of information and joint research between researchers by aggregating various information of stem cells (iPS cells, iPS cells derived from patients, etc.) to stimulate research on disease and regenerative medicine.

SKOS is an area of work developing specifications and standards to support the use of knowledge organization systems (KOS) such as thesauri, classification schemes, subject heading lists and taxonomies within the framework of the Semantic Web