AmelieSchreiber/esm2_t12_35M_lora_binding_sites_v2_cp3

https://huggingface.co/AmelieSchreiber/esm2_t12_35M_lora_binding_sites_v2_cp3
Staleby AmelieSchreiber317updated 2 years ago
Python

This model may be overfit to some extent (see below). Try running this notebook on the datasets linked to in the notebook. See if you can figure out why the metrics differ so much on the datasets. Is it due to something like sequence similarity in the train/test split?

Sourced from

  • HuggingFaceAmelieSchreiber/esm2_t12_35M_lora_binding_sites_v2_cp3

Related resources

This model was finetuned on concatenated pairs of interacting proteins in much the same way as PepMLM. It is meant to generate interaction partners for proteins using the masked language modeling capabilities of ESM-2. The model is not well tested, so use with caution.

Stale82 years ago
Python

This set of model weights was released with the GitHub-compatible esm package format. The models here are kept for backwards compatibility, but we recommend you use the HuggingFace-compatible model weights at biohub/ESMC-6B (or biohub/ESMC-300M / biohub/ESMC-600M) instead.

Active2.3K2 months ago
Python

A frontier protein-language generative model — because proteins deserve better small talk.

Active164 months ago

Sparse Autoencoder (SAE) trained on residue-level embeddings from ESM-2 (650M, layer 33) for interpretability research on protein language models.

Active183 months ago

SERAPH is a deep learning model designed for 3-state (Q3) protein secondary structure prediction. It processes raw single amino acid sequences and predicts residue-level secondary structure states: Alpha Helix (H), Beta Sheet (E), or Coil/Loop (C).

Active03 weeks ago

Fine-tuned ESM-2 650M with LoRA for predicting protein subcellular localization (10 classes).

Active181 month ago
Python