GenerTeam/GENERator-v2-eukaryote-1.2b-base

https://huggingface.co/GenerTeam/GENERator-v2-eukaryote-1.2b-base
Activeby GenerTeam3.5K4updated 3 months ago
Python

## Important Notice If you are using GENERator for sequence generation, please ensure that the length of each input sequence is a multiple of 6. This can be achieved by either: 1. Padding the sequence on the left with 'A' (left padding); 2. Truncating the sequence from the left (left truncation).

Sourced from

  • HuggingFace — GenerTeam/GENERator-v2-eukaryote-1.2b-base

Related resources

Evo2-1B-Base (Transformers port)

Active1.7K3 weeks ago
Python

Evo2-7B (Transformers port)

Active1.3K3 weeks ago
Python

The plant DNA large language models (LLMs) contain a series of foundation models based on different model architectures, which are pre-trained on various plant reference genomes. All the models have a comparable model size between 90 MB and 150 MB, BPE tokenizer is used for tokenization and 8000…

Idle3171 year ago
Python

This 1,120,772,224-parameter nucleotide-level causal language model is a member of the eight-model MarinDNA v0.5 parameter-scaling ladder developed with Marin. This repository contains only the final step-215573 checkpoint from run dna-bolinas-scaling-v0.5-h1920-p1B-0dc6f4, with its tokenizer…

Active2132 months ago
Python

GitHub homepage: Cell2Sentence GitHub

Idle1.5K11 months ago
Python

# Overview This is the CellHermes model, based on the LLaMA-3.1-8B-instruct architecture developed by Meta, fine-tuned using single-cell RNA sequencing (scRNA-seq) datasets from CellxGene and PPI network from BioGRID. CellHermes is an innovative framework for adapting existing large language models…

Active773 months ago
Python