awesome-repositories.com
Blog
MCP
awesome-repositories.com

Descubre los mejores repositorios open-source con nuestra búsqueda potenciada por IA.

ExplorarBúsquedas curadasAlternativas open-sourceSoftware autohospedableBlogMapa del sitio
ProyectoServidor MCPAcerca deCómo clasificamosPrensa
Aviso legalPrivacidadTérminos
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·

1 repositorio

Awesome GitHub RepositoriesGenomic Tokenization

Processes raw nucleotide sequences into discrete tokens for structured genomic analysis.

Distinct from Genomic Sequence Interpreters: Focuses on the encoding/tokenization process for genomic data, whereas Genomic Sequence Interpreters focus on the analysis of the resulting data.

Explore 1 awesome GitHub repository matching data & databases · Genomic Tokenization. Refine with filters or upvote what's useful.

Awesome Genomic Tokenization GitHub Repositories

Encuentra los mejores repositorios con IA.Buscaremos los repositorios que mejor coincidan usando IA.
  • arcinstitute/evo2Avatar de ArcInstitute

    ArcInstitute/evo2

    3,951Ver en GitHub↗

    evo2 is a genomic large language model and foundation model designed to predict, generate, and analyze genetic information across different species. It functions as a nucleotide sequence modeler and a DNA sequence generator, using transformer-based sequence modeling to process genomic data. The system provides capabilities for synthetic DNA generation, creating new genetic sequences based on biological prompts or species-specific tags. It also performs nucleotide likelihood prediction to score genomic variants and analyze biological properties within DNA sequences. The model supports genomic

    Converts raw nucleotide sequences into discrete tokens that the model processes as a structured vocabulary.

    Jupyter Notebook
    Ver en GitHub↗3,951
  1. Home
  2. Data & Databases
  3. Data Analysis & Visualization
  4. Analytical Platforms and Engines
  5. Sequence Analysis
  6. Genomic Sequence Interpreters
  7. Genomic Tokenization