awesome-repositories.com
Blog
MCP
awesome-repositories.com

Descubre los mejores repositorios open-source con nuestra búsqueda potenciada por IA.

ExplorarBúsquedas curadasAlternativas open-sourceSoftware autohospedableBlogMapa del sitio
ProyectoServidor MCPAcerca deCómo clasificamosPrensa
Aviso legalPrivacidadTérminos
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·

1 repositorio

Awesome GitHub RepositoriesHeading Level Classifiers

Capabilities for identifying heading levels (H1-H6) from font size clustering and semantic analysis in documents.

Distinct from Structured Document Extraction: Distinct from Structured Document Extraction: focuses specifically on classifying heading hierarchy rather than general layout-to-markdown conversion.

Explore 1 awesome GitHub repository matching artificial intelligence & ml · Heading Level Classifiers. Refine with filters or upvote what's useful.

Awesome Heading Level Classifiers GitHub Repositories

Encuentra los mejores repositorios con IA.Buscaremos los repositorios que mejor coincidan usando IA.
  • kreuzberg-dev/kreuzbergAvatar de kreuzberg-dev

    kreuzberg-dev/kreuzberg

    8,527Ver en GitHub↗

    Kreuzberg is a document extraction engine that converts PDFs, Office files, images, and over 90 other formats into clean, structured text and metadata. It is built around a compiled Rust core that can be used as a native library, a command-line tool, a REST API server, or a WebAssembly module for browser-based processing. The system is designed to run entirely on self-hosted infrastructure, with no data leaving the user's environment. What distinguishes Kreuzberg is its breadth of integration surfaces and its pipeline architecture. It exposes extraction capabilities through native bindings fo

    Identifies heading levels (H1-H6) from font size clustering and semantic analysis in PDFs.

    Rustdocument-intelligenceelixirffi
    Ver en GitHub↗8,527
  1. Home
  2. Artificial Intelligence & ML
  3. Natural Language Processing
  4. Structured Document Extraction
  5. Heading Level Classifiers