awesome-repositories.com
ब्लॉग
MCP
awesome-repositories.com

AI-संचालित खोज के साथ बेहतरीन ओपन-सोर्स रिपॉजिटरी खोजें।

एक्सप्लोर करेंक्यूरेटेड खोजेंओपन-सोर्स विकल्पसेल्फ-होस्टेड सॉफ्टवेयरब्लॉगसाइटमैप
प्रोजेक्टMCP सर्वरहमारे बारे मेंहम रैंकिंग कैसे करते हैंप्रेस
कानूनीगोपनीयताशर्तें
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
katanaml avatar

katanaml/sparrow

0
View on GitHub↗
5,162 स्टार्स·517 फोर्क्स·Python·GPL-3.0·17 व्यूज़sparrow.katanaml.io↗

Sparrow

Sparrow एक LLM डॉक्यूमेंट एक्सट्रैक्शन प्लेटफॉर्म और विज़न-आधारित इन्फरेंस इंजन है जिसे छवियों और PDFs को वैलिडेटेड स्ट्रक्चर्ड डेटा में बदलने के लिए डिज़ाइन किया गया है। यह एक एजेंटिक वर्कफ़्लो ऑर्केस्ट्रेटर के रूप में कार्य करता है जो वर्गीकरण, निष्कर्षण और सत्यापन कार्यों को मल्टी-स्टेप पाइपलाइनों में जोड़ता है।

यह सिस्टम एक बैकएंड-अज्ञेयवादी इन्फरेंस लेयर के माध्यम से खुद को अलग करता है जो लोकल GPUs, Apple Silicon और क्लाउड प्रदाताओं के बीच मॉडल का प्रबंधन करता है। यह एक्सट्रैक्ट किए गए टेक्स्ट को सटीक बाउंडिंग बॉक्स निर्देशांकों पर मैप करने के लिए कोऑर्डिनेट-आधारित विजुअल ग्राउंडिंग का उपयोग करता है और ध्यान केंद्रित करने और डेटा फॉर्मेट को सामान्य करने के लिए हिंट-आधारित मॉडल स्टीयरिंग का उपयोग करता है।

यह प्लेटफॉर्म डॉक्यूमेंट इंटेलिजेंस वर्कफ़्लो को कवर करता है, जिसमें संरचनात्मक अखंडता बनाए रखने के लिए विशेष इमेज-आधारित टेबल प्रोसेसिंग और एक्सट्रैक्ट किए गए फील्ड्स की शुद्धता को सत्यापित करने के लिए स्कीमा-आधारित सत्यापन शामिल है। यह API प्रदर्शन, उपयोग एनालिटिक्स और सिस्टम हेल्थ की निगरानी के लिए एक डॉक्यूमेंट एनालिसिस डैशबोर्ड भी प्रदान करता है।

आर्किटेक्चर में इंडेक्सिंग और ऑर्केस्ट्रेशन के लिए उपयोग की जाने वाली थर्ड-पार्टी लाइब्रेरीज़ को एकीकृत करने के लिए एक प्लगइन-आधारित एक्सटेंशन सिस्टम शामिल है।

Features

  • Intelligent Document Processing - Provides a platform for intelligent document processing, combining classification, extraction, and validation into multi-step pipelines.
  • Vision-Language Model Backends - Uses vision-capable language models to parse document layouts and convert visual content into structured data.
  • Agentic Workflow Pipelines - Implements automated pipelines that chain LLM instructions and external tools for complex document analysis and error recovery.
  • Hardware Acceleration Backends - Manages extraction pipelines across diverse hardware accelerators including local and cloud-based backends.
  • Image Text Extractions - Recognizes and extracts text and key-value pairs from images and PDFs as structured data.
  • Vision-Language Orchestrators - Manages model inference across local GPUs, Apple Silicon, and cloud providers to process visual document data.
  • Hardware-Agnostic Inference Layers - Provides a hardware-agnostic inference layer that routes processing to local GPUs, Apple Silicon, or cloud providers.
  • Structured Document Extraction - Converts visual document layouts into machine-readable structured formats using vision-capable models.
  • Vision-Language Inference - Implements a vision-language inference engine that executes multimodal models across various hardware backends.
  • Workflow Orchestration - Combines document classification and data extraction into a single AI workflow pipeline with visual monitoring.
  • Document Field Validations - Checks extracted document fields against schemas to verify the presence and correctness of required data.
  • Structured Data Extraction - Parses complex tables and text from documents into predefined schemas with bounding box coordinate mapping.
  • PDF Coordinate Extraction - Extracts precise bounding box coordinates for recognized text regions within PDF pages.
  • Agentic Workflow Orchestrators - Functions as an orchestrator that chains LLM classification, extraction, and validation tasks with integrated error recovery.
  • Pipeline Orchestrators - Chains classification, extraction, and validation tasks into sequenced pipelines with error recovery.
  • Schema-Driven Validations - Verifies extracted document fields against predefined structural definitions to ensure data correctness.
  • Extraction Coordinate Annotations - Generates bounding box coordinates for extracted elements to provide visual grounding for the data.
  • Visual Coordinate Mapping - Maps extracted text to precise bounding box coordinates for visual grounding within documents.
  • Tabular Data Extraction - Extracts complex tabular data from documents while maintaining structural integrity through specialized vision processing.
  • OCR Document Conversion - Converts images and PDFs into validated structured data using vision models and schema-based validation.
  • Local Model Backends - Supports running inference across a variety of backends including local GPUs, Apple Silicon, and cloud providers.
  • Attention Steering Hints - Uses hint-based configuration files to steer model attention and normalize extracted data formats.
  • Extraction Hinting - Uses hint-based model steering to guide attention and normalize data formats during the extraction process.
  • Document Processing Pipelines - Extracts and analyzes data across documents containing multiple pages using orchestrated pipelines.
  • Table Structure Detections - Identifies tabular grids and merges cells within document layouts to crop them for specialized inference.
  • Image-Based Table Extractors - Identifies tabular regions and crops them into images for specialized inference to preserve structural integrity.
  • Document Table Extractors - Maps large or multi-column tables from documents to structured schemas using an intermediate processing pipeline.
  • Document Analysis Dashboards - Provides a visual dashboard for monitoring API performance, usage analytics, and the operational health of extraction pipelines.
  • Data Processing - Solution for efficient data extraction from documents and images.
  • Data Processing Tools - Solution for efficient data extraction from documents and images.

स्टार हिस्ट्री

katanaml/sparrow के लिए स्टार हिस्ट्री चार्टkatanaml/sparrow के लिए स्टार हिस्ट्री चार्ट

AI सर्च

और अधिक बेहतरीन रिपॉजिटरी खोजें

अपनी ज़रूरत को सरल भाषा में बताएं — AI हजारों क्यूरेटेड ओपन-सोर्स प्रोजेक्ट्स को प्रासंगिकता के आधार पर रैंक करता है।

Start searching with AI

Sparrow के ओपन-सोर्स विकल्प

समान ओपन-सोर्स प्रोजेक्ट्स, जो Sparrow के साथ साझा की गई सुविधाओं के आधार पर रैंक किए गए हैं।
  • datalab-to/chandradatalab-to का अवतार

    datalab-to/chandra

    4,833GitHub पर देखें↗

    sChandra is a document processing platform that converts images, PDFs, Word documents, spreadsheets, and other formats into structured output such as HTML, Markdown, or JSON while preserving layout. It can also extract specific data fields from invoices, contracts, or reports using user-defined JSON schemas, with citations back to source locations. The service supports form filling in PDF and image documents, document generation from Markdown, and extraction of tracked changes from Word files. The platform distinguishes itself with pipeline-based processing chains that combine multiple proces

    Pythonaiocr
    GitHub पर देखें↗4,833
  • kreuzberg-dev/kreuzbergkreuzberg-dev का अवतार

    kreuzberg-dev/kreuzberg

    8,527GitHub पर देखें↗

    Kreuzberg is a document extraction engine that converts PDFs, Office files, images, and over 90 other formats into clean, structured text and metadata. It is built around a compiled Rust core that can be used as a native library, a command-line tool, a REST API server, or a WebAssembly module for browser-based processing. The system is designed to run entirely on self-hosted infrastructure, with no data leaving the user's environment. What distinguishes Kreuzberg is its breadth of integration surfaces and its pipeline architecture. It exposes extraction capabilities through native bindings fo

    Rustdocument-intelligenceelixirffi
    GitHub पर देखें↗8,527
  • bytedance/dolphinbytedance का अवतार

    bytedance/Dolphin

    8,820GitHub पर देखें↗

    Dolphin is a multimodal layout analyzer and image-to-structure converter that transforms photographed or digital document images into machine-readable structured data. It functions as an LLM document parser, utilizing vision-language models to simultaneously predict spatial layout and text content. The system is designed as a concurrent document processor, employing parallel document parsing to process multiple elements across distributed compute nodes. This high-throughput approach reduces the total time required to convert large volumes of images into structured formats. The project covers

    Pythondocument-analysislayout-analysisocr
    GitHub पर देखें↗8,820
  • chonkie-inc/chonkiechonkie-inc का अवतार

    chonkie-inc/chonkie

    4,170GitHub पर देखें↗

    Chonkie is a text chunking library designed for retrieval-augmented generation pipelines. It functions as a semantic text splitter and RAG ingestion pipeline, transforming raw text into embedded segments for storage in vector databases. The project distinguishes itself through specialized splitting strategies, including an AST-based code splitter for preserving logical boundaries in source code and a semantic text splitter that uses embedding models to determine boundaries based on meaning. It also provides a vector database ingestor to automate the generation of embeddings and their export t

    Pythonaichonkiechunker
    GitHub पर देखें↗4,170
Sparrow के सभी 30 विकल्प देखें→

अक्सर पूछे जाने वाले प्रश्न

katanaml/sparrow क्या करता है?

Sparrow एक LLM डॉक्यूमेंट एक्सट्रैक्शन प्लेटफॉर्म और विज़न-आधारित इन्फरेंस इंजन है जिसे छवियों और PDFs को वैलिडेटेड स्ट्रक्चर्ड डेटा में बदलने के लिए डिज़ाइन किया गया है। यह एक एजेंटिक वर्कफ़्लो ऑर्केस्ट्रेटर के रूप में कार्य करता है जो वर्गीकरण, निष्कर्षण और सत्यापन कार्यों को मल्टी-स्टेप पाइपलाइनों में जोड़ता है।

katanaml/sparrow की मुख्य विशेषताएं क्या हैं?

katanaml/sparrow की मुख्य विशेषताएं हैं: Intelligent Document Processing, Vision-Language Model Backends, Agentic Workflow Pipelines, Hardware Acceleration Backends, Image Text Extractions, Vision-Language Orchestrators, Hardware-Agnostic Inference Layers, Structured Document Extraction।

katanaml/sparrow के कुछ ओपन-सोर्स विकल्प क्या हैं?

katanaml/sparrow के ओपन-सोर्स विकल्पों में शामिल हैं: datalab-to/chandra — sChandra is a document processing platform that converts images, PDFs, Word documents, spreadsheets, and other formats… kreuzberg-dev/kreuzberg — Kreuzberg is a document extraction engine that converts PDFs, Office files, images, and over 90 other formats into… bytedance/dolphin — Dolphin is a multimodal layout analyzer and image-to-structure converter that transforms photographed or digital… chonkie-inc/chonkie — Chonkie is a text chunking library designed for retrieval-augmented generation pipelines. It functions as a semantic… run-llama/liteparse — A fast, helpful, and open-source document parser. getomni-ai/zerox — Zerox is a multimodal document parser and OCR tool that uses vision models to convert PDF files and images into…