awesome-repositories.com
Blog
MCP
awesome-repositories.com

Discover the best open-source repositories with AI-powered search.

ExploreCurated searchesOpen-source alternativesSelf-hosted softwareBlogSitemap
ProjectMCP serverAboutHow we rankPress
LegalPrivacyTerms
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
Frooodle avatar

Frooodle/Stirling-PDF

0
View on GitHub↗
81,168 stars·7,124 forks·Java·38 viewsstirling.com↗

Stirling PDF

Stirling-PDF is a web-based PDF management suite used for editing, merging, splitting, and converting PDF documents. It functions as a self-hosted document manager, providing a centralized interface for users to manipulate files on a private server.

The system features a workflow automation engine that allows for the creation of processing pipelines to handle large volumes of documents without writing custom code. It also includes an optical character recognition tool to convert scanned PDFs into searchable and editable text.

Access is managed through single sign-on integration and OIDC compatibility, which supports secure authentication and the maintenance of audit logs for compliance.

The application is delivered as a container-based deployment and exposes its functions through a REST API for external software integration.

Features

  • PDF Manipulation Utilities - Offers a unified interface for merging, splitting, editing, and converting PDF files.
  • OCR Engines - Integrates the Tesseract OCR engine to convert scanned image data into machine-readable text.
  • Optical Character Recognition - Extracts editable and searchable text from scanned PDF documents using OCR technology.
  • PDF Processing Tools - Converts scanned PDFs into searchable documents by extracting text via optical character recognition.
  • Self-Hosted PDF Suites - Functions as a self-hosted suite for managing sensitive PDF documents on private infrastructure.
  • PDF Workflow Orchestrators - Provides a workflow engine to chain multiple PDF operations into automated processing pipelines.
  • Optical Character Recognition - Includes a utility to turn scanned PDF documents into searchable and editable text.
  • Container Deployment - Delivered as a containerized application to ensure consistent execution across different operating systems.
  • Single Sign-On Integrations - Integrates with OIDC identity providers for secure single sign-on access and compliance auditing.
  • Stateless Architectures - Implements a stateless architecture for processing document transformations to ensure scalability across concurrent sessions.
  • PDF - Listed in the “PDF 工具” section of the Great Open Source Project awesome list.
  • Document Management - Web-based utility for merging, splitting, and converting PDF files.

Star history

Star history chart for frooodle/stirling-pdfStar history chart for frooodle/stirling-pdf

How this analysis was created: This summary and feature list were written by an AI model that read the project's README and public documentation pages. Each feature links to the documentation it came from; stars, license and language come straight from the GitHub API. The model does not read the source code, and the analysis is refreshed when the project is re-analysed. Learn more on our About page.

AI search

Explore more awesome repositories

Describe what you need in plain English — the AI ranks thousands of curated open-source projects by relevance.

Start searching with AI

Open-source alternatives to Stirling PDF

Similar open-source projects, ranked by how many features they share with Stirling PDF.
  • stirling-tools/stirling-pdfStirling-Tools avatar

    Stirling-Tools/Stirling-PDF

    81,109View on GitHub↗

    Stirling-PDF is a self-hosted document processing suite designed for secure, private file management. It functions as a comprehensive transformation engine that executes complex operations—such as merging, splitting, converting, and redacting documents—directly on the host machine. The platform provides both a browser-based interface for interactive editing and a programmatic, API-first architecture that allows for the automation of document workflows through standard HTTP requests. The project distinguishes itself through its focus on private, infrastructure-agnostic deployment and granular

    TypeScriptdockerhacktoberfestjava
    View on GitHub↗81,109
  • pdfcrafttool/pdfcraftPDFCraftTool avatar

    PDFCraftTool/pdfcraft

    3,113View on GitHub↗

    Pdfcraft is a containerized service for self-managed PDF processing, editing, and conversion. It provides a toolkit for document manipulation, a multi-format converter, and OCR software to transform scanned documents into searchable and editable text. The project features a visual, node-based workflow editor that allows users to build automated pipelines by chaining together various PDF conversion and optimization operations. The service covers a broad range of capabilities, including document management for merging and splitting files, format conversion between PDFs and office documents or

    JavaScript
    View on GitHub↗3,113
  • awesome-selfhosted/awesome-selfhostedawesome-selfhosted avatar

    awesome-selfhosted/awesome-selfhosted

    299,516View on GitHub↗

    This project is a community-curated directory of open-source software designed for deployment in private server environments and home labs. It serves as a comprehensive resource for discovering independent, self-hosted alternatives to mainstream cloud services, enabling users to maintain full data ownership and control over their digital infrastructure. The directory is structured through a hierarchical taxonomy that organizes a vast collection of applications into logical categories, ranging from media management and data analytics to private communication and team productivity tools. It dis

    awesomeawesome-listcloud
    View on GitHub↗299,516
  • pymupdf/pymupdfpymupdf avatar

    pymupdf/PyMuPDF

    9,086View on GitHub↗

    PyMuPDF is a comprehensive PDF manipulation library and document analysis tool. It serves as a text extraction tool, OCR engine, and image converter, providing a programmatic interface to edit, merge, split, and optimize PDF and Office documents. The project distinguishes itself through high-performance capabilities, including the use of C-bindings for low-level manipulation and parallelized page processing to accelerate workloads. It provides specialized conversion paths, such as transforming PDF content into Markdown for retrieval-augmented generation and large language model pipelines. It

    Pythondata-scienceepubextract-data
    View on GitHub↗9,086
See all 30 alternatives to Stirling PDF→

Frequently asked questions

What does frooodle/stirling-pdf do?

Stirling-PDF is a web-based PDF management suite used for editing, merging, splitting, and converting PDF documents. It functions as a self-hosted document manager, providing a centralized interface for users to manipulate files on a private server.

What are the main features of frooodle/stirling-pdf?

The main features of frooodle/stirling-pdf are: PDF Manipulation Utilities, OCR Engines, Optical Character Recognition, PDF Processing Tools, Self-Hosted PDF Suites, PDF Workflow Orchestrators, Container Deployment, Single Sign-On Integrations.

What are some open-source alternatives to frooodle/stirling-pdf?

Open-source alternatives to frooodle/stirling-pdf include: stirling-tools/stirling-pdf — Stirling-PDF is a self-hosted document processing suite designed for secure, private file management. It functions as… pdfcrafttool/pdfcraft — Pdfcraft is a containerized service for self-managed PDF processing, editing, and conversion. It provides a toolkit… awesome-selfhosted/awesome-selfhosted — This project is a community-curated directory of open-source software designed for deployment in private server… pymupdf/pymupdf — PyMuPDF is a comprehensive PDF manipulation library and document analysis tool. It serves as a text extraction tool,… the-paperless-project/paperless — Paperless is a self-hosted document management system designed to digitize, index, and archive paper documents. It… jaidedai/easyocr — EasyOCR is a deep learning-based computer vision library designed to perform optical character recognition on images…