awesome-repositories.com
Blog
MCP
awesome-repositories.com

Entdecke die besten Open-Source-Repositories mit KI-gestützter Suche.

EntdeckenKuratierte SuchenOpen-Source-AlternativenSelf-hosted SoftwareBlogSitemap
ProjektMCP-ServerÜber unsRanking-MethodikPresse
RechtlichesDatenschutzAGB
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
languagetool-org avatar

languagetool-org/languagetool

0
View on GitHub↗
14,597 Stars·1,527 Forks·Java·LGPL-2.1·13 Aufrufelanguagetool.org↗

Languagetool

LanguageTool is a multilingual grammar and style checking engine designed to detect spelling, grammar, and writing errors across multiple languages. It provides automated proofreading capabilities that can be deployed as a self-hosted server or executed as a standalone local desktop application.

The project distinguishes itself through a flexible rule development framework, allowing linguistic patterns to be defined via XML or implemented as custom Java classes. It utilizes n-gram frequency modeling for confused word detection and supports neural word embeddings to improve disambiguation between similar word pairs.

The engine covers a broad range of linguistic analysis capabilities, including part-of-speech tagging, sentence boundary definition, and markup-aware proofreading for documents containing HTML, XML, or LaTeX. It also features comprehensive dictionary management tools for compiling binary dictionaries and managing custom spell-checking rules.

Developers can integrate these capabilities into external applications using a REST-based API, which includes support for automatic language detection and access restriction controls.

Features

  • Proofreading Tools - Ships a comprehensive suite of proofreading tools for detecting and correcting grammatical and stylistic errors in text.
  • Spell and Grammar Checkers - Detects spelling, grammar, and style errors across multiple languages to improve overall writing quality.
  • N-Gram Co-occurrence Models - Utilizes n-gram co-occurrence modeling to identify frequently confused words and rank spelling suggestions.
  • Sentence Boundary Detection - Implements sentence boundary detection using regular expressions to split text into individual sentences for analysis.
  • Text Error Detection APIs - Exposes a REST API for analyzing text and identifying grammar and style mistakes across multiple languages.
  • Programmatic Proofreading Interfaces - Integrates automated text analysis and error detection into external applications via a REST API.
  • Writing Error Detection - Analyzes text against collections of common misspellings and grammatical mistakes to identify writing errors.
  • Spelling and Language Tools - Identifies misspelled words and manages custom dictionaries to ensure spelling accuracy across domains.
  • Linguistic Rule Development - Provides advanced detection patterns using unification and specialized filters to target complex linguistic errors.
  • Rule-Based Grammar Definitions - Defines linguistic patterns and error detection logic using XML schemas to match tokens and part-of-speech tags.
  • Grammar Rule Specification - Allows creating custom grammar and style detection patterns using XML, regular expressions, and part-of-speech tags.
  • Linguistic Token Matching - Analyzes text by breaking it into discrete tokens and applying unification and filter logic to detect errors.
  • Character Offset Trackers - Uses character offset trackers to identify the exact positions of errors relative to the original raw text containing markup.
  • Confused Word Detection - Uses large n-gram datasets to identify and correct frequently swapped words based on linguistic patterns.
  • Language Detection Tools - Includes utilities for identifying the language of provided text content to apply the correct grammar rules.
  • Text Analysis APIs - Provides a text analysis API to detect grammar and style errors based on provided language specifications.
  • Linguistic Language Modules - Provides a framework for adding new languages by defining modules and writing custom error detection rules.
  • Markup-Aware Analysis - Identifies errors in HTML, XML, or LaTeX documents by mapping text positions and ignoring tags.
  • Spell Check Whitelisting - Allows marking specific words or patterns to be ignored by the spell checker in designated contexts.
  • Spell Checking Rules - Provides mechanisms to define custom lists of ignored and prohibited words to refine spelling validation.
  • Markup-Aware Text Analysis - Analyzes text in HTML, XML, or LaTeX by ignoring tags while maintaining character positions for accurate error highlighting.
  • Linguistic Variable Capturing - Provides the ability to store matched text in variables for use in subsequent patterns and context-aware suggestions.
  • Analysis Integration Interfaces - Provides programmatic interfaces for embedding grammar and style analysis tools into custom workflows and automated pipelines.
  • Desktop Applications - Offers a standalone desktop application for local proofreading without requiring a remote server.
  • Linguistic Extension Frameworks - Enables adding new languages via custom tagger dictionaries, disambiguators, and sentence segmentation rules.
  • Proofreading Rule Specification - Implements specialized grammar and style checks using unification and tone tags to detect linguistic patterns.
  • Self-Hosted Instances - Provides a local instance of the grammar checking engine to process text without relying on cloud services.
  • Programmatic API Interfaces - Provides Java and HTTP interfaces that expose the full functionality of the proofreading engine for automation.
  • Rule Logic Extensions - Allows implementation of complex validation logic by extending base Java classes for linguistic rules.
  • Custom Dictionaries - Supports the use of user-defined word lists to improve and refine spell-checking accuracy.
  • Regex-Based Exclusion Rules - Provides a system for specifying antipatterns or exclusion regexes to prevent conflicting rule triggers.
  • Client-Server Architecture - Employs a client-server architecture where a central HTTP server handles the heavy lifting of text analysis.
  • Java-Based Rule Implementation - Supports developing complex validation logic by extending base classes when XML patterns are insufficient.
  • False Positive Filtering - Uses regex anti-patterns to prevent specific text sequences from triggering grammar or style rules.
  • Suggestion Ranking - Integrates word usage frequency data to ensure the most common correct spellings appear first in suggestions.
  • Rule Antipatterns - Prevents false positives by defining token sequences that deactivate specific matching rules.
  • Proofreading Servers - Allows running a self-contained HTTP instance locally to process text without relying on external cloud services.

Star-Verlauf

Star-Verlauf für languagetool-org/languagetoolStar-Verlauf für languagetool-org/languagetool

KI-Suche

Entdecke weitere awesome Repositories

Beschreibe in einfachen Worten, was du brauchst — die KI bewertet tausende kuratierte Open-Source-Projekte nach Relevanz.

Start searching with AI

Open-Source-Alternativen zu Languagetool

Ähnliche Open-Source-Projekte, sortiert nach der Anzahl der gemeinsamen Funktionen mit Languagetool.
  • vale-cli/valeAvatar von vale-cli

    vale-cli/vale

    5,477Auf GitHub ansehen↗

    Vale is a markup-aware prose linter and command-line interface tool designed to enforce editorial style guides and grammar rules across various document formats. It functions as a YAML-based style guide engine that analyzes text for consistency in tone, spelling, and terminology while ignoring non-prose elements like code blocks. The project distinguishes itself through a flexible extensibility model that allows users to define custom linting rules using YAML configurations, regular expressions, and external scripts for complex validation logic. It supports a wide array of documentation forma

    Golinterlintingvale
    Auf GitHub ansehen↗5,477
  • amperser/proselintAvatar von amperser

    amperser/proselint

    4,542Auf GitHub ansehen↗

    Proselint is a prose linter and rule-based text analyzer designed to identify stylistic errors, clichés, and jargon in written text. It scans documents against a curated registry of linguistic and typographic rules to maintain professional editorial standards and improve writing quality. The project functions as a command line text processor, a programmable analysis library, and a git pre-commit hook. Its modular architecture allows the core engine to be embedded into other applications, exposed via a REST API, or integrated into text editors. The tool supports recursive directory traversal

    JavaScript
    Auf GitHub ansehen↗4,542
  • amzxyz/rime_wanxiangAvatar von amzxyz

    amzxyz/rime_wanxiang

    2,863Auf GitHub ansehen↗

    This project is a CJK input method framework and configuration set designed for the Rime input engine. It provides a comprehensive system of schemas and dictionary packs to optimize Chinese character entry through pinyin and double-pinyin workflows. The framework is distinguished by its use of Lua-powered extensions that add dynamic utilities, such as inline mathematical calculators, automated timestamps, and text formatting, directly to the input interface. It also features refined word libraries and language models specifically tuned to improve prediction accuracy and first-choice hit rates

    Luadictsrimerime-config
    Auf GitHub ansehen↗2,863
  • crate-ci/typosAvatar von crate-ci

    crate-ci/typos

    4,002Auf GitHub ansehen↗

    Typos is a source code spell checker and automated typo fixer designed to detect and correct spelling errors across programming languages and project files. It functions as a CI spelling validator and SARIF compatible linter, allowing projects to prevent misspelled text from reaching production. The tool features a customizable dictionary engine that utilizes TOML configuration and locale-specific dictionaries to manage project-specific terminology. It differentiates itself by splitting programming language identifiers into individual words for validation and verifying the spelling of filenam

    Rust
    Auf GitHub ansehen↗4,002
Alle 30 Alternativen zu Languagetool anzeigen→

Häufig gestellte Fragen

Was macht languagetool-org/languagetool?

LanguageTool is a multilingual grammar and style checking engine designed to detect spelling, grammar, and writing errors across multiple languages. It provides automated proofreading capabilities that can be deployed as a self-hosted server or executed as a standalone local desktop application.

Was sind die Hauptfunktionen von languagetool-org/languagetool?

Die Hauptfunktionen von languagetool-org/languagetool sind: Proofreading Tools, Spell and Grammar Checkers, N-Gram Co-occurrence Models, Sentence Boundary Detection, Text Error Detection APIs, Programmatic Proofreading Interfaces, Writing Error Detection, Spelling and Language Tools.

Welche Open-Source-Alternativen gibt es zu languagetool-org/languagetool?

Open-Source-Alternativen zu languagetool-org/languagetool sind unter anderem: vale-cli/vale — Vale is a markup-aware prose linter and command-line interface tool designed to enforce editorial style guides and… amperser/proselint — Proselint is a prose linter and rule-based text analyzer designed to identify stylistic errors, clichés, and jargon in… amzxyz/rime_wanxiang — This project is a CJK input method framework and configuration set designed for the Rime input engine. It provides a… crate-ci/typos — Typos is a source code spell checker and automated typo fixer designed to detect and correct spelling errors across… pbek/qownnotes — QOwnNotes is a desktop note editor that stores each note as a plain-text Markdown file on the local filesystem,… hankcs/hanlp — HanLP is a natural language processing library and deep learning framework specifically optimized for the Chinese…