awesome-repositories.com
المدونة
awesome-repositories.com

اكتشف أفضل مستودعات المصادر المفتوحة باستخدام بحث مدعوم بالذكاء الاصطناعي.

استكشفعمليات بحث منسقةبدائل مفتوحة المصدربرمجيات ذاتية الاستضافةالمدونةخريطة الموقع
المشروعحولكيفية ترتيب النتائجالصحافةخادم MCP
قانونيالخصوصيةالشروط
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
Back to watson-developer-cloud/natural-language-understanding-nodejs

Open-source alternatives to Natural Language Understanding Nodejs

30 open-source projects similar to watson-developer-cloud/natural-language-understanding-nodejs, ranked by how many features they have in common. Compare stars, activity and what each one does to find the best Natural Language Understanding Nodejs alternative.

  • 01walid/goarabicالصورة الرمزية لـ 01walid

    01walid/goarabic

    117عرض على GitHub↗

    A Go Lang package for dealing with Arabic text.

    Go
    عرض على GitHub↗117
  • adbar/german-nlpالصورة الرمزية لـ adbar

    adbar/German-NLP

    526عرض على GitHub↗

    Curated list of open-access/open-source/off-the-shelf resources and tools developed with a particular focus on German

    عرض على GitHub↗526
  • adobe/nlp-cubeالصورة الرمزية لـ adobe

    adobe/NLP-Cube

    562عرض على GitHub↗

    05 August 2021 - We are releasing version 3.0 of NLPCube and models and introducing FLAVOURS. This is a major update, but we did our best to maintain the same API, so previous implementation will not crash. The supported language list is smaller, but you can open an issue for unsupported…

    HTML
    عرض على GitHub↗562
  • alexandrainst/danlpالصورة الرمزية لـ alexandrainst

    alexandrainst/danlp

    209عرض على GitHub↗

    Part of Speech Tagging | Dependency Parsing Named Entity Recognition | Named Entity Disambiguation | Coreference Resolution Sentiment Analysis | Hatespeech Detection Embeddings | Datasets | Tutorials

    Python
    عرض على GitHub↗209
  • alirezatheh/perkeالصورة الرمزية لـ AlirezaTheH

    AlirezaTheH/perke

    73عرض على GitHub↗

    Perke is a Python keyphrase extraction package for Persian language. It provides an end-to-end keyphrase extraction pipeline in which each component can be easily modified or extended to develop new models.

    Python
    عرض على GitHub↗73

بحث بالذكاء الاصطناعي

استكشف المزيد من المستودعات الرائعة

صف ما تحتاجه بلغة بسيطة — وسيقوم الذكاء الاصطناعي بترتيب آلاف المشاريع مفتوحة المصدر المنسقة حسب الصلة.

Find more with AI search
  • amir-zeldes/rftokenizerالصورة الرمزية لـ amir-zeldes

    amir-zeldes/RFTokenizer

    31عرض على GitHub↗

    A character-wise tokenizer for morphologically rich languages

    Lex
    عرض على GitHub↗31
  • aziz/virastarالصورة الرمزية لـ aziz

    aziz/virastar

    88عرض على GitHub↗

    #ویراستار نوشته‌های فارسی شما را ویرایش می‌کند

    Ruby
    عرض على GitHub↗88
  • botcenter/spanishsent2vecالصورة الرمزية لـ BotCenter

    BotCenter/spanishSent2Vec

    4عرض على GitHub↗

    Spanish Sentence Embeddings trained using sent2vec on the Spanish Unannotated Corpora.

    عرض على GitHub↗4
  • botcenter/spanishwordembeddingsالصورة الرمزية لـ BotCenter

    BotCenter/spanishWordEmbeddings

    9عرض على GitHub↗

    Spanish words embeddings computed using fastText on the Spanish Unannotated Corpora.

    عرض على GitHub↗9
  • calmdownkarm/sivareddydependencyparserالصورة الرمزية لـ CalmDownKarm

    CalmDownKarm/sivareddydependencyparser

    0عرض على GitHub↗

    Your input file should have the extension .input.txt e.g. hindi.input.txt To dependency tag your input file run "make .output" e.g.

    Lex
    عرض على GitHub↗0
  • dccuchile/betoالصورة الرمزية لـ dccuchile

    dccuchile/beto

    505عرض على GitHub↗

    BETO is a BERT model trained on a big Spanish corpus. BETO is of size similar to a BERT-Base and was trained with the Whole Word Masking technique. Below you find Tensorflow and Pytorch checkpoints for the uncased and cased versions, as well as some results for Spanish benchmarks comparing BETO…

    عرض على GitHub↗505
  • dccuchile/spanish-word-embeddingsالصورة الرمزية لـ dccuchile

    dccuchile/spanish-word-embeddings

    365عرض على GitHub↗

    Below you find links to Spanish word embeddings computed with different methods and from different corpora. Whenever it is possible, a description of the parameters used to compute the embeddings is included, together with simple statistics of the vectors, vocabulary, and description of the…

    عرض على GitHub↗365
  • ejtaal/jsastemالصورة الرمزية لـ ejtaal

    ejtaal/jsastem

    26عرض على GitHub↗

    JSASTEM - JavaScript Arabic Stemmer

    JavaScript
    عرض على GitHub↗26
  • fighting41love/funnlpالصورة الرمزية لـ fighting41love

    fighting41love/funNLP

    81,299عرض على GitHub↗

    This project is a community-driven knowledge base and curated repository focused on natural language processing and large language model development. It serves as a centralized index for high-quality tools, libraries, and research materials, organizing technical resources into structured, version-controlled documentation to assist developers in navigating the evolving artificial intelligence ecosystem. The repository distinguishes itself by acting as an aggregator for AI model evaluation and benchmarking. It provides access to tools that enable the simultaneous comparison of multiple conversa

    Python
    عرض على GitHub↗81,299
  • fnielsen/awesome-danishالصورة الرمزية لـ fnielsen

    fnielsen/awesome-danish

    195عرض على GitHub↗

    A curated list of awesome resources for Danish language technology

    عرض على GitHub↗195
  • fudannlp/fnlpالصورة الرمزية لـ FudanNLP

    FudanNLP/fnlp

    2,690عرض على GitHub↗

    FudanNLP (FNLP)

    Java
    عرض على GitHub↗2,690
  • fxsjy/jiebaالصورة الرمزية لـ fxsjy

    fxsjy/jieba

    35,027عرض على GitHub↗

    This project is a Chinese text segmentation library and tokenizer designed to split Chinese sentences into individual words. It serves as a natural language processing tool for splitting characters into words, tagging parts of speech, and extracting keywords using statistical analysis. The library distinguishes itself through support for custom dictionary configuration and vocabulary file management, allowing users to override default segmentation rules for domain-specific accuracy. It also includes a TF-IDF keyword extractor to identify significant words and core topics within documents. Th

    Python
    عرض على GitHub↗35,027
  • galuhsahid/indonesian-word-embeddingالصورة الرمزية لـ galuhsahid

    galuhsahid/indonesian-word-embedding

    20عرض على GitHub↗

    A web application that demonstrates Indonesian word embedding, inspired by Word embedding demo.

    JavaScript
    عرض على GitHub↗20
  • goru001/inltkالصورة الرمزية لـ goru001

    goru001/inltk

    840عرض على GitHub↗

    iNLTK aims to provide out of the box support for various NLP tasks that an application developer might need for Indic languages.

    Python
    عرض على GitHub↗840
  • hankcs/hanlpالصورة الرمزية لـ hankcs

    hankcs/HanLP

    36,413عرض على GitHub↗

    HanLP is a natural language processing library and deep learning framework specifically optimized for the Chinese language, while also functioning as a multilingual text processor. It serves as a toolkit for performing linguistic analysis, semantic understanding, and script conversion. The project distinguishes itself through a dedicated focus on Chinese linguistic structures, including a specialized script converter for transforming text between Simplified Chinese, Traditional Chinese, and Pinyin. It further supports domain-specific model training to improve the recognition of professional t

    Pythondependency-parserhanlpnamed-entity-recognition
    عرض على GitHub↗36,413
  • ictrc/parsivarالصورة الرمزية لـ ICTRC

    ICTRC/Parsivar

    247عرض على GitHub↗

    parsivar

    Python
    عرض على GitHub↗247
  • isnowfy/snownlpالصورة الرمزية لـ isnowfy

    isnowfy/snownlp

    6,631عرض على GitHub↗

    SnowNLP is a Python library for Chinese natural language processing. It provides tools for text segmentation, sentiment analysis, document classification, and phonetic transliteration. The library includes capabilities for training and saving custom machine learning models for tokenization and sentiment analysis using raw training datasets. It covers a range of linguistic processing areas, including parts of speech tagging, sentence splitting, and text similarity measurement. The toolkit also provides utilities for extracting key information through text summarization and calculating word im

    Python
    عرض على GitHub↗6,631
  • jfreddypuentes/spanlpالصورة الرمزية لـ jfreddypuentes

    jfreddypuentes/spanlp

    41عرض على GitHub↗

    spanlp es una librería escrita en Python para detectar, censurar y limpiar groserías, vulgaridades, palabras de odio, racismo, xenofobia y bullying en textos escritos en Español .

    Python
    عرض على GitHub↗41
  • jonsafari/perstemالصورة الرمزية لـ jonsafari

    jonsafari/perstem

    19عرض على GitHub↗

    Persian (Farsi) stemmer, morphological analyzer, transliterator, and partial part-of-speech tagger. Input may be encoded as Perso-Arabic script UTF-8, ISIRI 3342, Windows-1256, SGML/HTML/XML-style numeric character references (ncr), or dehdari-transliterated latin-script text. Use the -i flag to…

    Perl
    عرض على GitHub↗19
  • kangfend/bahasaالصورة الرمزية لـ kangfend

    kangfend/bahasa

    20عرض على GitHub↗

    BAHASA

    Python
    عرض على GitHub↗20
  • kenjiroai/synthaiالصورة الرمزية لـ KenjiroAI

    KenjiroAI/SynThai

    41عرض على GitHub↗

    Thai Word Segmentation and Part-of-Speech Tagging with Deep Learning

    Python
    عرض على GitHub↗41
  • ksopyla/awesome-nlp-polishالصورة الرمزية لـ ksopyla

    ksopyla/awesome-nlp-polish

    308عرض على GitHub↗

    A curated list of resources dedicated to Natural Language Processing (NLP) in polish. Models, tools, datasets.

    عرض على GitHub↗308
  • mikahama/uralicnlpالصورة الرمزية لـ mikahama

    mikahama/uralicNLP

    98عرض على GitHub↗

    Natural language processing for many languages

    Python
    عرض على GitHub↗98
  • narimann2/parsianalyzerالصورة الرمزية لـ NarimanN2

    NarimanN2/ParsiAnalyzer

    166عرض على GitHub↗

    Persian Analyzer for Elasticsearch.

    Java
    عرض على GitHub↗166
  • phuonglh/vn.vitkالصورة الرمزية لـ phuonglh

    phuonglh/vn.vitk

    218عرض على GitHub↗

    NOTE: This repos is now obsolete. Interested programmers should consider to use the new repo vlp (github.com/phuonglh/vlp) We have preferred using Scala instead of Java since 2016.

    Java
    عرض على GitHub↗218