awesome-repositories.com
Blog
awesome-repositories.com

Descubre los mejores repositorios open-source con nuestra búsqueda potenciada por IA.

ExplorarBúsquedas curadasAlternativas open-sourceSoftware autohospedableBlogMapa del sitio
ProyectoAcerca deCómo clasificamosPrensaServidor MCP
Aviso legalPrivacidadTérminos
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
ownthink avatar

ownthink/KnowledgeGraphData

0
View on GitHub↗
5,181 estrellas·737 forks·Python·7 vistaswww.ownthink.com↗

KnowledgeGraphData

KnowledgeGraphData es una colección de conjuntos de datos estructurados y corpora diseñados para proporcionar una capa fundamental para sistemas de inteligencia cognitiva e inteligencia artificial. Consiste principalmente en conjuntos de datos de grafos de conocimiento chinos a gran escala, incluyendo datos de relación de entidades y conjuntos de entrenamiento de NLP utilizados para impulsar la comprensión semántica y la respuesta automática a preguntas.

El proyecto se centra en la construcción y exportación de grafos masivos de entidad-atributo-valor, organizando el conocimiento en formatos portátiles. Proporciona partición de dominio especializada para adaptar la recuperación de información a campos profesionales como la salud, el ejército y la seguridad pública.

El repositorio cubre una amplia gama de capacidades, incluyendo procesamiento de lenguaje natural en chino, búsqueda semántica y sistemas de diálogo cognitivo. Su conjunto de herramientas abarca análisis lingüístico, extracción de entidades, detección de sentimientos y resumen de texto, así como análisis de contenido visual para auditoría de sitios web y conversión de voz a texto.

Features

  • Knowledge Graph Construction - Assembles massive datasets of interconnected entities to create a foundational layer for cognitive artificial intelligence.
  • Chinese Natural Language Processing - Provides a natural language processing pipeline that segments text and assigns parts of speech for Chinese characters.
  • Cognitive Intelligence Bases - Provides a foundational layer of structured knowledge used to drive conversational responses and semantic understanding in AI agents.
  • Entity and Relation Extraction - Identifies and categorizes named entities and their relationships from unstructured text using linguistic patterns.
  • Structured Entity Datasets - Provides an organized dataset of entities and attributes in a format suitable for automated question answering.
  • Knowledge Graph Extraction - Implements automated processes for identifying entities and relationships to build structured knowledge representations.
  • Knowledge Graph Construction - Provides automated processes for constructing large-scale graph data structures to serve as a foundation for cognitive AI.
  • Conversational Response Generation - Combines semantic understanding with knowledge graph data to produce natural and accurate answers for dialogue agents.
  • Training Datasets - Ships a corpus of structured Chinese text used to train models for entity extraction and sentiment analysis.
  • Large Scale Knowledge Graph Datasets - Builds and exports massive entity relation datasets to provide a foundational data layer for artificial intelligence systems.
  • Domain-Specific Graph Modeling - Organizes entity-attribute data tailored for professional domains such as healthcare, military, and public security.
  • Conversational Dialogue Systems - Combines semantic understanding with knowledge graph data to power conversational agents.
  • Semantic Search - Implements search functionality that leverages knowledge graphs and semantic perception to understand user intent beyond keywords.
  • Entity-Attribute-Value Models - Implements a data model that stores knowledge as triples of entities, attributes, and values to handle heterogeneous data.
  • Knowledge Graphs - Provides a large scale collection of entity relation data used for building cognitive intelligence systems.
  • Article Tagging - The product analyzes text content to identify and apply relevant labels that categorize the subject matter for better organization.
  • Automated Text Analysis - Extracts keywords, summarizes articles, and performs sentiment analysis to surface key insights from large volumes of text.
  • Causal Relationship Mapping - Provides mechanisms to link entities through defined causal relationships to power automated question answering agents.
  • Text Parsing Tools - Segments Chinese text into words and identifies parts of speech to prepare raw text for analysis.
  • Keyword and Phrase Extraction - Identifies the most significant terms and phrases that represent the primary topic of a document.
  • Text Document Classification - Categorizes text-based documents into predefined classes using machine learning models.
  • Named Entity Recognition - Identifies and classifies named entities such as people and organizations within unstructured text.
  • Natural Language Processing - Provides libraries and techniques for analyzing, processing, and extracting insights from human language data.
  • Text Summarization - Implements methods and tools that use language models to generate concise summaries of documents.
  • Part-of-Speech Taggers - Implements systems for assigning grammatical labels to words based on context and linguistic rules.
  • Semantic Similarity Calculation - Calculates the conceptual distance between text fragments to determine meaning regardless of specific wording.
  • Sentiment Analysis Tools - Provides software for classifying the emotional tone of text as positive, negative, or neutral.
  • Targeted Entity Sentiment - Determines the emotional polarity associated with specific entities mentioned within a text.
  • Multilingual Datasets - Supplies structured entity-relation datasets in specific languages to power intelligent applications.
  • Semantic Search - Uses knowledge graphs and vector similarity to retrieve conceptually relevant information based on user intent.
  • Semantic Insight Extraction - Identifies keywords, generates summaries, and performs sentiment analysis to surface key information.
  • Chinese Language Segmenters - Provides specialized tools for tokenizing and segmenting continuous Chinese text streams.
  • Knowledge Domain Partitioning - Organizes entity data into specialized professional domains to tailor information retrieval for specific industries.
  • Knowledge Graphs - Large-scale Chinese knowledge graph dataset with billions of entities.
  • Corpus and Datasets - Large-scale Chinese knowledge graph dataset.

Historial de estrellas

Gráfico del historial de estrellas de ownthink/knowledgegraphdataGráfico del historial de estrellas de ownthink/knowledgegraphdata

Búsqueda con IA

Explora más repositorios increíbles

Describe lo que necesitas en lenguaje sencillo: la IA clasifica miles de proyectos open-source curados por relevancia.

Start searching with AI

Alternativas open-source a KnowledgeGraphData

Proyectos open-source similares, clasificados según cuántas características comparten con KnowledgeGraphData.
  • hankcs/hanlpAvatar de hankcs

    hankcs/HanLP

    36,413Ver en GitHub↗

    HanLP is a natural language processing library and deep learning framework specifically optimized for the Chinese language, while also functioning as a multilingual text processor. It serves as a toolkit for performing linguistic analysis, semantic understanding, and script conversion. The project distinguishes itself through a dedicated focus on Chinese linguistic structures, including a specialized script converter for transforming text between Simplified Chinese, Traditional Chinese, and Pinyin. It further supports domain-specific model training to improve the recognition of professional t

    Pythondependency-parserhanlpnamed-entity-recognition
    Ver en GitHub↗36,413
  • mesolitica/nlp-models-tensorflowAvatar de mesolitica

    mesolitica/NLP-Models-Tensorflow

    1,778Ver en GitHub↗

    This repository provides a collection of deep learning models and neural network architectures built for natural language processing tasks. It functions as a library of pre-trained models designed to process, analyze, and generate human language data using the TensorFlow framework. The project utilizes sequence-to-sequence modeling and layered neural architectures to handle variable-length language data. By employing static dataflow graphing and tensor-based representations, the models execute mathematical operations to transform input features into abstract linguistic meanings. Users can loa

    Jupyter Notebookattentionchatbotdeep-learning
    Ver en GitHub↗1,778
  • dongrixinyu/jionlpAvatar de dongrixinyu

    dongrixinyu/JioNLP

    3,847Ver en GitHub↗

    JioNLP is a Chinese natural language processing toolkit designed for cleaning, normalizing, and extracting structured information from unstructured text. It functions as a linguistic analyzer for Chinese characters and a rule-based named entity extractor, providing a specialized system for sentiment scoring and synthetic data generation for machine learning workflows. The project features a lexicon-based sentiment analysis engine that computes numerical emotional tone scores and a data augmentation library that uses back-translation and synonym replacement to expand training datasets. It incl

    Python
    Ver en GitHub↗3,847
  • isnowfy/snownlpAvatar de isnowfy

    isnowfy/snownlp

    6,631Ver en GitHub↗

    SnowNLP is a Python library for Chinese natural language processing. It provides tools for text segmentation, sentiment analysis, document classification, and phonetic transliteration. The library includes capabilities for training and saving custom machine learning models for tokenization and sentiment analysis using raw training datasets. It covers a range of linguistic processing areas, including parts of speech tagging, sentence splitting, and text similarity measurement. The toolkit also provides utilities for extracting key information through text summarization and calculating word im

    Python
    Ver en GitHub↗6,631
Ver las 30 alternativas a KnowledgeGraphData→

Preguntas frecuentes

¿Qué hace ownthink/knowledgegraphdata?

KnowledgeGraphData es una colección de conjuntos de datos estructurados y corpora diseñados para proporcionar una capa fundamental para sistemas de inteligencia cognitiva e inteligencia artificial. Consiste principalmente en conjuntos de datos de grafos de conocimiento chinos a gran escala, incluyendo datos de relación de entidades y conjuntos de entrenamiento de NLP utilizados para impulsar la comprensión semántica y la respuesta automática a preguntas.

¿Cuáles son las características principales de ownthink/knowledgegraphdata?

Las características principales de ownthink/knowledgegraphdata son: Knowledge Graph Construction, Chinese Natural Language Processing, Cognitive Intelligence Bases, Entity and Relation Extraction, Structured Entity Datasets, Knowledge Graph Extraction, Conversational Response Generation, Training Datasets.

¿Qué alternativas de código abierto existen para ownthink/knowledgegraphdata?

Las alternativas de código abierto para ownthink/knowledgegraphdata incluyen: hankcs/hanlp — HanLP is a natural language processing library and deep learning framework specifically optimized for the Chinese… mesolitica/nlp-models-tensorflow — This repository provides a collection of deep learning models and neural network architectures built for natural… dongrixinyu/jionlp — JioNLP is a Chinese natural language processing toolkit designed for cleaning, normalizing, and extracting structured… isnowfy/snownlp — SnowNLP is a Python library for Chinese natural language processing. It provides tools for text segmentation,… fxsjy/jieba — This project is a Chinese text segmentation library and tokenizer designed to split Chinese sentences into individual… huyingxi/synonyms — Synonyms is a Chinese natural language processing tool focused on semantic analysis. It provides capabilities for…