1 Repo
Utilities for converting complex annotation objects into simplified or flattened formats.
Distinct from Annotation Metadata Transformers: Distinct from Annotation Metadata Transformers: focuses on flattening and simplifying data structures for consumption rather than modifying bean class metadata.
Explore 1 awesome GitHub repository matching development tools & productivity · Annotation Data Structuring. Refine with filters or upvote what's useful.
Spark NLP is a toolkit for scalable text analysis and machine learning built on the Apache Spark distributed computing framework. It provides a multimodal machine learning framework and a distributed pipeline system for sequencing annotators to process large-scale linguistic data. The library includes a transformer text processor for generating contextual vector embeddings and a dedicated inference engine for managing large language models. The project distinguishes itself through its ability to process heterogeneous data types, including text, audio, and images, within a unified vision-langu
Provides utilities to flatten and simplify complex annotation objects into arrays or dataframes for easier consumption.