How this analysis was created: This summary and feature list were written by an AI model that read the project's README and public documentation pages. Each feature links to the documentation it came from; stars, license and language come straight from the GitHub API. The model does not read the source code, and the analysis is refreshed when the project is re-analysed. Learn more on our About page.
| Status | Stable | Latest | Source code|Spark compatibility| |:----------:|:-------------:|:------|:------:|:------| | GeoSpark | | | |Spark 2.X, 1.X| | GeoSparkSQL | | | | Spark SQL 2.1, 2.2| | GeoSparkViz | | | |Spark 2.X, 1.X|
ADAM is a genomics analysis platform with specialized file formats built using Apache Avro, Apache Spark, and Apache Parquet. Apache 2 licensed.
Cloud-native genomic dataframes and batch computing
The Archives Unleashed Toolkit is an open-source platform for analyzing web archives using Apache Spark, and makes use of Sparkling for parsing W/ARC records. The toolkit provides powerful tools for analytics and data processing. It is part of the Archives Unleashed Project.
The main features of archivesunleashed/aut are: Domain Specific Processing.
Open-source alternatives to archivesunleashed/aut include: apache/incubator-sedona — | Status | Stable | Latest | Source code|Spark compatibility| |:----------:|:-------------:|:------|:------:|:------|… bigdatagenomics/adam — ADAM is a genomics analysis platform with specialized file formats built using Apache Avro, Apache Spark, and Apache… graphframes/graphframes. hail-is/hail — Cloud-native genomic dataframes and batch computing. neo4j-contrib/neo4j-spark-connector — This repository contains the Neo4j Connector for Apache Spark.