3 repository-uri
Systems that arbitrate simultaneous write attempts to prevent data loss and maintain metadata synchronization.
Distinct from Atomic Write Coordinators: Focuses on arbitrating simultaneous writes via a catalog rather than grouping tasks into a single transaction.
Explore 3 awesome GitHub repositories matching software engineering & architecture · Concurrent Write Coordinators. Refine with filters or upvote what's useful.
This project is a collection of Python scripts and tools designed for web scraping, browser automation, and large-scale data extraction. It provides a set of implementations for retrieving information from websites and private APIs, including tools for multimedia downloading and social media data archiving. The toolset includes specialized mechanisms for bypassing anti-scraping measures through IP proxy pool rotation and multi-threaded crawlers. It also features capabilities for simulating browser sessions to handle authentication, intercepting session cookies, and decrypting network payloads
Employs thread locks to synchronize concurrent writes from multiple workers to shared CSV files.
Delta is a lakehouse table format that brings ACID transactions and data warehouse consistency to large scale data lakes on cloud object storage. It serves as an ACID transaction manager, coordinating atomic commits and serializable isolation for concurrent reads and writes across distributed compute engines. The project provides a multi-engine interoperability layer that uses format translation to allow diverse SQL engines and processing frameworks to read and write the same tables. It functions as a data versioning system, utilizing a transaction log to enable time travel, historical snapsh
Arbitrates simultaneous write attempts via a centralized catalog to prevent data loss.
Noms is a distributed version control database and content-addressable data store. It identifies data by cryptographic hashes to ensure integrity and deduplication, while tracking dataset state changes through a sequence of immutable commits to enable branching, forking, and historical recovery. The system functions as a peer-to-peer data synchronizer, reconciling state between disconnected database instances to ensure all nodes converge on the same data. It distinguishes itself as a schema-flexible document store that supports self-describing types, allowing schemas to evolve and widen as ne
Uses concurrent write coordinators and optimistic locking to prevent data corruption during simultaneous insertions.