awesome-repositories.com
Blog
MCP
awesome-repositories.com

Discover the best open-source repositories with AI-powered search.

ExploreCurated searchesOpen-source alternativesSelf-hosted softwareBlogSitemap
ProjectMCP serverAboutHow we rankPress
LegalPrivacyTerms
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
datacharmer avatar

datacharmer/test_db

0
View on GitHub↗
4,388 stars·2,693 forks·Shell·17 views

Test Db

test_db is a collection of tools for validating database integrity, benchmarking system throughput, and generating synthetic schemas and datasets. It includes a sample corporate employee database for MySQL, a SQL dataset generator for creating representative records, and an integrity validator that uses checksums and record counts to verify data consistency across different database engines.

The project provides a database performance benchmark consisting of complex queries and stored procedures designed to measure system response times and throughput. These tools simulate real-world workloads through multi-table joins and aggregations.

The toolset covers broad capability areas including database schema simulation, synthetic data generation, and data load testing. It specifically measures data load duration by calculating the elapsed time from the initial table creation to the final update.

Features

  • Database Integrity Verification - Provides a tool that uses checksums and record counts to verify data consistency across different database engines.
  • General Synthetic Data Generators - Populates database tables with representative records to simulate real-world workloads for software testing.
  • Sample Datasets - Provides scripts and pre-defined datasets for populating database tables for performance benchmarking.
  • Environment Simulation Schemas - Creates tables, views, and stored procedures to simulate corporate database environments for testing.
  • Database Population Tools - Provides tools for populating database tables with large, pre-configured datasets from external dumps.
  • Dataset Generation Scripts - Provides a series of scripts that populate database tables with representative records to simulate real-world workloads.
  • Join Performance Benchmarking - Evaluates system throughput by executing complex aggregation queries that force heavy resource usage across multiple tables.
  • Environment Simulation - Creates tables, views, and stored procedures to build consistent environments for testing application logic.
  • Data Integrity Verification - Uses checksums and record counts to ensure datasets remain consistent across different database engines.
  • Data Integrity Checksums - Implements checksum-based verification to ensure data consistency across different database engines.
  • Query Performance Benchmarks - Measures system response times and throughput using complex queries and aggregations on standardized datasets.
  • Load Testing - Evaluates the efficiency of data ingestion by measuring the time required to create schemas and populate tables.
  • Logic Deployment - Enables the deployment of stored procedures and functions to simulate complex server-side business logic.
  • Sample Databases - Ships a sample corporate employee database for MySQL used for testing and benchmarking.
  • SQL Script Execution - Executes predefined SQL scripts from external files to provision database tables and views.
  • Stored Procedures - Deploys complex functions and stored procedures to simulate server-side business logic.
  • Data Load Duration Measurement - Calculates the total elapsed time from initial table creation to final update to evaluate load efficiency.
  • Data Load Durations - Calculates the total elapsed time from initial table creation to the final update to evaluate data ingestion efficiency.

Star history

Star history chart for datacharmer/test_dbStar history chart for datacharmer/test_db

How this analysis was created: This summary and feature list were written by an AI model that read the project's README and public documentation pages. Each feature links to the documentation it came from; stars, license and language come straight from the GitHub API. The model does not read the source code, and the analysis is refreshed when the project is re-analysed. Learn more on our About page.

AI search

Explore more awesome repositories

Describe what you need in plain English — the AI ranks thousands of curated open-source projects by relevance.

Start searching with AI

Open-source alternatives to Test Db

Similar open-source projects, ranked by how many features they share with Test Db.
  • joke2k/fakerjoke2k avatar

    joke2k/faker

    19,278View on GitHub↗

    Faker is a Python library designed to generate realistic synthetic data for software testing, database prototyping, and privacy-preserving anonymization. It provides a comprehensive suite of tools to create diverse information types, including personal identities, financial records, geographic locations, and technical system metadata, allowing developers to populate environments with mock data that mimics real-world structures. The library is built on a modular provider architecture that supports dynamic method dispatch, enabling users to extend functionality by registering custom data genera

    Pythondatasetfakefake-data
    View on GitHub↗19,278
  • pawelsalawa/sqlitestudiopawelsalawa avatar

    pawelsalawa/sqlitestudio

    6,428View on GitHub↗

    SQLiteStudio is an open-source graphical tool for browsing, editing, and managing SQLite database files. It combines a full-featured SQL editor with syntax highlighting, a visual database schema designer for creating entity-relationship diagrams, and a plugin-based extensibility platform that allows adding custom functionality through C/C++, JavaScript, Tcl, or Python. The application distinguishes itself through its multi-language scripting engine, which embeds JavaScript, Tcl, and Python interpreters to enable user-defined functions and scripts within SQL queries. It supports encrypted data

    Ccppdatabasedatabase-management
    View on GitHub↗6,428
  • h2database/h2databaseh2database avatar

    h2database/h2database

    4,607View on GitHub↗

    H2 is a JDBC-compliant relational database management system written in Java. It functions as an embeddable SQL database that can run directly within an application process to remove network latency, or as an in-memory database for high-performance volatile storage. It also includes a web-based console for executing SQL commands and administering schemas. The system is characterized by its flexible deployment modes, including a standalone server mode for remote TCP/IP access and a mixed mode for simultaneous local and remote connectivity. It features a dialect emulation layer and compatibilit

    Javadatabasejavajdbc
    View on GitHub↗4,607
  • prestodb/prestoprestodb avatar

    prestodb/presto

    16,711View on GitHub↗

    Presto is a distributed SQL query engine designed for high-performance analytical processing across heterogeneous data sources. It functions as a data federation platform and massively parallel processing engine, allowing users to execute interactive queries against diverse storage systems without requiring data migration. By mapping remote metadata and structures to a unified relational namespace, it enables seamless cross-platform analysis through a standard SQL interface. The engine distinguishes itself through a pluggable connector architecture and a shared-nothing distributed processing

    Javabig-datadatahadoop
    View on GitHub↗16,711
See all 30 alternatives to Test Db→

Frequently asked questions

What does datacharmer/test_db do?

test_db is a collection of tools for validating database integrity, benchmarking system throughput, and generating synthetic schemas and datasets. It includes a sample corporate employee database for MySQL, a SQL dataset generator for creating representative records, and an integrity validator that uses checksums and record counts to verify data consistency across different database engines.

What are the main features of datacharmer/test_db?

The main features of datacharmer/test_db are: Database Integrity Verification, General Synthetic Data Generators, Sample Datasets, Environment Simulation Schemas, Database Population Tools, Dataset Generation Scripts, Join Performance Benchmarking, Environment Simulation.

What are some open-source alternatives to datacharmer/test_db?

Open-source alternatives to datacharmer/test_db include: joke2k/faker — Faker is a Python library designed to generate realistic synthetic data for software testing, database prototyping,… pawelsalawa/sqlitestudio — SQLiteStudio is an open-source graphical tool for browsing, editing, and managing SQLite database files. It combines a… h2database/h2database — H2 is a JDBC-compliant relational database management system written in Java. It functions as an embeddable SQL… prestodb/presto — Presto is a distributed SQL query engine designed for high-performance analytical processing across heterogeneous data… lk-geimfari/mimesis — Mimesis is a Python synthetic data generator used to create realistic fake datasets and mock data for software testing… stympy/faker — Faker is a synthetic data generation library used to create realistic but fake information, such as names, addresses,…