awesome-repositories.com
ब्लॉग
awesome-repositories.com

AI-संचालित खोज के साथ बेहतरीन ओपन-सोर्स रिपॉजिटरी खोजें।

एक्सप्लोर करेंक्यूरेटेड खोजेंओपन-सोर्स विकल्पसेल्फ-होस्टेड सॉफ्टवेयरब्लॉगसाइटमैप
प्रोजेक्टहमारे बारे मेंहम रैंकिंग कैसे करते हैंप्रेसMCP सर्वर
कानूनीगोपनीयताशर्तें
© 2026 Bringes Technology SRL·VAT RO45896025·hello@awesome-repositories.com
·
apache avatar

apache/hbase

0
View on GitHub↗
5,540 स्टार्स·3,397 फोर्क्स·Java·Apache-2.0·13 व्यूज़hbase.apache.org↗

Hbase

HBase एक वितरित, वाइड-कॉलम NoSQL स्टोर और बिग डेटा स्टोरेज इंजन है जिसे स्पार्स डेटासेट के लिए डिज़ाइन किया गया है। यह विशाल मात्रा में स्ट्रक्चर्ड और अनस्ट्रक्चर्ड डेटा तक रीयल-टाइम रीड और राइट एक्सेस प्रदान करने के लिए Hadoop Distributed File System के ऊपर निर्मित एक स्केलेबल कॉलम-आधारित डेटाबेस के रूप में कार्य करता है।

सिस्टम एक क्रॉस-लैंग्वेज डेटाबेस गेटवे के रूप में कार्य करता है, जो नेटिव रिमोट प्रोसीजर कॉल्स, REST और Thrift इंटरफेस के माध्यम से कनेक्टिविटी प्रदान करता है। यह एक मास्टर-वर्कर समन्वय मॉडल के माध्यम से खुद को अलग करता है जो एक क्लस्टर में क्षैतिज स्केलिंग और फॉल्ट टॉलरेंस को सक्षम बनाता है।

प्रोजेक्ट सेल-लेवल विजिबिलिटी लेबल्स, प्लगेबल डेटा कम्प्रेशन और सर्वर-साइड डेटा एग्रीगेशन के माध्यम से फाइन-ग्रेन्ड एक्सेस कंट्रोल सहित क्षमताओं के एक व्यापक सेट को कवर करता है। यह मैप-रिड्यूस एकीकरण के माध्यम से बिग डेटा एनालिटिक्स वर्कफ़्लो का भी समर्थन करता है और कस्टम सर्वर-साइड लॉजिक के निष्पादन की अनुमति देता है।

ऑपरेशनल मॉनिटरिंग सिस्टम मेट्रिक ट्रैकिंग और प्लगइन-आधारित मेट्रिक एक्सपोर्टिंग के माध्यम से प्रदान की जाती है।

Features

  • Columnar Databases - Implements a distributed NoSQL wide-column store built on top of the Hadoop ecosystem for sparse datasets.
  • Big Data Storage - Functions as a distributed engine for storing and querying massive volumes of structured and unstructured data.
  • Column Family Management - Organizes sparse data into grouped column families for efficient distributed storage and retrieval.
  • Hadoop - Integrates with the Hadoop Distributed File System to provide a columnar store for large-scale data analysis.
  • Distributed File System Backends - Relies on the Hadoop Distributed File System for durable, replicated persistent storage of data files.
  • Sparse Dataset Management - Provides scalable storage and versioning for massive, sparse, column-oriented datasets across a cluster.
  • LSM-Tree Storage Engines - Utilizes an LSM-tree storage engine to provide high write throughput via in-memory buffering and sorted flushes.
  • Region-Based Partitioning - Implements region-based partitioning by splitting the sorted keyspace into contiguous ranges for horizontal scaling.
  • Wide-Column Stores - Organizes data into column families to provide real-time read and write access to high-scale datasets.
  • Column-Oriented Disk Storage - Organizes sparse datasets into column-oriented disk storage for scalable, versioned data management.
  • Distributed Data Stores - Provides a cluster-based storage system with horizontal scaling and fault tolerance for scalable data retrieval.
  • Cell-Level Controls - Enforces fine-grained access control using visibility labels at the individual data cell level.
  • Master-Worker Coordination - Employs a master-worker coordination model to manage cluster metadata and region assignments.
  • Distributed Storage Clusters - Implements a scalable architecture that aggregates multiple nodes into a unified storage system for massive datasets.
  • Distributed File Systems - Relies on a distributed file system like HDFS for durable and replicated storage of underlying data files.
  • Big Data Processing - Supports big data processing workflows using map-reduce patterns for large-scale data transformation.
  • Cross-Language Data Interfaces - Provides consistent data interaction interfaces via native RPC, REST, and Thrift APIs for clients in multiple programming languages.
  • MapReduce Processing Engines - Integrates with MapReduce processing engines to transform and migrate large volumes of data between tables.
  • Server-Side Aggregations - Calculates summaries and statistics directly on the server to minimize data transfer to the client.
  • Application REST API Gateways - Exposes database operations and cluster status through a standardized REST API gateway.
  • Thrift RPC Servers - Ships a dedicated Thrift server to enable cross-language connectivity for database operations.
  • Cross-Language Service Gateways - Acts as an entry point that translates REST, Thrift, and RPC requests into internal database protocols.
  • Storage Block Compression - Applies pluggable block compression to reduce the physical storage footprint of datasets on disk.
  • Multi-Protocol Communication Bridges - Provides a multi-protocol gateway allowing clients to connect via RPC, HTTP, and Thrift.
  • Remote Procedure Calls - Uses remote procedure calls for low-latency communication between clients, master nodes, and region servers.
  • Remote Procedure Call Protocols - Implements structured messaging protocols for standardized communication between cluster nodes and clients.
  • Database Systems - Distributed big data store modeled after Bigtable.

स्टार हिस्ट्री

apache/hbase के लिए स्टार हिस्ट्री चार्टapache/hbase के लिए स्टार हिस्ट्री चार्ट

AI सर्च

और अधिक बेहतरीन रिपॉजिटरी खोजें

अपनी ज़रूरत को सरल भाषा में बताएं — AI हजारों क्यूरेटेड ओपन-सोर्स प्रोजेक्ट्स को प्रासंगिकता के आधार पर रैंक करता है।

Start searching with AI

Hbase के ओपन-सोर्स विकल्प

समान ओपन-सोर्स प्रोजेक्ट्स, जो Hbase के साथ साझा की गई सुविधाओं के आधार पर रैंक किए गए हैं।
  • apache/hadoopapache का अवतार

    apache/hadoop

    15,567GitHub पर देखें↗

    Hadoop is a big data infrastructure suite and distributed data processing framework designed to store and process massive datasets across clusters of computers. It consists of a distributed storage system for managing large files across multiple nodes and a parallel computing engine for processing data across a distributed cluster. The framework implements a distributed file system to ensure fault tolerance and high throughput, paired with a programming model that processes large datasets in parallel. It manages the underlying hardware and software environment required for distributed big dat

    Java
    GitHub पर देखें↗15,567
  • apache/hiveapache का अवतार

    apache/hive

    6,012GitHub पर देखें↗

    Apache Hive is a SQL-on-Hadoop data warehouse that enables querying and managing petabytes of data stored in distributed storage such as HDFS and cloud storage services. It provides a familiar SQL interface for batch analytics and reporting, supported by a core set of components including the HiveServer2 Thrift service for remote query execution, the Hive Metastore Service for central metadata management, the Hive ACID Transaction Engine for concurrent read-write operations, and the Hive LLAP Interactive Engine for low-latency analytical processing. The WebHCat REST API offers an HTTP interfac

    Javaapachebig-datadatabase
    GitHub पर देखें↗6,012
  • ravendb/ravendbravendb का अवतार

    ravendb/ravendb

    3,961GitHub पर देखें↗

    RavenDB is a multi-model NoSQL document database designed for high-performance, ACID-compliant data storage. It persists structured information as schema-flexible JSON documents and utilizes a unit-of-work session pattern to track entity changes and batch modifications into atomic transactions. The platform is built on a distributed architecture that supports horizontal scaling through sharding and ensures high availability via multi-node, master-to-master cluster replication. The database distinguishes itself through a self-optimizing query engine that automatically creates and maintains ind

    C#csharpdatabasedocument-database
    GitHub पर देखें↗3,961
  • deepseek-ai/3fsdeepseek-ai का अवतार

    deepseek-ai/3FS

    9,970GitHub पर देखें↗

    3FS is a distributed file system and RDMA storage cluster designed for high-performance AI training and inference workloads. It functions as a strongly consistent storage layer that utilizes a disaggregated architecture to pool SSDs and memory resources across multiple nodes. The system provides specialized storage implementations including an AI training checkpoint store for parallel state preservation and a distributed key-value cache store for decoder layer vectors to optimize inference processing. It ensures data integrity through chain replication and apportioned query distribution. The

    C++
    GitHub पर देखें↗9,970
Hbase के सभी 30 विकल्प देखें→

अक्सर पूछे जाने वाले प्रश्न

apache/hbase क्या करता है?

HBase एक वितरित, वाइड-कॉलम NoSQL स्टोर और बिग डेटा स्टोरेज इंजन है जिसे स्पार्स डेटासेट के लिए डिज़ाइन किया गया है। यह विशाल मात्रा में स्ट्रक्चर्ड और अनस्ट्रक्चर्ड डेटा तक रीयल-टाइम रीड और राइट एक्सेस प्रदान करने के लिए Hadoop Distributed File System के ऊपर निर्मित एक स्केलेबल कॉलम-आधारित डेटाबेस के रूप में कार्य करता है।

apache/hbase की मुख्य विशेषताएं क्या हैं?

apache/hbase की मुख्य विशेषताएं हैं: Columnar Databases, Big Data Storage, Column Family Management, Hadoop, Distributed File System Backends, Sparse Dataset Management, LSM-Tree Storage Engines, Region-Based Partitioning।

apache/hbase के कुछ ओपन-सोर्स विकल्प क्या हैं?

apache/hbase के ओपन-सोर्स विकल्पों में शामिल हैं: apache/hadoop — Hadoop is a big data infrastructure suite and distributed data processing framework designed to store and process… apache/hive — Apache Hive is a SQL-on-Hadoop data warehouse that enables querying and managing petabytes of data stored in… ravendb/ravendb — RavenDB is a multi-model NoSQL document database designed for high-performance, ACID-compliant data storage. It… gluster/glusterfs — GlusterFS is a software-defined distributed file system and scale-out storage cluster that aggregates disk resources… deepseek-ai/3fs — 3FS is a distributed file system and RDMA storage cluster designed for high-performance AI training and inference… hazelcast/hazelcast — Hazelcast is a distributed data platform that combines an in-memory data grid with a stream processing engine to…