4 रिपॉजिटरी
Capabilities for locating data assets and identifying their owners using semantic search and descriptions.
Distinct from Search and Discovery: Existing candidates focused on either general search libraries or specific quality instrumentation rather than asset-centric discovery.
Explore 4 awesome GitHub repositories matching data & databases · Data Asset Discoveries. Refine with filters or upvote what's useful.
OpenMetadata is an enterprise data catalog, metadata platform, and governance suite that functions as a knowledge graph for data assets. It serves as an AI-ready metadata layer, providing governed context and organizational memory to large language model agents via the Model Context Protocol. The platform distinguishes itself by capturing institutional knowledge, linking conversations, decisions, and remediation notes directly to data assets to preserve tribal knowledge. It integrates AI agents to automate metadata governance, such as suggesting descriptions and identifying sensitive data thr
Locates data through semantic search and descriptions to identify ownership and sample data.
Naabu is a port scanner library and tool that probes hosts for open ports using SYN, CONNECT, and UDP methods to identify active services. It functions as a Go library for embedding port scanning into programs, and as a standalone tool that accepts targets as hostnames, IP addresses, CIDR ranges, or ASN numbers. The tool discovers live hosts before scanning, filters ports by range or top lists, and can integrate with Nmap for service version detection. The project distinguishes itself through its SYN-based port probing approach that sends TCP SYN packets and analyzes responses without complet
Transforms raw asset data into detailed profiles through DNS analysis, HTTP probing, screenshots, and technology fingerprinting.
Amundsen is a data catalog and discovery platform that provides a centralized directory for indexing tables and dashboards. It functions as a metadata management system and search engine, allowing users to locate and understand available data assets across diverse distributed sources. The platform includes capabilities for data lineage tracking to map the origin and movement of datasets between systems. It also serves as a data profiling tool, calculating distribution and quality statistics for individual table columns to provide automated insights into the nature of the data. The system man
Locates specific tables and dashboards across an organization using a centralized searchable index.
nit एक ब्लॉकचेन एसेट प्रोवेनेंस प्लेटफ़ॉर्म और विकेंद्रीकृत एसेट रजिस्ट्री है। यह फ़ाइलों को क्रिप्टोग्राफ़िक रूप से अद्वितीय पहचानकर्ता सौंपकर और लेज़र पर उनके मूल, स्वामित्व और संशोधन इतिहास को रिकॉर्ड करके डिजिटल मीडिया के लिए हिरासत की एक सत्यापन योग्य श्रृंखला स्थापित करता है। यह प्रोजेक्ट विकेंद्रीकृत स्टोरेज के लिए IPFS और एक कंटेंट वर्ज़निंग सिस्टम को एकीकृत करके खुद को अलग करता है जो अपरिवर्तनीय कमिट के माध्यम से एसेट विकास को ट्रैक करता है। इसमें जेनरेटिव AI प्रोवेनेंस ट्रैकिंग के लिए विशेष टूलिंग शामिल है, जो पारदर्शी मेटाडेटा ट्री बनाए रखने के लिए सिंथेटिक मीडिया में उपयोग किए गए रचनाकारों और टूल को रिकॉर्ड करने की अनुमति देती है। सिस्टम डिजिटल अधिकार प्रबंधन, स्मार्ट कॉन्ट्रैक्ट्स के माध्यम से स्वचालित रॉयल्टी वितरण और कंटेंट प्रामाणिकता सत्यापन सहित क्षमताओं की एक विस्तृत श्रृंखला को कवर करता है। यह एक टोकन-भारित गवर्नेंस मॉडल भी लागू करता है जहां उपयोगकर्ता विकेंद्रीकृत वोटिंग के माध्यम से प्रोटोकॉल दिशा को प्रभावित करने के लिए टोकन स्टेक कर सकते हैं। प्रोजेक्ट TypeScript में विकसित किया गया है।
Enriches asset profiles with AI-generated metadata and creator information for detailed discovery.