datahub-project/datahub

The Context Platform for your Data and AI Stack

View on GitHub ↗Jump to charts ↓

Summary Information

Updated 45 minutes ago
Added to GitGenius on September 4th, 2026
Created on November 18th, 2015
Open Issues & Pull Requests: 1,283 (+2)
GitHub issues: Enabled
Number of forks: 3,686
Total Stargazers: 12,644 (+0)
Total Subscribers: 251 (+0)

Repository Insights (GitGenius)

Most active contributors

Sign in to see contributor activity.

Related repositories by overlapping contributors

No overlapping-contributor repos identified yet.

Charts & Analytics

Fetching additional details & charts...

Issue Activity (beta)

Open issues: 508
New in 7 days: 10
Closed in 7 days: 8
Avg open age: 411 days
Stale 30+ days: 438
Stale 90+ days: 345

Recent activity

Opened in 7 days: 9
Closed in 7 days: 6
Comments in 7 days: 7
Events in 7 days: 12

Top labels

  • bug (547)
  • stale (159)
  • ingestion (34)
  • accepted (33)
  • feature-request (24)
  • datahub-v1.0-rc (15)
  • product (13)
  • internal (8)

Detailed Description

DataHub is a metadata platform and data catalog that enables discovery, governance, and observability across data ecosystems. The tool solves the problem of fragmented metadata across complex data stacks by providing a centralized platform where teams can search, discover, and understand their data assets. It works by ingesting metadata from various data sources and systems, then surfacing that information through a unified interface that supports both discovery and governance workflows.

Organizations should adopt DataHub when they need enterprise-grade metadata management across multiple data systems and want to enable self-service data discovery for their teams. It suits projects ranging from mid-sized data operations to large enterprises managing complex data ecosystems with many interconnected sources. The platform supports data observability use cases, allowing teams to track data quality and lineage alongside discovery and governance capabilities.

The project maintains active development with regular updates to its core functionality. The codebase shows consistent refinement of existing features and expansion of integrations with external data platforms. Community engagement remains strong through multiple channels including documentation, demonstrations, and collaborative spaces. The tool continues to evolve its capabilities for handling metadata at scale across diverse data environments.