Skip to main content

Amundsen Alternatives – DataHub, Metacat, and Apache Atlas

Emily Winks, Data Governance Expert, Atlan
Data Governance Expert
Updated:
|
Published:
4 min read

Key takeaways

  • Amundsen was archived in September 2026 and its GitHub repo is read-only: no releases, no security patches
  • DataHub and OpenMetadata are the maintained open-source options; Metacat and Apache Atlas are narrower
  • Each alternative has distinct strengths in metadata management, lineage, and governance
  • Choosing the right tool depends on your data stack, governance needs, and team resources

Quick Answer: What are the Best Amundsen Alternatives?

Amundsen was archived in September 2026, so the question is where to move rather than whether to stay. DataHub by LinkedIn and OpenMetadata are the maintained open-source options: DataHub offers stream-based metadata ingestion with Kafka and 70+ connectors, and OpenMetadata ships discovery, lineage, and quality testing in one platform. Metacat federates metadata access across diverse data stores at Netflix scale; Apache Atlas focuses on Hadoop-ecosystem governance with classification hierarchies and Apache Ranger integration. Atlan is a managed alternative built on a hardened Atlas fork, adding automated context generation, governance workflows, and 100+ cloud-native connectors.

Key components:

  • DataHub by LinkedIn for generalized metadata search and discovery
  • OpenMetadata for a maintained catalog with built-in quality testing and lineage
  • Metacat by Netflix for federated metadata access across data stores
  • Apache Atlas for Hadoop ecosystem governance and classification
  • Atlan as a managed alternative with automated context generation, governance workflows, and 100+ connectors

Ready to see how Atlan compares?

See Context Layer in Action

What are some alternatives to Amundsen?

Status, September 2026: Amundsen is archived. The Amundsen repository carries the notice “Due to inactivity, this project was archived in September 2026. The contents will remain available for historical purposes.” GitHub records the archive on September 10, 2026. The code stays available under Apache-2.0, but the repo is read-only: no new releases, no security patches, no maintainer review of issues or pull requests. So the question on this page is no longer whether Amundsen is good enough. It is where you move. For a maintained open-source catalog, start with DataHub or OpenMetadata.

Amundsen was an open-source data discovery platform and metadata engine built by the Lyft engineering team.

Introduced and open-sourced in 2019 for adoption outside of Lyft, the platform was initially built to improve the productivity of engineers, data scientists, and analysts at the ride-hailing app.


Adoption inside Lyft ran high, and the project drew an open-source community of ~100 contributors and 39 organizations before activity dried up.

If you run Amundsen today, or you were about to stand it up, these are the projects to look at instead.


Is Open Source really free? Estimate the cost of deploying an open-source data catalog 👉 Download Free Calculator


3 open source Amundsen alternatives


  • LinkedIn’s DataHub
  • Netflix’s Metacat
  • Hortonworks’ Apache Atlas

A Guide to Building a Business Case for a Data Catalog

Download Ebook

DataHub: what it means in practice

Built by LinkedIn, DataHub is an open-source metadata search and discovery tool.

Open-sourced in 2020, this tool is LinkedIn’s second attempt at solving cataloging and data discovery problems. Its first attempt, WhereHows, was in 2016.

Popular use cases of DataHub include:

  • Automating metadata ingestion from multiple data sources
  • Streamlining data discovery via searching and browsing data assets
  • Enhancing understanding of data with context

Further reading about DataHub as an Amundsen alternative:




Metacat: what it means in practice

An open-source metadata management platform, Metacat powers data discovery and metadata interoperability at Netflix.

With a single access layer for data across a diverse mesh of data sources, Metacat simplifies data discovery, cataloging, processing, and management.

Metacat’s best-known capabilities include:

  • Easy data discovery
  • Data change notifications
  • One common abstraction layer
  • Provisions for the user-and business-defined metadata storage

Further reading about Metacat as an Amundsen alternative:



Apache Atlas: what it means in practice

Apache Atlas is a popular open-source software used to build catalogs of data assets.

With an active community of committers from Hortonworks, Aetna, Merck, IBM, Target, and more leading companies, the Apache Atlas project expands year by year.

Apache Atlas has the following main capabilities:

  • Visualizing metadata lineage
  • Adding entities to metadata to streamline searches
  • Creating classifications for data

Further reading about Apache Atlas as an Amundsen alternative:


Evaluating open source data catalog tools

One of the crucial steps in enabling a collaborative and efficient data culture in your organization is deploying a data catalog software. Discovering the right one for your organization requires answering many different questions — all at once.

  • Will this platform support all our primary use cases?
  • Will it still work if our data stack changes?
  • Will our users feel comfortable using it?
  • Should we build it or buy it?
  • How do we justify the money we’re asking for?

We know thinking about all of that can be overwhelming. It helps to have a framework to objectively evaluate. Here’s how we could help:

The Ultimate Guide to Evaluating an Enterprise Data Catalog

Download Ebook

Share this article

signoff-panel-logo

Atlan is the Context Layer for AI — the infrastructure that connects business definitions, lineage, quality signals, and governance policies across 100+ source systems into one traversable graph. Human teams use Atlan's Data Marketplace for conversational search, automated governance, and self-service data products. AI agents query the same graph via MCP, SQL, and API to get the context they need before they act on enterprise data.

Bridge the context gap.
Ship AI that works.