Open-source data discovery and metadata catalog from Lyft
Amundsen is an open-source data discovery and metadata platform originally built at Lyft. It indexes tables, columns, dashboards and owners into a searchable catalog, and shows popularity and usage data alongside each asset so people can find the tables their colleagues actually rely on. Lineage, column descriptions and ownership records are displayed in one page, and the metadata layer is modular so individual components can be adopted on their own. It is worth knowing that development has slowed considerably in recent years, so teams evaluating it should weigh the active community and available maintenance against its solid core design.
Last updated: 2026-09-20. This site only provides an index; for exact features, pricing, and licensing, see the official website.
data discovery, metadata cataloging, lineage exploration and self-hosted governance
If you're comparing similar products, check the alternatives below, or browse all tools in the AI Data Analysis category.
Community activity has decreased noticeably compared with its early years, and much of the ecosystem has moved on. It still works and can be self-hosted, but plan for in-house maintenance capability or evaluate more actively maintained alternatives.
A metadata database, a search backend such as Elasticsearch, a graph store for lineage, and the web and metadata services, usually containerised. Expect to write ingestion jobs for the metadata sources you care about.
It answers which table to use, who owns it and what depends on it. Without that, analysts rebuild existing datasets and break downstream jobs unknowingly. That is a human problem first and a tooling problem second.
Unified lakehouse platform for data engineering, analytics and AI
Microsoft business intelligence service with Copilot-assisted reporting
Visual analytics platform known for exploratory drag-and-drop charting
Open-source BI platform you can self-host with SQL-first dashboards