Data lineage tools for Amazon S3
Data lineage tools are software that allows to extract, view and analyze data lineage. Data lineage is the process of understanding and visualizing data flow from the source to different destinations. It allows to create a map of the data journey through the entire ecosystem.
Dataedo
Dataedo allows you to extract lineage automatically or design flows manually and visualize how data moves through the system with interactive diagrams. Dataedo supports object and column-level data lineage. It improves transparency, supports impact analysis, and ensures data integrity across an organization.
BI Tools lineage: |
|
---|---|
Commercial: | Commercial |
Data migration tools lineage: |
|
Data warehouses lineage: |
|
ETLs: |
|
Free edition: |
|
Hadoop: |
|
NoSQL: |
|
Pipelines lineage: |
|
RDBMS: |
|
Global IDs
Global IDs Data Lineage provides automated analysis of the actual flow of data through your enterprise, enabling you to understand – in real-time – where data originates, how it flows through the ecosystem, and how it is transformed en route.
BI Tools lineage: |
|
---|---|
Commercial: | Commercial |
Data migration tools lineage: |
|
Data warehouses lineage: |
|
ETLs: |
|
Free edition: |
|
Hadoop: |
|
NoSQL: |
|
Pipelines lineage: |
|
RDBMS: |
|
Alteryx Connect
Alteryx Connect uses powerful search capabilities to find and reuse information contained in data files, databases, visualizations, dashboards, workflows, analytic apps, and more. It lets you automatically capture and visualize data lineage between assets, improving the overall quality and reliability of shared information between data, process, and people. You can get technical data lineage by loading metadata from source and target systems and interpreting Alteryx workflows.
BI Tools lineage: |
|
---|---|
Commercial: | Commercial |
Data migration tools lineage: |
|
Data warehouses lineage: |
|
ETLs: |
|
Free edition: |
|
Hadoop: |
|
NoSQL: |
|
Pipelines lineage: |
|
RDBMS: |
|
IBM Watson Knowledge Catalog
IBM Watson Knowledge Catalog is an open and intelligent data catalog for managing enterprise data that also lets you ensure well-structured and maintained data lineage. It lets you track where data originated and how it’s consumed, increasing trust when accessing data across many sources and destinations
BI Tools lineage: |
|
---|---|
Commercial: | Commercial |
Data migration tools lineage: |
|
Data warehouses lineage: |
|
ETLs: |
|
Free edition: |
|
Hadoop: |
|
NoSQL: |
|
Pipelines lineage: |
|
RDBMS: |
|
OvalEdge
OvalEdge offers a comprehensive lineage solution to show a complete the complete data cycle. OvalEdge algorithms parse various kinds of source code to build the lineage automatically and then it is enhanced by experts with proper descriptions.
BI Tools lineage: |
|
---|---|
Commercial: | Commercial |
Data migration tools lineage: |
|
Data warehouses lineage: |
|
ETLs: |
|
Free edition: |
|
Hadoop: |
|
NoSQL: |
|
Pipelines lineage: |
|
RDBMS: |
|
Truedat
Truedat is an open source data governance business solution tool that lets you have an end to end vision of your data from a business and technical point of view. Truedat data lineage module allows the visualization of the information life cycle, as well as the interconnection between each system of the organization, which allows to have a complete traceability of the data, as well as impact analysis in the event of possible changes in data structures or processes.
BI Tools lineage: |
|
---|---|
Commercial: | Free |
Data migration tools lineage: |
|
Data warehouses lineage: |
|
ETLs: |
|
Free edition: |
|
Hadoop: |
|
NoSQL: |
|
Pipelines lineage: |
|
RDBMS: |
|
Talend Data Catalog
Talend Data Catalog gives your organization a single, secure point of control for your data. Its data flow lineage feature allows you to narrow in on specific objects and shows you how these objects are related to each other, within a model, an external metadata repository, or a configuration. The data flow lineage is based upon connection definitions to data stores and physical transformation rules which transform and move the data.
BI Tools lineage: |
|
---|---|
Commercial: | Commercial |
Data migration tools lineage: |
|
Data warehouses lineage: |
|
ETLs: |
|
Free edition: |
|
Hadoop: |
|
NoSQL: |
|
Pipelines lineage: |
|
RDBMS: |
|
Secoda
Secoda is a data discovery tool that offers an intuitive, collaborative, and easy to implement data discovery built. It automatically extracts queries to generate data lineage. Currently, it is supporting table lineage for Snowflake, dbt, Redshift, and BigQuery, with support for Postgres, MySQL, and Microsoft SQL Server coming soon. Secoda data lineage can help data teams identify the downstream and upstream dependencies of a table easily. On each dependency, you will be able to see how many levels away a particular table is, with the ability to view the data in a visual form coming soon.
BI Tools lineage: |
|
---|---|
Commercial: | Commercial |
Data migration tools lineage: |
|
Data warehouses lineage: |
|
ETLs: |
|
Free edition: |
|
Hadoop: |
|
NoSQL: |
|
Pipelines lineage: |
|
RDBMS: |
|
Cloudera Navigator
Cloudera Navigator is the complete data governance solution for Hadoop. Cloudera Navigator automatically collects audit logs from across the entire platform and maintains a full history, with a unified, searchable audit dashboard for simple, point-in-time visibility. With automatic collection and visualization of column-level lineage, users can also quickly identify the origin, usage, and impact of a dataset.
BI Tools lineage: |
|
---|---|
Commercial: | Commercial |
Data migration tools lineage: |
|
Data warehouses lineage: |
|
ETLs: |
|
Free edition: |
|
Hadoop: |
|
NoSQL: |
|
Pipelines lineage: |
|
RDBMS: |
|
erwin Data Catalog
erwin Data Catalog automates enterprise metadata management, data mapping, code generation, and data lineage for faster time to value and greater accuracy for data movement and deployment projects. It lets you generate on-demand lineage down to the column level and visualize data flows from source systems all the way to the reporting layers, including all transformations. Fully configurable and navigable lineage diagrams provide high-level business views as well as detailed technical depictions.
BI Tools lineage: |
|
---|---|
Commercial: | Commercial |
Data migration tools lineage: |
|
Data warehouses lineage: |
|
ETLs: |
|
Free edition: |
|
Hadoop: |
|
NoSQL: |
|
Pipelines lineage: |
|
RDBMS: |
|
IBM InfoSphere Information Governance Catalog
IBM InfoSphere Information Governance Catalog is a web-based tool that allows you to explore, understand, and analyze information. It lets you run data lineage to create trusted information that supports data governance and compliance efforts. You can perform lineage analysis to understand where data comes from or goes to by using shared table information, job design information, or operational metadata from job runs.
BI Tools lineage: |
|
---|---|
Commercial: | Commercial |
Data migration tools lineage: |
|
Data warehouses lineage: |
|
ETLs: |
|
Free edition: |
|
Hadoop: |
|
NoSQL: |
|
Pipelines lineage: |
|
RDBMS: |
|
Alation Data Catalog
Alation is a powerful data lineage tool that helps organizations visualize and understand the flow of data across various systems and processes. It automatically captures metadata, tracks how data moves and transforms from source to destination, and offers intuitive visualizations that make it easier for users to trace data origins, transformations, and usage.
BI Tools lineage: |
|
---|---|
Commercial: | Commercial |
Data migration tools lineage: |
|
Data warehouses lineage: |
|
ETLs: |
|
Free edition: |
|
Hadoop: |
|
NoSQL: |
|
Pipelines lineage: |
|
RDBMS: |
|
Data lineage forms the foundation for accurate data analytics and management. The core features of data lineage focuses on:
• Identifying data quality issues.
• Performing root cause analysis.
• Enabling to understand which data sources are outdated or which datasets are relevant.
• Minimizing the risk of migration projects.
• Providing transparency over the life cycle of data.
Data lineage tools map the data flow and help you understand where the data originated, how it flows and transforms. To help you find the right tool for your company, we prepared a list that includes some of the best data lineage tools.