---
title: "An integrated energy company built one metadata framework before moving 110 million files | DataXray"
description: "The company was moving more than 110 million unstructured files into a new data catalog. It needed to know what was in them first."
url: "https://www.dataxray.io/case-studies/enterprise-wide-data-classification-for-legacy-file-migration/"
language: "en-US"
---

1. [Home](https://www.dataxray.io/)
2. [Resources](https://www.dataxray.io/resources/)
3. [Case studies](https://www.dataxray.io/resources/?type=case-studies)
4. Energy metadata framework

CASE STUDY

# An integrated energy company built one metadata framework before moving 110 million files.

The company was moving more than 110 million unstructured files into a new data catalog. It needed to know what was in them first.

110M+files from a legacy content system

$800Ksaved per petabyte deleted or archived

Millionsin expected multi-year savings

CLIENTIntegrated energy company, North America

DATA ESTATEWindows file servers, Microsoft 365, OpenText Livelink, and an Azure data warehouse

DEPLOYMENTIn the company’s own environment

INTEGRATED WITHCollibra Data Catalog

QUOTE

> “We had requirements from multiple data owners and business functions. Having a partner that could deliver across the use cases was paramount. DataXray ticked all the boxes.”

_PRODUCT OWNER, DATA ANALYTICS PLATFORM · INTEGRATED ENERGY COMPANY_

THE CHALLENGE

## The files had to be classified before they could move

The company had chosen Collibra as its data catalog, to improve governance, security, and privacy across the business. First it had to move more than 110 million unstructured files out of OpenText Livelink, its legacy content system.

Moving files it couldn’t describe would have pulled sensitive data into the catalog. The company needed one metadata framework across the enterprise first, and that framework also had to work for privacy requests, security tools, and a new data lake.

At that scale, every step that needed a person would have slowed the migration down.

THE APPROACH

## DataXray read the legacy content first

The company chose DataXray after a competitive proof of concept. It needed something that would scale to hundreds of millions of files across many kinds of source.

DataXray discovered and classified the files in OpenText Livelink, and the company built its framework on the results. The DataXray team ran workshops with security, privacy, data custodians, and other stakeholders, so every team agreed on what the framework had to do.

DataXray also built a custom connector for the company’s legacy sources, so a critical use case was ready in time for the rollout.

THE RESULT

## The framework was ready for Collibra

Discovery and classification in OpenText Livelink ran automatically. The company expects the work to save millions of dollars over the term, through automated disposition, lower storage costs, automated privacy processes, and license savings.

One enterprise-wide metadata framework

Discovery and classification automated in OpenText Livelink

One framework for the 110M+ file move into Collibra

Millions in expected savings over the term

WHAT HAPPENED NEXT

## Then the company went after storage costs

With budgets tight, the company turned DataXray on its legacy file servers, which held several dozen petabytes. A first scan by file age found files it could delete straight away. A second scan by content found more to delete or archive.

For every petabyte deleted or moved to cold storage, the company saved around $800,000.

[EXPLORE THE PLATFORM](https://www.dataxray.io/platform/)

## Talk to the DataXray team

[Book a demo](https://www.dataxray.io/demo/#book)
