Data Archiving
Data archiving is the practice of identifying data that is no longer actively used and moving it out of production systems into separate, long-term storage. The goal is to preserve information that remains important for retention or reference purposes while keeping active systems lighter and less cluttered. Archived data is typically accessed infrequently, if at all, but is retained because it still matters to the organization.
Data archiving is the process of identifying inactive or historical data and relocating it from production systems into a dedicated long-term storage tier for retention. Distinct from backup (which creates recoverable copies of active data), archiving moves data out of primary systems for durable preservation, and archived data is generally optimized for infrequent access rather than frequent modification. In some implementations, archiving mechanisms also capture and manage changes to data over time to support historical analysis. This definition addresses the storage and lifecycle concept only; it does not cover jurisdiction-specific retention periods, lawful retention obligations, deletion or disposal requirements, cross-border transfer controls, or the security controls applied to archived stores, each of which must be determined by applicable law, policy, and implementation context. Archiving personal data does not remove it from the scope of applicable data protection obligations.
Why it matters
Data archiving matters because production systems accumulate large volumes of inactive or historical data over time, and retaining that data in primary systems increases cost, complexity, and management overhead. By moving data that is no longer actively used into a dedicated long-term storage tier, organizations can keep active systems lighter while preserving information that remains important for retention or reference. This supports both operational efficiency and the organization's ability to retain information it still considers valuable.
From an information governance perspective, archiving is a lifecycle activity that intersects with, but does not by itself satisfy, retention and disposal requirements. Archiving personal data does not remove it from the scope of applicable data protection obligations; archived personal data generally remains personal data and continues to be subject to the same accountability, security, and lawful-basis considerations as data held in production. Organizations should not treat the act of archiving as equivalent to compliance with retention law, and separate policy and legal analysis is needed to determine how long data may or must be kept.
Archiving is also distinct from backup, and conflating the two can create governance gaps. Backup creates recoverable copies of active data to guard against loss, while archiving moves data out of primary systems for durable, long-term preservation of information that is accessed infrequently. Treating an archive as a backup, or a backup as an archive, can lead to incorrect assumptions about recoverability, retention, and the controls that should apply to each store.
Who it's relevant to
Inside Data Archiving
Common questions
Answers to the questions practitioners most commonly ask about Data Archiving.