# Data Deduplication Disabled on FSx for Windows File Server

Canonical: https://www.pointfive.co/efficiency-hub/inefficiencies/data-deduplication-disabled-on-fsx-for-windows-file-server

FSx for Windows File Server supports Microsoft Data Deduplication, which stores repeated chunks of data only once and compresses them, but it is not...

By: PointFive

Updated: 2026-09-28

[Cloud Efficiency Hub](https://www.pointfive.co/efficiency-hub) 

The short version

FSx for Windows File Server supports Microsoft Data Deduplication, which stores repeated chunks of data only once and compresses them, but it is not enabled by default.

PointFive Research

Cloud cost research at PointFive

AWS service

[Amazon FSx](https://www.pointfive.co/efficiency-hub/cloud-services/amazon-fsx)

Category

[Storage](https://www.pointfive.co/efficiency-hub/service-category/storage)

Reference

CER-0400

Type

Inefficient Configuration

## Explanation

Why the waste happens and who it affects.

User home directories, departmental shares and software build shares tend to hold many copies of the same or similar files, so without deduplication the file system needs far more provisioned storage than the unique data requires. AWS documents typical savings of 50-60 percent for general-purpose shares, 30-50 percent for user documents and 70-80 percent for software development datasets.

Because FSx for Windows bills on provisioned storage capacity, and that capacity can be increased but never decreased, skipping deduplication compounds: teams add capacity as shares fill up, and each increase is permanent for the life of the file system. The pattern is common on file systems migrated from on-premises servers where deduplication was never configured. Deduplication is a poor fit for highly dynamic data, so the savings are concentrated on shares with mostly static content.

## Billing model

The pricing dimensions that drive this cost.

FSx for Windows File Server bills for what you provision, prorated by the hour, not for the data actually stored.

Storage capacity

Billed per GB-month of provisioned SSD or HDD capacity, by deployment type, whether or not it is used

Throughput capacity

Billed per MBps-month of provisioned throughput, which also sets the memory available for deduplication jobs

Backups

Billed per GB-month of backup storage consumed

Capacity increase only

Storage capacity can be increased but never decreased on an existing file system

## How to detect

4 checks to find it in your estate.

- Run Get-FSxDedupStatus through the FSx remote PowerShell endpoint on each Windows file system; file systems with no deduplication status have it disabled

- Use Measure-FSxDedupFileMetadata on representative folders to estimate how much space deduplication would reclaim before enabling it

- Prioritize file systems whose FreeStorageCapacity CloudWatch metric is trending down and that are candidates for a storage capacity increase

- Classify shares by content type: user documents, home directories and software build or binary shares are the strongest candidates, while databases and rapidly changing data are not

## How to fix

5 ways to remove the waste.

- Enable deduplication with Enable-FSxDedup and adjust settings with Set-FSxDedupConfiguration, for example restricting it to certain file types or folders, or setting a minimum file age

- Schedule optimization and garbage collection jobs with Set-FSxDedupSchedule during idle periods; space is freed only after garbage collection runs

- Check that the file system's throughput capacity provides enough memory for the logical data size (Microsoft recommends about 1 GB of memory per 1 TB of logical data) and increase it if deduplication jobs fail for lack of memory

- Use the reclaimed space to defer or avoid future storage capacity increases; to actually lower provisioned capacity, create a smaller file system and migrate the data, sizing it with headroom until deduplication has run on the new file system

- Monitor job results with Get-FSxDedupStatus, and avoid Robocopy options that AWS and Microsoft flag as unsafe with deduplication

## Documentation

Vendor references for pricing and configuration.

- [Amazon FSx for Windows File Server Pricing  aws.amazon.com](https://aws.amazon.com/fsx/windows/pricing/)

- [Enable data deduplication in Amazon FSx  docs.aws.amazon.com](https://docs.aws.amazon.com/prescriptive-guidance/latest/optimize-costs-microsoft-workloads/storage-fsx-deduplication.html)

- [Managing storage on FSx for Windows File Server  docs.aws.amazon.com](https://docs.aws.amazon.com/fsx/latest/WindowsGuide/managing-storage-configuration.html)

## Related inefficiencies

[Browse the library](https://www.pointfive.co/efficiency-hub)

- Amazon FSx  CER-0328

### [Infrequently Accessed Data Retained on High-Performance FSx File Systems](https://www.pointfive.co/efficiency-hub/inefficiencies/infrequently-accessed-data-retained-on-high-performance-fsx-file-systems)

Amazon FSx file systems are designed for performance-sensitive workloads such as shared enterprise file systems, high-performance computing, analytics, and machine learning. Storage costs are driven by provisioned capacity (measured in...

Storage

- Amazon FSx  CER-0401

### [Inactive FSx File System](https://www.pointfive.co/efficiency-hub/inefficiencies/inactive-fsx-file-system)

Amazon FSx file systems, whether FSx for Windows File Server, FSx for NetApp ONTAP, FSx for OpenZFS or FSx for Lustre, bill for the storage and throughput capacity configured on them, prorated by the hour, regardless of whether any client...

Storage

- AWS ECR  CER-0308

### [ECR Archive Storage Class Used Below 150 TB Threshold](https://www.pointfive.co/efficiency-hub/inefficiencies/ecr-archive-storage-class-used-below-150-tb-threshold)

In November 2025, AWS introduced an Archive storage class for private ECR repositories, marketed as a way to reduce storage costs for large volumes of rarely used container images. However, Archive storage pricing is identical to Standard...

Storage

---
Source: the public page above. Product screenshots and illustrative interfaces are examples, not live customer data.

