# Burstable EC2 Instances Incurring Unlimited Mode Surplus Credit Charges

Canonical: https://www.pointfive.co/efficiency-hub/inefficiencies/burstable-ec2-instances-incurring-unlimited-mode-surplus-credit-charges

Burstable performance instances (the T family) are priced for workloads that run below a baseline CPU level most of the time.

By: PointFive

Updated: 2026-09-28

[Cloud Efficiency Hub](https://www.pointfive.co/efficiency-hub) 

The short version

Burstable performance instances (the T family) are priced for workloads that run below a baseline CPU level most of the time.

PointFive Research

Cloud cost research at PointFive

AWS service

[AWS EC2](https://www.pointfive.co/efficiency-hub/cloud-services/aws-ec2)

Category

[Compute](https://www.pointfive.co/efficiency-hub/service-category/compute)

Reference

CER-0344

Type

Inefficient Configuration

## Explanation

Why the waste happens and who it affects.

T3, T3a, T4g and T8i instances launch in unlimited mode by default, which lets them burst above the baseline for as long as needed by spending surplus credits. If the instance's average CPU over a 24-hour period stays above its baseline, the surplus credits that are not paid down are billed at a flat additional rate per vCPU-hour, on top of the instance's hourly price.

This is common when a T instance chosen for a light workload becomes a steadily busy one, or when T instances are used as a cheap default for build agents, batch jobs or small production services. The extra charge does not appear as a separate resource and performance never degrades, so it is easy to miss. AWS documents a breakeven point beyond which a burstable instance in unlimited mode costs more than an equivalently sized fixed-performance instance; for example, a t3.large costs more than an m5.large above 42.5 percent average CPU, and a T3 bursting continuously at 100 percent costs about 1.5 times the equivalent M5 (us-east-1, Linux).

## Billing model

The pricing dimensions that drive this cost.

Instance hourly price

Covers CPU usage up to the instance's baseline utilization per vCPU

Surplus credits

Credits spent after the earned CPUCreditBalance reaches zero; they are paid down by credits earned when CPU drops below baseline

Surplus credit charge

Surplus credits not paid down are billed per vCPU-hour; AWS lists $0.05 for T2 and T3 and $0.04 for T4g on Linux, RHEL and SLES, the same across sizes and Regions

When charges post

When spent surplus credits exceed the maximum the instance can earn in 24 hours, when the instance stops or terminates, or when it switches from unlimited to standard

## How to detect

5 checks to find it in your estate.

- Query the CloudWatch metric CPUSurplusCreditsCharged per instance; any sustained non-zero value means the instance is paying for CPU above its hourly price

- Track CPUSurplusCreditBalance for unlimited instances; a balance that stays high means the instance is not earning enough credits to pay down its bursting

- Compare 24-hour average CPUUtilization with the instance type's baseline utilization per vCPU from the credit table, and with the breakeven CPU percentage for the equivalent fixed-performance type

- List credit settings with describe-instance-credit-specifications and check the account-level default credit specification per T family and Region

- Review Compute Optimizer recommendations for T instances, whose CPU utilization graph shows the burstable baseline alongside actual usage

## How to fix

5 ways to remove the waste.

- Move instances whose average CPU stays above the breakeven point to a fixed-performance type of suitable size (for example M or C family), which removes the surplus charge

- Where occasional throttling to baseline is acceptable, switch the instance to standard credit mode; note that any outstanding CPUSurplusCreditBalance is charged immediately at the switch

- Consider a larger T size only if the workload's average CPU would stay below that size's baseline

- Set the account default credit specification for each T family deliberately, and set the credit option explicitly in launch templates for workloads with known sustained load

- Create CloudWatch alarms on CPUSurplusCreditsCharged so new cases are caught within hours rather than at month end

## Documentation

Vendor references for pricing and configuration.

- [Unlimited mode concepts for burstable instances  docs.aws.amazon.com](https://docs.aws.amazon.com/AWSEC2/latest/UserGuide/burstable-performance-instances-unlimited-mode-concepts.html)

- [Monitor CPU credits for burstable instances  docs.aws.amazon.com](https://docs.aws.amazon.com/AWSEC2/latest/UserGuide/burstable-performance-instances-monitoring-cpu-credits.html)

- [Key concepts for burstable performance instances  docs.aws.amazon.com](https://docs.aws.amazon.com/AWSEC2/latest/UserGuide/burstable-credits-baseline-concepts.html)

- [Amazon EC2 On-Demand Pricing  aws.amazon.com](https://aws.amazon.com/ec2/pricing/on-demand/)

## Related inefficiencies

[Browse the library](https://www.pointfive.co/efficiency-hub)

- AWS EC2  CER-0327

### [Spot Instance Overreliance Without Effective Cost-Per-Performance Analysis](https://www.pointfive.co/efficiency-hub/inefficiencies/spot-instance-overreliance-without-cost-per-performance-analysis)

Organizations frequently pursue aggressive Spot Instance adoption based on headline discount percentages - up to 90% off On-Demand pricing - without evaluating the effective cost per unit of work completed. While Spot pricing can deliver...

Compute

- AWS EC2  CER-0211

### [Unnecessary Multi-AZ Deployment for Non-Production EC2 Instances](https://www.pointfive.co/efficiency-hub/inefficiencies/unnecessary-multi-az-deployment-for-non-production-ec2-instances)

Multi-AZ deployment is often essential for production workloads, but its use in non-production environments (e.g., development, test, QA) offers minimal value. These environments typically do not require high availability, yet still incur...

Compute

- AWS EC2  CER-0096

### [Missing Scheduled Shutdown for Non-Production EC2 Instances](https://www.pointfive.co/efficiency-hub/inefficiencies/missing-scheduled-shutdown-for-non-production-ec2-instances)

Non-production EC2 instances are often provisioned for daytime-only usage but remain running 24/7 out of convenience or oversight. This results in unnecessary compute charges, even if the workload is inactive for 16+ hours per day. AWS...

Compute

---
Source: the public page above. Product screenshots and illustrative interfaces are examples, not live customer data.

