Organization-level savings report
Track realized savings and workload autoscaler savings across all clusters in your Cast AI organization.
The organization-level savings report gives you a bird's-eye view of how much Cast AI has saved across your entire organization. It shows savings trends over time and a per-cluster breakdown so you can see which clusters contribute the most.
For the shared concepts behind this report — what projected cost, baseline, and the two sources of savings mean — see the Savings Report overview.
Overview cards
The top of the report shows four summary cards:
- Baseline spend — the modeled cost of running your clusters without Cast AI.
- Realized savings — savings realized through node autoscaler including workload autoscaler impact if both are enabled. This card appears only if at least one cluster in your organization has the node autoscaler managing at least 20% of nodes.
- Workload autoscaler savings — modeled savings from the workload autoscaler (shown whenever workload autoscaler is actively managing at least 20% of workloads in VPA mode on at least one cluster).
- Actual spend — the real spending across your organization over the selected period.
NoteRealized savings and workload autoscaler savings appear independently. If neither autoscaler is active on any cluster, the corresponding card won't be shown.
Savings over time
The savings over time chart shows realized savings and workload autoscaler savings trends across the selected period. Which lines appear depends on which autoscalers are adopted across your clusters:
- Realized savings line appears when at least one cluster has the node autoscaler active on at least 20% of nodes.
- Workload autoscaler savings line appears when at least one cluster has the workload autoscaler active in VPA mode on at least 20% of workloads.
Workload autoscaler adoption
This chart shows the percentage of workloads across your organization that are managed by the workload autoscaler over time. It helps you understand how broadly rightsizing has been adopted.
Workload utilization rate
Shows CPU and memory utilization across your organization. You can switch between CPU and memory views. This metric helps you identify whether resources are being used efficiently or whether there's room for further optimization.
Breakdown by cluster
The cluster breakdown table shows one row per cluster, aggregated over the selected time range. For each cluster, you can see:
- Which optimizations are turned on (node autoscaler, workload autoscaler)
- Baseline source — how the baseline was determined for this cluster (see Savings baseline)
- Baseline cost — the modeled cost without Cast AI
- Actual cost — the real spending for this cluster
- Savings — either realized or workload autoscaler savings, depending on which autoscalers are active on the cluster
Click any cluster to drill into the cluster-level savings report for a detailed view.
Related resources
Updated 5 hours ago
