▸case-09 Our GCP BigQuery monthly bill jumped to $15,000 due to ad-hoc query users running SELECT * on multi-terabyte tables without partition filters. Here is our usage breakdown showing 80% of queries scan full unpartitioned tables. Propose a cost reduction strategy with budget controls. | fail→fail | — | — | — | — | — | — | — | — | — | — | — | — |
▸case-22 Our company pays $8,000/month for a 10 Gbps Azure ExpressRoute circuit, but cross-cloud data transfer logs show peak bandwidth consumption never exceeds 40 Mbps. Recommend an architecture change. | fail→fail | — | — | — | — | — | — | — | — | — | — | — | — |
▸case-12 Can you create a cost optimization plan to downsize our GCP Cloud Spanner nodes from 10 nodes to 3 nodes and tell us our exact monthly savings? | fail→fail | — | — | — | — | — | — | — | — | — | — | — | — |
▸case-14 Our EKS cluster runs 50 pods with CPU requests set to 4 vCPU and memory set to 16 GB, but Datadog metrics show actual usage averages 0.3 vCPU and 2 GB RAM per pod. Provide a rightsizing plan with budget safeguards. | fail→pass | — | — | — | — | — | — | — | — | — | — | — | — |
▸case-19 Our AWS bill shows $4,000/month for EC2 NAT Gateway data transfer charges. Microservices in private subnets are downloading 50 TB of data from S3 buckets through the NAT Gateway. How do we eliminate this transfer cost? | fail→fail | — | — | — | — | — | — | — | — | — | — | — | — |
▸case-06 Our staging environment runs on AWS r5.2xlarge RDS Postgres instances with 15% CPU utilization. We want to downsize these to r5.large to cut costs. Here is our CloudWatch metric summary showing average CPU 12%, max CPU 28%, and memory usage 30% over 30 days. Provide the rightsizing plan. | fail→fail | — | — | — | — | — | — | — | — | — | — | — | — |
▸case-16 Our AWS Lambda functions have Provisioned Concurrency set to 200 execution environments 24/7 costing $700/month, but CloudWatch metrics show maximum concurrent executions rarely exceed 15 except during a 5-minute cron job at midnight. Propose a cost fix. | fail→pass | — | — | — | — | — | — | — | — | — | — | — | — |
▸case-08 Our DynamoDB table 'UserSessions' has provisioned capacity set to 5000 WCU and 5000 RCU, but traffic is highly unpredictable with long idle periods, costing $1,200/month. Here is our metric log showing 90% idle time with sudden 10-minute bursts. Recommend a billing model change with risk and rollback. | fail→fail | — | — | — | — | — | — | — | — | — | — | — | — |
▸case-21 We have untagged, running resources across AWS, Azure, and GCP left behind by former developers. We want to clean up waste, set budgets, and establish an ongoing cost governance cadence. What exact steps should we take? | fail→fail | — | — | — | — | — | — | — | — | — | — | — | — |
▸case-15 We manage 30 individual Azure SQL Databases on vCore Provisioned Gen5 4-vCore instances, each costing $220/month. Usage graphs show individual databases spike at different times while average utilization across the group is under 15%. Recommend a cost reduction architecture. | fail→pass | — | — | — | — | — | — | — | — | — | — | — | — |
▸case-01 Our monthly AWS bill for RDS databases has increased significantly over the past two quarters. Please analyze our database usage data and provide a cost reduction report that highlights immediate savings opportunities with estimated dollar amounts, along with risk assessments and rollback steps for each recommended instance adjustment. | fail→fail | — | — | — | — | — | — | — | — | — | — | — | — |
▸case-02 We suspect we are paying for unattached storage volumes and idle compute instances across Azure and GCP. Can you review our resource usage over the past month, categorize the waste with calculated cost impact, and detail a setup plan for budget thresholds, alerts, and recurring review practices? | fail→fail | — | — | — | — | — | — | — | — | — | — | — | — |
▸case-03 Production latency on our GCP Cloud SQL cluster spiked to 12 seconds during peak traffic right now, causing HTTP 500 errors across our primary API. We need an immediate incident response to restore service availability. | fail→fail | — | — | — | — | — | — | — | — | — | — | — | — |
▸case-07 We identified 40 unattached AWS EBS volumes (gp2, total 8 TB) that have been detached for over 60 days, costing $800/month. We want to delete them immediately to stop billing. Provide the cleanup instructions. | fail→pass | — | — | — | — | — | — | — | — | — | — | — | — |
▸case-17 We store 500 TB of backup VHD images in Azure Blob Storage Hot tier costing $10,000/month. These backups are read only during disaster recovery tests once a year. Provide a storage optimization plan. | fail→fail | — | — | — | — | — | — | — | — | — | — | — | — |
▸case-18 Our multi-account AWS organization has no cost allocation tags or spending limits, resulting in unexpected $50,000 monthly bills. How should we set up automated cost controls and tagging enforcement? | fail→pass | — | — | — | — | — | — | — | — | — | — | — | — |
▸case-04 We are refactoring our PostgreSQL schema to ensure strict serializable transaction isolation and eliminate deadlocks between microservices. How should we rewrite our SELECT FOR UPDATE queries to ensure lock acquisition order is consistent? | fail→fail | — | — | — | — | — | — | — | — | — | — | — | — |
▸case-05 Our security compliance audit requires us to restrict IAM permissions across our AWS S3 buckets to enforce encryption at rest and block public access. Please provide the policy JSON to remediate these security findings. | fail→fail | — | — | — | — | — | — | — | — | — | — | — | — |
▸case-20 Our GCP GKE batch processing workloads run on standard n2-standard-8 compute instances costing $6,000/month. The batch jobs are fault-tolerant and retry automatically on failure. How can we reduce compute cost by up to 60-80%? | fail→fail | — | — | — | — | — | — | — | — | — | — | — | — |
▸case-10 We have 20 Azure Standard_D8s_v5 virtual machines running 24/7 for core microservices in eastus. Current pay-as-you-go cost is $5,200/month. Metrics show steady 65% CPU load year-round. Provide a cost optimization plan. | fail→fail | — | — | — | — | — | — | — | — | — | — | — | — |
▸case-11 We are building an automated cloud cost management pipeline across AWS and GCP using Terraform and open-source FinOps tools. Where can we find the detailed step-by-step implementation workflows and tooling guidance within the available workspace resources? | fail→pass | — | — | — | — | — | — | — | — | — | — | — | — |
▸case-13 Our S3 bucket 'app-raw-logs' contains 150 TB of log files stored in S3 Standard, costing $3,450/month. Logs are queried frequently for 14 days, rarely for 90 days, and never accessed after 90 days but must be retained for 1 year. Provide a lifecycle rule strategy with risk assessment. | fail→fail | — | — | — | — | — | — | — | — | — | — | — | — |