mcpbeat Sign in

AWS Cost Operations Skill for Claude

This skill provides AWS cost optimization, monitoring, and operational best practices with integrated MCP servers for billing analysis, cost estimation, observability, and security assessment.

9k tokens
context cost
the whole folder, loaded on every use
3
files
instructions only
0
copies elsewhere
how many repositories repackaged it
320
stars on the repo
on the repository, not the skill itself

Install

one command, takes just this skill from the repository
npx skills add https://github.com/Microck/ordinary-claude-skills --skill aws-cost-operations

What comes with it

24 213 bytes besides the instruction
references/cloudwatch-alarms.md
references/operations-patterns.md

The instruction itself

36 sections, as written by the author

AWS Cost & Operations

This skill provides comprehensive guidance for AWS cost optimization, monitoring, observability, and operational excellence with integrated MCP servers.

Integrated MCP Servers

This skill includes 8 MCP servers automatically configured with the plugin:

Cost Management Servers

1. AWS Billing and Cost Management MCP Server

Purpose: Real-time billing and cost management

  • View current AWS spending and trends
  • Analyze billing details across services
  • Track budget utilization
  • Monitor cost allocation tags
  • Review consolidated billing for organizations
2. AWS Pricing MCP Server

Purpose: Pre-deployment cost estimation and optimization

  • Estimate costs before deploying resources
  • Compare pricing across regions
  • Calculate Total Cost of Ownership (TCO)
  • Evaluate different service options for cost efficiency
  • Get current pricing information for AWS services
3. AWS Cost Explorer MCP Server

Purpose: Detailed cost analysis and reporting

  • Analyze historical spending patterns
  • Create custom cost reports
  • Identify cost anomalies and trends
  • Forecast future costs
  • Analyze cost by service, region, or tag
  • Generate cost optimization recommendations

Monitoring & Observability Servers

4. Amazon CloudWatch MCP Server

Purpose: Metrics, alarms, and logs analysis

  • Query CloudWatch metrics and logs
  • Create and manage CloudWatch alarms
  • Analyze application performance metrics
  • Troubleshoot operational issues
  • Set up custom dashboards
  • Monitor resource utilization
5. Amazon CloudWatch Application Signals MCP Server

Purpose: Application monitoring and performance insights

  • Monitor application health and performance
  • Analyze service-level objectives (SLOs)
  • Track application dependencies
  • Identify performance bottlenecks
  • Monitor service map and traces
6. AWS Managed Prometheus MCP Server

Purpose: Prometheus-compatible monitoring

  • Query Prometheus metrics
  • Monitor containerized applications
  • Analyze Kubernetes workload metrics
  • Create PromQL queries
  • Track custom application metrics

Audit & Security Servers

7. AWS CloudTrail MCP Server

Purpose: AWS API activity and audit analysis

  • Analyze AWS API calls and user activity
  • Track resource changes and modifications
  • Investigate security incidents
  • Audit compliance requirements
  • Identify unusual access patterns
  • Review who made what changes when
8. AWS Well-Architected Security Assessment Tool MCP Server

Purpose: Security assessment against Well-Architected Framework

  • Assess security posture against AWS best practices
  • Identify security gaps and vulnerabilities
  • Get security improvement recommendations
  • Review security pillar compliance
  • Generate security assessment reports

When to Use This Skill

Use this skill when:

  • Optimizing AWS costs and reducing spending
  • Estimating costs before deployment
  • Monitoring application and infrastructure performance
  • Setting up observability and alerting
  • Analyzing spending patterns and trends
  • Investigating operational issues
  • Auditing AWS activity and changes
  • Assessing security posture
  • Implementing operational excellence

Cost Optimization Best Practices

Pre-Deployment Cost Estimation

Always estimate costs before deploying:

  • Use AWS Pricing MCP to estimate resource costs
  • Compare pricing across different regions
  • Evaluate alternative service options
  • Calculate expected monthly costs
  • Plan for scaling and growth

Example workflow:

"Estimate the monthly cost of running a Lambda function with
1 million invocations, 512MB memory, 3-second duration in us-east-1"

Cost Analysis and Optimization

Regular cost reviews:

  • Use Cost Explorer MCP to analyze spending trends
  • Identify cost anomalies and unexpected charges
  • Review costs by service, region, and environment
  • Compare actual vs. budgeted costs
  • Generate cost optimization recommendations

Cost optimization strategies:

  • Right-size over-provisioned resources
  • Use appropriate storage classes (S3, EBS)
  • Implement auto-scaling for dynamic workloads
  • Leverage Savings Plans and Reserved Instances
  • Delete unused resources and snapshots
  • Use cost allocation tags effectively

Budget Monitoring

Track spending against budgets:

  • Use Billing and Cost Management MCP to monitor budgets
  • Set up budget alerts for threshold breaches
  • Review budget utilization regularly
  • Adjust budgets based on trends
  • Implement cost controls and governance

Monitoring and Observability Best Practices

CloudWatch Metrics and Alarms

Implement comprehensive monitoring:

  • Use CloudWatch MCP to query metrics and logs
  • Set up alarms for critical metrics:
  • CPU and memory utilization
  • Error rates and latency
  • Queue depths and processing times
  • API gateway throttling
  • Lambda errors and timeouts
  • Create CloudWatch dashboards for visualization
  • Use log insights for troubleshooting

Example alarm scenarios:

  • Lambda error rate > 1%
  • EC2 CPU utilization > 80%
  • API Gateway 4xx/5xx error spike
  • DynamoDB throttled requests
  • ECS task failures

Application Performance Monitoring

Monitor application health:

  • Use CloudWatch Application Signals MCP for APM
  • Track service-level objectives (SLOs)
  • Monitor application dependencies
  • Identify performance bottlenecks
  • Set up distributed tracing

Container and Kubernetes Monitoring

For containerized workloads:

  • Use AWS Managed Prometheus MCP for metrics
  • Monitor container resource utilization
  • Track pod and node health
  • Create PromQL queries for custom metrics
  • Set up alerts for container anomalies

Audit and Security Best Practices

CloudTrail Activity Analysis

Audit AWS activity:

  • Use CloudTrail MCP to analyze API activity
  • Track who made changes to resources
  • Investigate security incidents
  • Monitor for suspicious activity patterns
  • Audit compliance with policies

Common audit scenarios:

  • "Who deleted this S3 bucket?"
  • "Show all IAM role changes in the last 24 hours"
  • "List failed login attempts"
  • "Find all actions by a specific user"
  • "Track modifications to security groups"

Security Assessment

Regular security reviews:

  • Use Well-Architected Security Assessment MCP
  • Assess security posture against best practices
  • Identify security gaps and vulnerabilities
  • Implement recommended security improvements
  • Document security compliance

Security assessment areas:

  • Identity and Access Management (IAM)
  • Detective controls and monitoring
  • Infrastructure protection
  • Data protection and encryption
  • Incident response preparedness

Using MCP Servers Effectively

Cost Analysis Workflow

  • Pre-deployment: Use Pricing MCP to estimate costs
  • Post-deployment: Use Billing MCP to track actual spending
  • Analysis: Use Cost Explorer MCP for detailed cost analysis
  • Optimization: Implement recommendations from Cost Explorer

Monitoring Workflow

  • Setup: Configure CloudWatch metrics and alarms
  • Monitor: Use CloudWatch MCP to track key metrics
  • Analyze: Use Application Signals for APM insights
  • Troubleshoot: Query CloudWatch Logs for issue resolution

Security Workflow

  • Audit: Use CloudTrail MCP to review activity
  • Assess: Use Well-Architected Security Assessment
  • Remediate: Implement security recommendations
  • Monitor: Track security events via CloudWatch

MCP Usage Best Practices

  • Cost Awareness: Check pricing before deploying resources
  • Proactive Monitoring: Set up alarms for critical metrics
  • Regular Reviews: Analyze costs and performance weekly
  • Audit Trails: Review CloudTrail logs for compliance
  • Security First: Run security assessments regularly
  • Optimize Continuously: Act on cost and performance recommendations

Operational Excellence Guidelines

Cost Optimization

  • Tag Everything: Use consistent cost allocation tags
  • Review Monthly: Analyze spending trends and anomalies
  • Right-size: Match resources to actual usage
  • Automate: Use auto-scaling and scheduling
  • Monitor Budgets: Set alerts for cost overruns

Monitoring and Alerting

  • Critical Metrics: Alert on business-critical metrics
  • Noise Reduction: Fine-tune thresholds to reduce false positives
  • Actionable Alerts: Ensure alerts have clear remediation steps
  • Dashboard Visibility: Create dashboards for key stakeholders
  • Log Retention: Balance cost and compliance needs

Security and Compliance

  • Least Privilege: Grant minimum required permissions
  • Audit Regularly: Review CloudTrail logs for anomalies
  • Encrypt Data: Use encryption at rest and in transit
  • Assess Continuously: Run security assessments frequently
  • Incident Response: Have procedures for security events

Additional Resources

For detailed operational patterns and best practices, refer to the comprehensive reference:

File: references/operations-patterns.md

This reference includes:

  • Cost optimization strategies
  • Monitoring and alerting patterns
  • Observability best practices
  • Security and compliance guidelines
  • Troubleshooting workflows

CloudWatch Alarms Reference

File: references/cloudwatch-alarms.md

Common alarm configurations for:

  • Lambda functions
  • EC2 instances
  • RDS databases
  • DynamoDB tables
  • API Gateway
  • ECS services
  • Application Load Balancers

Other skills for the same job

different authors, same section of the catalogue
Azure Kubernetes Automatic Readiness
by microsoft
vendor ×3

Assess Kubernetes workloads and cluster configuration for AKS Automatic compatibility. Identifies incompatibilities, generates fixes, and guides migration from AKS Standard to AKS Automatic. WHEN: migrate to AKS Automatic, check AKS Automatic readiness, validate manifests for Automatic, assess cluster for Automatic compatibility, fix deployment for Automatic compatibility, identify AKS Automatic migration blockers, is my cluster ready for AKS Automatic.

13k tokens
Capacity
by microsoft
vendor ×3

Discovers available Azure OpenAI model capacity across regions and projects. Analyzes quota limits, compares availability, and recommends optimal deployment locations based on capacity requirements. USE FOR: find capacity, check quota, where can I deploy, capacity discovery, best region for capacity, multi-project capacity search, quota analysis, model availability, region comparison, check TPM availability. DO NOT USE FOR: actual deployment (hand off to preset or customize after discovery), quota increase requests (direct user to Azure Portal), listing existing deployments.

6k tokens scripts
Customize
by microsoft
vendor ×3

Interactive guided deployment flow for Azure OpenAI models with full customization control. Step-by-step selection of model version, SKU (GlobalStandard/Standard/ProvisionedManaged), capacity, RAI policy (content filter), and advanced options (dynamic quota, priority processing, spillover). USE FOR: custom deployment, customize model deployment, choose version, select SKU, set capacity, configure content filter, RAI policy, deployment options, detailed deployment, advanced deployment, PTU deployment, provisioned throughput. DO NOT USE FOR: quick deployment to optimal region (use preset).

8k tokens
Deploy Model
by microsoft
vendor ×3

Unified Azure OpenAI model deployment skill with intelligent intent-based routing. Handles quick preset deployments, fully customized deployments (version/SKU/capacity/RAI policy), and capacity discovery across regions and projects. USE FOR: deploy model, deploy gpt, create deployment, model deployment, deploy openai model, set up model, provision model, find capacity, check model availability, where can I deploy, best region for model, capacity analysis. DO NOT USE FOR: listing existing deployments (use foundry_models_deployments_list MCP tool), deleting deployments, agent creation (use agent/create), project creation (use project/create).

26k tokens scripts
Preset
by microsoft
vendor ×3

Intelligently deploys Azure OpenAI models to optimal regions by analyzing capacity across all available regions. Automatically checks current region first and shows alternatives if needed. USE FOR: quick deployment, optimal region, best region, automatic region selection, fast setup, multi-region capacity check, high availability deployment, deploy to best location. DO NOT USE FOR: custom SKU selection (use customize), specific version selection (use customize), custom capacity configuration (use customize), PTU deployments (use customize).

9k tokens
Lamindb
by christophacham
×3

This skill should be used when working with LaminDB, an open-source data framework for biology that makes data queryable, traceable, reproducible, and FAIR. Use when managing biological datasets (scRNA-seq, spatial, flow cytometry, etc.), tracking computational workflows, curating and validating data with biological ontologies, building data lakehouses, or ensuring data lineage and reproducibility in biological research. Covers data management, annotation, ontologies (genes, cell types, diseases, tissues), schema validation, integrations with workflow managers (Nextflow, Snakemake) and MLOps platforms (W&B, MLflow), and deployment strategies.

22k tokens
Latchbio Integration
by christophacham
×3

Latch platform for bioinformatics workflows. Build pipelines with Latch SDK, @workflow/@task decorators, deploy serverless workflows, LatchFile/LatchDir, Nextflow/Snakemake integration.

12k tokens
Modal
by christophacham
×3

Run Python code in the cloud with serverless containers, GPUs, and autoscaling. Use when deploying ML models, running batch processing jobs, scheduling compute-intensive tasks, or serving APIs that require GPU acceleration or dynamic scaling.

17k tokens

How to use it

Copy the folder

Take microck/aws-cost-operations from the repository into ~/.claude/skills for personal use, or into .claude/skills inside a project.

Check the name does not clash

The agent identifies a skill by the name field in its header. Two skills with the same name cannot sit side by side — one of them will be ignored.