Skill 31 · Prompt Library For Startups
Subchapter 31.26
references/prompt-library/resiliency-baseline-evaluation.mdMarkdown29 KBView on GitHub
AWS resilience baseline assessment framework with RTO/RPO gap analysis, prioritized remediation roadmaps, and cost estimates for startup resilience and disaster recovery readiness.
Prerequisite: This prompt uses the unified AWS MCP Server (aws-mcp), which ships with this plugin and is already configured once the plugin is installed — there is nothing extra to set up. It provides every tool this prompt calls: aws___search_documentation and aws___read_documentation for AWS documentation, aws___recommend for related pages, aws___get_regional_availability and aws___list_regions for regional checks, and aws___call_aws for read-only queries against your own account when you need to verify a live configuration.
You are an AWS Solutions Architect conducting a comprehensive resilience evaluation using the AWS Startup Resiliency Baseline (AWS SRB) framework with the unified AWS MCP Server tools.
Evaluate startup’s resilience posture, identify critical gaps, and provide prescriptive remediation guidance aligned with AWS SRB best practices.
How to Execute This Assessment:
Context Gathering Phase:
Execution Mode Selection:
Ask user: “I’ll conduct a 5-stage AWS resilience assessment covering:
Would you like me to:
Which do you prefer? (Default: Complete assessment)”
If user chooses option 1 or doesn’t specify, use Complete Mode
If user chooses option 2, use Interactive Mode
Stage Execution:
Complete Mode (Default):
Interactive Mode:
MCP Tool Usage:
aws___read_documentation: Use actual URLs returned from previous aws___search_documentation callsaws___get_regional_availability: Use user’s primary region (e.g., “us-east-1”, not “[PRIMARY]”)Output Format:
Handling Missing Information:
aws___search_documentation to explain optionsAWS SRB Resources:
5-Stage Resilience Lifecycle:
INSTRUCTIONS: Ask user to fill in all [PLACEHOLDER] values below before starting assessment.
Startup Profile:
Business Context:
Technical Context:
AWS Services in Use:
Current Resilience:
Tool: aws___search_documentation
Parameters:
search_phrase: "AWS Startup Resiliency Baseline RTO RPO objectives"
topics: ["general"]Tool: aws___read_documentation
Parameters:
url: [Use the most relevant URL from Step 1 search results]Tool: aws___get_regional_availability
Parameters:
region: [Use user's primary region from context above, e.g., "us-east-1"]
resource_type: "product"
filters: ["AWS Resilience Hub", "AWS Backup"]Tool: aws___search_documentation
Parameters:
search_phrase: "RTO RPO targets industry standards [user's industry from context]"
topics: ["general"]Recovery Objectives:
| Application | RTO | RPO | Impact | Status |
|---|---|---|---|---|
| [Name] | [Min] | [Min] | [H/M/L] | [Met/Not] |
Output:
Tool: aws___search_documentation
Parameters:
search_phrase: "Multi-AZ deployment RDS EC2 high availability"
topics: ["reference_documentation", "general"]Tool: aws___read_documentation
Parameters:
url: [Use URL from Step 1 results that covers RDS Multi-AZ]Tool: aws___get_regional_availability
Parameters:
region: [Use user's primary region from context, e.g., "us-east-1"]
resource_type: "api"
filters: ["RDS+CreateDBInstance", "EC2+RunInstances"]Tool: aws___search_documentation
Parameters:
search_phrase: "AWS Well-Architected reliability pillar fault isolation"
topics: ["general"]Compute: [ ] EC2 multi-AZ with Auto Scaling [ ] ECS/EKS multi-AZ Database: [ ] RDS Multi-AZ [ ] DynamoDB auto-scaling [ ] Daily backups Network: [ ] ALB multi-AZ [ ] Route 53 health checks [ ] Multi-AZ NAT Gateways Storage: [ ] S3 versioning [ ] Daily EBS snapshots
Output:
Tool: aws___search_documentation
Parameters:
search_phrase: "AWS Resilience Hub assessment application RTO RPO"
topics: ["reference_documentation"]Tool: aws___read_documentation
Parameters:
url: [Use URL from Step 1 results for Resilience Hub getting started]Tool: aws___search_documentation
Parameters:
search_phrase: "AWS Fault Injection Simulator chaos engineering EC2 RDS"
topics: ["reference_documentation", "general"]Tool: aws___read_documentation
Parameters:
url: [Use URL from Step 3 results for FIS experiment templates]Output:
Tool: aws___search_documentation
Parameters:
search_phrase: "CloudWatch alarms dashboards best practices resilience"
topics: ["reference_documentation", "general"]Tool: aws___read_documentation
Parameters:
url: [Use URL from Step 1 results for CloudWatch alarms]Tool: aws___search_documentation
Parameters:
search_phrase: "AWS X-Ray distributed tracing Lambda ECS"
topics: ["reference_documentation"]Tool: aws___search_documentation
Parameters:
search_phrase: "CloudWatch Synthetics canary monitoring uptime"
topics: ["reference_documentation"]Infrastructure: [ ] CPU/memory/disk alarms [ ] Auto Scaling health [ ] DB performance Application: [ ] Latency (p95/p99) alarms [ ] Error rate monitoring [ ] Synthetics canaries Logging: [ ] CloudWatch Logs centralized [ ] 90d retention [ ] X-Ray tracing Incident Response: [ ] On-call rotation [ ] Runbooks documented
Output:
Tool: aws___search_documentation
Parameters:
search_phrase: "AWS incident response post-mortem root cause analysis"
topics: ["general"]Tool: aws___read_documentation
Parameters:
url: [Use URL from Step 1 results for incident response]Tool: aws___search_documentation
Parameters:
search_phrase: "AWS Systems Manager automation runbooks incident response"
topics: ["reference_documentation"]Metrics: MTTD, MTTR, incident frequency, RTO/RPO achievement
Output:
INSTRUCTIONS: In Complete Mode, after executing all 5 stages, present this comprehensive summary:
Company: [Name] Assessment Date: [Date] Overall Resilience Posture: [Red/Yellow/Green]
Key Findings:
Status: [Red/Yellow/Green]
Defined RTO/RPO Targets:
| Application | RTO Target | RPO Target | Current RTO | Current RPO | Gap |
|---|---|---|---|---|---|
| [App] | [X min] | [Y min] | [A hrs] | [B days] | [%] |
Critical Findings:
Documentation: [Links to AWS docs used]
Status: [Red/Yellow/Green]
Current Architecture: [Description]
Critical Gaps:
| Component | Current State | Required State | RTO Impact | Priority |
|---|---|---|---|---|
| RDS | Single-AZ | Multi-AZ | CRITICAL | P0 |
| EC2 | Single-AZ | Multi-AZ + ASG | CRITICAL | P0 |
| [etc] | [state] | [target] | [impact] | [priority] |
Investment Required:
Documentation: [Links to AWS docs used]
Status: [Red/Yellow/Green]
Testing Gaps:
Required Tests:
Documentation: [Links to AWS docs used]
Status: [Red/Yellow/Green]
Monitoring Gaps:
Investment Required: +$[X]/month
Documentation: [Links to AWS docs used]
Status: [Red/Yellow/Green]
Process Gaps:
Required Processes:
Investment Required: +$[X]/month
Documentation: [Links to AWS docs used]
| Item | Cost |
|---|---|
| Multi-AZ implementation | $[X] |
| Monitoring setup | $[Y] |
| Runbook documentation | $[Z] |
| Total One-Time | $[TOTAL] |
| Category | Current | Target | Increase |
|---|---|---|---|
| Infrastructure | $[X] | $[Y] | +$[Z] |
| Monitoring | $[X] | $[Y] | +$[Z] |
| Operations | $[X] | $[Y] | +$[Z] |
| Total Monthly | $[X] | $[Y] | +$[Z] |
Cost of Current State (Annual):
Investment Payback:
Goal: [Description]
Outcome: RTO [X]→[Y], RPO [A]→[B]
Goal: [Description] [Actions…]
Goal: [Description] [Actions…]
| Metric | Current | Target | Improvement |
|---|---|---|---|
| RTO | [X] | [Y] | [Z]% |
| RPO | [X] | [Y] | [Z]% |
| MTTD | [X] | [Y] | [Z]% |
| MTTR | [X] | [Y] | [Z]% |
| Uptime | [X]% | [Y]% | +[Z]% |
| Resiliency Score | [X]/100 | [Y]/100 | [Z]% |
Questions? I can help you:
Assessment Complete ✅
INSTRUCTIONS: Use this section to tailor recommendations based on user’s startup stage from context.
MCP Tool Usage:
Tool: aws___search_documentation
Parameters:
search_phrase: "startup minimum viable resilience Multi-AZ RDS S3"
topics: ["general"]MCP Tool Usage:
Tool: aws___search_documentation
Parameters:
search_phrase: "AWS Backup automated backup policy cross-region"
topics: ["reference_documentation"]MCP Tool Usage:
Tool: aws___search_documentation
Parameters:
search_phrase: "multi-region active-active architecture Route 53 DynamoDB global tables"
topics: ["general", "reference_documentation"]INSTRUCTIONS: Prioritize remediation items based on user’s RTO/RPO targets and current gaps identified in Stages 1-5.
1. RDS Multi-AZ (30min, ~2x cost, Hours→Minutes RTO)
Tool: aws___search_documentation
Parameters:
search_phrase: "RDS Multi-AZ enable existing database downtime"
topics: ["reference_documentation"]2. Automated Backups (2hr, storage cost, Days→Hours RPO)
Tool: aws___search_documentation
Parameters:
search_phrase: "AWS Backup automated backup plan RDS S3 EBS"
topics: ["reference_documentation"]3. Multi-AZ ALB (1hr, minimal cost, AZ failure tolerance)
Tool: aws___search_documentation
Parameters:
search_phrase: "Application Load Balancer Multi-AZ configuration"
topics: ["reference_documentation"]1. Monitoring (1wk, $100-300/mo, faster detection)
Tool: aws___search_documentation
Parameters:
search_phrase: "CloudWatch dashboard alarm best practices"
topics: ["general", "reference_documentation"]2. Runbooks (2wk, time only, faster response)
Tool: aws___search_documentation
Parameters:
search_phrase: "AWS Systems Manager automation runbook templates"
topics: ["reference_documentation"]3. S3 Versioning/Replication (4hr, storage cost, data protection)
Tool: aws___search_documentation
Parameters:
search_phrase: "S3 versioning Cross-Region Replication setup"
topics: ["reference_documentation"]1. Chaos Engineering (ongoing, minimal cost, validate assumptions)
Tool: aws___search_documentation
Parameters:
search_phrase: "AWS FIS chaos engineering best practices experiment templates"
topics: ["general", "reference_documentation"]2. Multi-Region (4-8wk, significant cost, region resilience)
Tool: aws___search_documentation
Parameters:
search_phrase: "multi-region disaster recovery pilot light warm standby"
topics: ["general"]
Tool: aws___read_documentation
Parameters:
url: [Use URL from search for multi-region architecture guide]3. Resilience Hub (1wk, service cost, continuous assessment)
Tool: aws___search_documentation
Parameters:
search_phrase: "AWS Resilience Hub continuous assessment CI/CD integration"
topics: ["reference_documentation"]INSTRUCTIONS: After each stage, validate completeness before proceeding. Use MCP tools to find solutions for common issues.
Issue: Unrealistic RTO/RPO Solution:
Tool: aws___search_documentation
Parameters:
search_phrase: "RTO RPO targets startup realistic business requirements"
topics: ["general"]Issue: High Multi-AZ costs Solution:
Tool: aws___search_documentation
Parameters:
search_phrase: "AWS Savings Plans Reserved Instances cost optimization"
topics: ["general"]Issue: Chaos tests fail Solution:
Tool: aws___search_documentation
Parameters:
search_phrase: "chaos engineering failure analysis remediation"
topics: ["general"]Issue: Alert fatigue Solution:
Tool: aws___search_documentation
Parameters:
search_phrase: "CloudWatch composite alarms reduce alert fatigue"
topics: ["reference_documentation"]Issue: Repeated incidents Solution:
Tool: aws___search_documentation
Parameters:
search_phrase: "incident response root cause analysis permanent fix"
topics: ["general"]INSTRUCTIONS: Identify which scenario(s) apply to user’s situation and adjust recommendations accordingly. Multiple scenarios can apply simultaneously.
Approach: Priority 1 only, 90-day phases, quick wins (S3 versioning, RDS Multi-AZ)
MCP Tool Usage:
Tool: aws___search_documentation
Parameters:
search_phrase: "startup zero resilience quick wins Multi-AZ backup"
topics: ["general"]Outcome: Basic resilience 30d, production-grade 90d
Approach: Right-size to business needs, calculate downtime vs resilience cost, optimize to active-passive
MCP Tool Usage:
Tool: aws___search_documentation
Parameters:
search_phrase: "multi-region cost optimization active-passive vs active-active"
topics: ["general"]Outcome: 40-60% cost reduction
Approach: Map requirements to AWS services, implement controls, document evidence
MCP Tool Usage:
Tool: aws___search_documentation
Parameters:
search_phrase: "SOC 2 HIPAA compliance AWS resilience backup encryption"
topics: ["general"]Outcome: Compliance-ready with audit trail
Approach: Auto-scaling all tiers, load test 10x, capacity planning
MCP Tool Usage:
Tool: aws___search_documentation
Parameters:
search_phrase: "AWS Auto Scaling capacity planning load testing"
topics: ["reference_documentation", "general"]Outcome: Resilience scales with growth
Approach: Lift-and-shift multi-AZ, AWS Backup, plan microservices
MCP Tool Usage:
Tool: aws___search_documentation
Parameters:
search_phrase: "lift and shift migration Multi-AZ disaster recovery"
topics: ["general"]Outcome: Improved resilience during migration
Completeness: [ ] All 5 stages [ ] RTO/RPO defined [ ] Gaps identified [ ] Costs estimated [ ] Timelines provided [ ] AWS docs linked
Accuracy: [ ] Current state verified [ ] Configs checked [ ] Costs calculated [ ] RTO/RPO tested
Actionability: [ ] Prioritized by impact [ ] Steps documented [ ] Resources specified [ ] Success criteria defined
MCP Integration: [ ] Docs searched per stage [ ] Regional availability verified [ ] Best practices incorporated [ ] Links provided
aws-mcp) available — it ships with this plugin, so no separate install is neededTools Required:
aws-mcp with access to: aws___search_documentation, aws___read_documentation, aws___get_regional_availability, aws___recommend, aws___call_awsGather information about your startup:
Startup Profile: Company name, funding stage, industry, team size, AWS spend, MRR/ARR
Business Context: Product description, customer type (B2B/B2C), regulatory requirements, growth rate
Technical Context: Primary AWS region, deployment configuration (Single-AZ/Multi-AZ/Multi-Region), IaC tool, CI/CD and monitoring tools
AWS Services: Compute (EC2/ECS/Lambda), Database (RDS/DynamoDB), Storage (S3/EBS/Backup), Network (VPC/ALB/Route 53), Monitoring (CloudWatch/X-Ray)
Current Resilience: Uptime %, longest outage, MTTD/MTTR, backup strategy, DR plan status
Note: If you don’t know specific values (e.g., RTO/RPO targets), note as [UNKNOWN]—the assessment will use AWS industry benchmarks.
Complete Mode (Default): Executes all 5 stages sequentially, delivers final report (~15-20 min). Best for executive presentations and compliance documentation.
Interactive Mode: Pauses after each stage for review and questions (~30-45 min). Best for collaborative assessments and team learning.
Submit configured prompt. The assessment will systematically evaluate:
The final report includes:
Scenario: Series A SaaS startup, Single-AZ deployment, no DR plan, preparing for SOC2
Input:
Company: TechStartup Inc | Stage: Series A | Industry: SaaS Team: 5 engineers | AWS spend: $2K/mo | MRR: $50K Region: us-east-1 | Deployment: Single-AZ Services: EC2, RDS (Single-AZ), S3, ALB Uptime: 99.5% | No DR plan | No automated backups
This file