Skip to main content

Claude/ChatGPT Prompt to Design Multi-Region AWS Disaster Recovery

Design a multi-region AWS disaster recovery strategy: pick pilot light, warm standby, or active-active to hit your RTO and RPO targets, with failover steps.

Fill in the placeholders

Edit the values, then copy your finished prompt.

Your Prompt
prompt.txt
Design a disaster recovery strategy for my SaaS platform with real-time data processing hosted in us-east-1 with failover to us-west-2. RTO target: 15 minutes, RPO target: 1 minute. 1) DR strategy selection: evaluate pilot light, warm standby, and multi-site active-active — recommend based on our RTO/RPO and DR infrastructure should be under 30% of primary cost. 2) Data replication: configure Aurora PostgreSQL Global Database cross-region replication with Aurora Global Database with sub-second replication and verify replication lag monitoring. 3) Application layer: set up ECS Fargate behind an ALB in the secondary region with scaled-down warm standby at 25% of primary capacity. 4) DNS failover: configure Route 53 health checks on /api/health on the primary region ALB with failover routing policy and 3 consecutive failed health checks over 90 seconds conditions. 5) Storage replication: configure S3 Cross-Region Replication for user-uploads and application-config buckets with replication rules and lifecycle alignment. 6) Secrets and config: replicate SSM parameters and Secrets Manager secrets to secondary region using Secrets Manager multi-region replicas plus a sync Lambda for SSM. 7) Failover runbook: create a step-by-step manual and automated failover procedure with validation checks at each step. 8) DR testing: design a quarterly DR test plan that validates regional outage, database failover, and DNS failover validation without impacting production.

What this prompt does

This prompt designs a multi-region AWS disaster recovery strategy matched to your actual recovery targets. You give it your [application_type], [primary_region], [secondary_region], an [rto], and an [rpo], and it evaluates pilot light, warm standby, and active-active before recommending one based on those targets and your [budget_constraint]. From there it specifies data replication for [database], application-layer setup in the secondary region, Route 53 DNS failover, S3 cross-region replication, and secrets replication.

The structure works because it refuses to skip the unglamorous parts. Most DR plans stop at the architecture diagram, but this prompt forces a step-by-step failover runbook and a quarterly DR test plan covering [test_scenarios] without touching production. The [rto] and [rpo] variables are the anchor — a 15-minute RTO and 1-minute RPO pushes the recommendation toward warm standby or active-active, while looser targets allow cheaper pilot light. The [failover_trigger] and [health_endpoints] shape exactly when Route 53 cuts over.

When to use it

  • An application has reached the point where downtime carries real cost and needs a deliberate DR posture
  • You need to choose between pilot light, warm standby, and active-active against honest [rto] and [rpo] targets
  • You want cross-region replication designed for [database] with replication-lag monitoring
  • You need a Route 53 failover policy tied to [health_endpoints] and a clear [failover_trigger]
  • Your DR plan lacks a written failover runbook with validation at each step
  • You want a repeatable quarterly DR test covering [test_scenarios] that does not impact production

Example output

Expect a strategy document: a comparison of the three DR patterns with a recommendation justified by your [rto], [rpo], and [budget_constraint]; replication configuration for your database and S3 buckets; a Route 53 failover design; a numbered failover runbook with validation checks at each step; and a quarterly test plan. The runbook and test plan are the operational deliverables teams most often skip.

Pro tips

  • Set [rto] and [rpo] honestly — tighter targets justify more expensive standby capacity, so don't claim a 1-minute RPO unless the business truly needs it
  • Use [budget_constraint] to keep the recommendation realistic; capping DR at a percentage of primary cost steers the model toward warm standby over full active-active
  • Be precise about [secondary_config] — "warm standby at 25% of primary" produces a very different cost and failover profile than full-capacity
  • Define [failover_trigger] carefully; too sensitive and you flap between regions, too lax and you exceed your RTO during a real outage
  • The runbook is only as good as your testing, so actually run the quarterly [test_scenarios] — an untested DR plan is a guess
  • Validate replication-lag monitoring against your real [rpo]; the model proposes the design, but you must confirm lag stays within target under load

Frequently Asked Questions

How does the prompt decide between pilot light, warm standby, and active-active?
It weighs your `[rto]` and `[rpo]` targets against `[budget_constraint]`. Tight recovery targets like a 15-minute RTO push toward warm standby or active-active, while looser targets and a smaller budget make pilot light a reasonable fit. You provide the targets; the prompt justifies the trade-off.
Does it include a failover runbook or just the architecture?
It produces both. The prompt explicitly requires a step-by-step failover procedure with validation checks at each step, plus a quarterly DR test plan covering your `[test_scenarios]`. These operational pieces are the parts most DR designs omit, which is why the prompt forces them.
Can it design DR for any database, not just Aurora?
Yes. Set `[database]` and `[replication_method]` to your actual stack, whether that is Aurora Global Database, RDS read replicas, or DynamoDB global tables. The replication design adapts, though achievable RPO varies by service, so confirm the proposed method can meet your stated target.
Will following this guarantee I hit my RTO and RPO?
No. The prompt designs toward your targets, but real recovery time depends on replication lag, automation quality, and how often you test. Treat the output as a design to validate through the quarterly tests it recommends, not a guarantee — an untested DR plan rarely performs as written.
Engr Mejba Ahmed

Need this built for real?

Engr Mejba Ahmed

AI Developer · Software Engineer

I'm Mejba — I design and ship production AI systems, automations, and full-stack apps. If you want this turned into a working solution for your team, let's talk.

More in AWS & Cloud Architecture Prompts

Engr Mejba Ahmed

Engr Mejba Ahmed

AI assistant · trained on my work

👋

Hey there!

Quick Actions

WhatsApp Direct line to me

Chat on WhatsApp

+880 1723 741224 · Replies within the hour on working days

Popular Questions

Engr Mejba Ahmed is connected
Engr Mejba Ahmed is typing...
Engr Mejba Ahmed avatar

✉ Want me to follow up? Drop your email

Engr Mejba Ahmed avatar

📞 Connect Directly

Choose how you'd like to reach me

WhatsApp

+880 1723 741224

Email

mejba.13@gmail.com

✓ Details sent! I'll get back to you shortly.

Powered by OpenAI

335+

Blog Posts

25

AI Courses

63

Projects

Services & Expertise

Pricing & Process

Learning & Resources

Connect & Support