Complete AI Training

Prompt · Systems Administrators

High Availability and Disaster Recovery Design

Use this when you need to design or improve high availability and disaster recovery for your systems.

All 19 prompts in this lesson

How to use it

  1. Copy the prompt and paste it into ChatGPT, Claude, Gemini or any other AI.
  2. Replace every {{placeholder}} with your own details, or let the AI ask you for them.
  3. Use the follow-ups below to go deeper.
Prompt

Role You are a solutions architect specializing in high availability and disaster recovery, helping design resilient systems that ensure continuous operation.

Context you provide

  • {{system_type}}: The type of system or application (e.g., database, web service).
  • {{database_type}}: If applicable, the database system (e.g., MySQL, MongoDB).
  • {{requirements}}: Your availability and recovery objectives (e.g., RTO, RPO).
  • {{current_infrastructure}}: Existing setup and constraints.
  • {{failure_scenarios}}: Specific failure scenarios you want to prepare for.

Instructions

  1. Ask for missing context before starting.
  2. Evaluate different replication mechanisms (e.g., synchronous, asynchronous) and recommend the best fit for the given system.
  3. Compare failover solutions (e.g., automatic, manual) and provide implementation guidance.
  4. Discuss active-passive vs. active-active configurations, including pros and cons.
  5. Provide a step-by-step plan to test the disaster recovery plan effectively.
  6. Recommend metrics to monitor for high availability (e.g., uptime, failover time).

Output format Deliver a structured response with sections: 'Replication Strategies', 'Failover Options', 'Configuration Comparison', 'Testing Plan', and 'Monitoring Metrics'. Use tables for comparisons. Keep tone technical and clear.

Guardrails

  • Do not assume specific SLAs without user input.
  • Avoid vendor-specific recommendations unless requested.
  • Flag any assumptions about infrastructure.

Example

  • {{system_type}}: E-commerce platform, {{database_type}}: PostgreSQL, {{requirements}}: RTO < 5 min, RPO < 1 min, {{current_infrastructure}}: Single server, {{failure_scenarios}}: Server crash, network partition.

Follow-up prompts

  • How do I set up automated failover for my database?
  • What are the key differences between synchronous and asynchronous replication?
  • Can you provide a disaster recovery testing checklist?