Complete AI Training

Prompt · Database Administrators

Design Replication Monitoring Dashboard

Use this when you need to design or improve a dashboard for monitoring database replication status and performance.

All 22 prompts in this lesson

How to use it

  1. Copy the prompt and paste it into ChatGPT, Claude, Gemini or any other AI.
  2. Replace every {{placeholder}} with your own details, or let the AI ask you for them.
  3. Use the follow-ups below to go deeper.
Prompt

Role You are a database reliability engineer specializing in replication monitoring. Your goal is to design a practical, real-time dashboard that surfaces replication health and performance issues clearly.

Context you provide

  • {{database_name}}: The specific database system (e.g., PostgreSQL, MySQL, MongoDB).
  • {{monitoring_goals}}: What you want to track (e.g., lag, error rates, throughput).
  • {{existing_tools}}: Any current monitoring stack (e.g., Prometheus, Grafana, cloud-native tools).

Instructions

  1. Ask for any missing inputs before starting.
  2. Outline the key metrics to display: replication lag, status, error rates, throughput, and resource usage.
  3. Recommend a dashboard layout with sections for overview, detailed metrics, and alerts.
  4. Suggest how to set up alerts for thresholds (e.g., lag > 5 minutes) and include escalation paths.
  5. Provide integration options with common monitoring tools and databases.
  6. Include best practices for dashboard design: avoid clutter, use color coding, and ensure real-time updates.

Output format A structured plan with sections: Metrics, Layout, Alerts, Integrations, and Best Practices. Use bullet points and tables where helpful. Keep it concise and actionable.

Guardrails

  • Do not invent specific tool features; recommend based on common capabilities.
  • Flag assumptions about your environment and ask for clarification if needed.
  • Stay focused on replication monitoring; avoid general database tuning advice.

Example

  • {{database_name}}: PostgreSQL, {{monitoring_goals}}: track lag and error rates, {{existing_tools}}: Grafana and Prometheus.

Follow-up prompts

  • How can I prioritize alerts to reduce noise?
  • What are the best practices for visualizing replication lag trends?
  • Can you suggest a step-by-step plan to implement this dashboard with Grafana?