Complete AI Training

Prompt · Global Heads of IT

Network Performance Monitoring System Design

Use this when you need to design a continuous network performance monitoring system that identifies anomalies and suggests improvements.

All 22 prompts in this lesson

How to use it

  1. Copy the prompt and paste it into ChatGPT, Claude, Gemini or any other AI.
  2. Replace every {{placeholder}} with your own details, or let the AI ask you for them.
  3. Use the follow-ups below to go deeper.
Prompt

Role — You are a network operations architect with deep experience in performance monitoring and analytics. Your goal is to design a robust, real-time monitoring system that detects anomalies and provides actionable insights for improvement.

Context you provide

  • {{network metrics}} — e.g., "latency, packet loss, throughput, jitter"
  • {{current tools}} — e.g., "SNMP, NetFlow, custom scripts"
  • {{infrastructure scale}} — e.g., "500 devices, 10 Gbps backbone"
  • {{goals}} — e.g., "reduce downtime by 20%, identify bottlenecks, automate alerts"
  • {{compliance requirements}} — e.g., "PCI-DSS, SOC 2"

Instructions

  1. Ask for any missing context, especially critical applications and peak usage times.
  2. Define the key performance indicators (KPIs) to monitor based on the provided metrics.
  3. Outline a system architecture that includes data collection, processing, storage, and alerting layers.
  4. Describe how to set up baseline thresholds for anomaly detection and how alerts should be generated (e.g., severity levels, escalation paths).
  5. Recommend proactive measures (e.g., capacity planning, traffic shaping) based on historical trends.
  6. Provide a phased implementation plan with milestones and success criteria.

Output format A monitoring system design document in markdown: Overview, KPI definitions, Architecture diagram (textual), Alerting rules, Proactive measures, Implementation plan. Tone: technical and precise. Length: 500–700 words.

Guardrails

  • Do not assume specific commercial tools; describe capabilities generically unless the user names a tool.
  • Ensure all recommendations respect given compliance requirements.
  • Avoid sharing actual network data; keep all examples hypothetical.

Example

  • {{network metrics}}: "latency, packet loss, throughput"
  • {{current tools}}: "SNMP, custom scripts"
  • {{infrastructure scale}}: "200 devices, 1 Gbps backbone"
  • {{goals}}: "detect anomalies within 5 minutes, reduce mean time to resolution"
  • {{compliance requirements}}: "SOC 2"

Follow-up prompts

  • How can we integrate this monitoring system with our existing incident management platform?
  • What are the best practices for setting dynamic thresholds instead of static ones?
  • Can you help me create a dashboard mockup for the key KPIs?