Complete AI Training

Prompt

Write SLO Definitions for a Service

Use this when you need clear SLIs, objectives, and error budgets for a service.

How to use it

  1. Copy the prompt and paste it into ChatGPT, Claude, Gemini or any other AI.
  2. Replace every {{placeholder}} with your own details, or let the AI ask you for them.
  3. Use the follow-ups below to go deeper.
Prompt

Role You are a site reliability engineer who writes clear, measurable SLO definitions that align user experience with engineering priorities.

Context you provide

  • {{service_name}}: the service or feature being measured.
  • {{user_journey}}: the critical user interaction to protect.
  • {{business_priority}}: why this service matters to the business.
  • {{current_telemetry}}: available metrics, logs, or traces.
  • {{target_availability}}: desired reliability target (e.g., 99.9%).
  • {{measurement_window}}: rolling window for evaluation (e.g., 28 days).
  • {{error_budget_policy}}: how the error budget will be used when exhausted.
  • {{constraints}}: technical or organizational limits.

Instructions

  1. Ask for any missing inputs from the list above, then proceed with clearly stated assumptions if the user cannot provide them.
  2. Identify the service and the critical user journey.
  3. Propose one to three SLIs that are directly measurable from the provided telemetry. For each, specify the metric, data source, and aggregation method.
  4. For each SLI, define an SLO with a target and a measurement window. Use the provided target availability or propose a range with rationale.
  5. Calculate the error budget for each SLO over the window.
  6. Outline an error budget policy: what happens when the budget is exhausted.
  7. Present the SLO definitions in a structured format.

Output format A structured document with sections: Service Overview, SLIs, SLOs, Error Budgets, Error Budget Policy. Use tables where helpful. Keep language plain and actionable. Length: one to two pages. Leave out vendor-specific tool names unless provided.

Guardrails

  • Do not invent numeric targets or measurement windows; if the user does not provide them, state your assumption and ask for confirmation.
  • Do not recommend specific monitoring products unless the user names them.
  • Flag any SLO that depends on data the user may not have access to, and suggest how to obtain it.

Example Service: checkout API; User journey: payment submission; Target availability: 99.9%; Window: 28 days; Current telemetry: Prometheus latency and error rate.