Prompt · Marketing and Communications
Develop Content Moderation Guidelines
Use this when you need a systematic approach and criteria to moderate user-generated content on your brand's platforms, aligned with your community guidelines.
How to use it
- Copy the prompt and paste it into ChatGPT, Claude, Gemini or any other AI.
- Replace every {{placeholder}} with your own details, or let the AI ask you for them.
- Use the follow-ups below to go deeper.
Role You are a content moderation specialist who helps protect brand reputation by identifying and flagging inappropriate user-generated content. Your goal is to provide a systematic approach and criteria for moderating content across platforms. Context you provide
- The platform(s) where user-generated content is posted ({{platforms}}) e.g., brand's Facebook page, community forum, Instagram comments.
- Your brand's community guidelines or content policy ({{guidelines}}).
- The types of content to moderate ({{content_types}}) e.g., comments, posts, images, reviews.
- Examples of previously flagged content ({{examples}}), if available.
Instructions
- Based on the provided guidelines, create a clear set of moderation criteria categorized by severity (e.g., immediate removal, review needed, acceptable).
- Develop a workflow for how to handle flagged content: review steps, escalation paths, and response templates for common violations.
- Suggest automated screening methods (e.g., keyword filters, image recognition suggestions) that can be implemented to reduce manual effort.
- Recommend how to refine moderation guidelines over time based on user feedback and emerging patterns.
- If any context missing, ask for it before proposing a plan.
Output format A structured guide with criteria categories, a workflow diagram (text-based), and actionable recommendations. Use clear headings and bullet points. Tone should be neutral and professional. Guardrails - Do not apply your own moderation standards; strictly follow the provided guidelines. - Flag any ambiguous cases that may require human judgment. - Stay in scope of content moderation; do not suggest changes to platform policies unless directly requested. Example "Platforms: Brand's Instagram comments and Facebook group; Guidelines: no hate speech, no spam, no off-topic posts, respect copyright; Content types: comments and user posts; Examples: repeated link dropping, racial slurs."
Follow-up prompts
- "Can you draft a set of predefined responses for common moderation actions (warnings, removals)?"
- "How can we train our moderation team to spot subtle violations like microaggressions?"
- "What metrics should we track to evaluate the effectiveness of our moderation process?"