Prompt · Bloggers
Inappropriate Comment Flagging System
Use this when you need to design a system to flag inappropriate comments on your online community or platform.
How to use it
- Copy the prompt and paste it into ChatGPT, Claude, Gemini or any other AI.
- Replace every {{placeholder}} with your own details, or let the AI ask you for them.
- Use the follow-ups below to go deeper.
Role You are an expert in online community moderation and content policy. Your goal is to design a practical system for flagging inappropriate comments that aligns with community guidelines.
Context you provide
- {{community_type}}: Describe the platform or community (e.g., blog, forum, social media page).
- {{existing_guidelines}}: If any, list the existing community guidelines or rules.
- {{specific_concerns}}: Mention any particular types of inappropriate content you want to focus on (e.g., hate speech, personal attacks, spam, illegal activity).
Instructions
- If the user has not provided the context, ask for the community type and existing guidelines.
- Based on the context, define clear criteria for flagging comments, covering the specific concerns mentioned.
- Outline a step-by-step process for implementing a flagging system, including how to train moderators, how to handle automated flagging (using AI or keyword lists), and how to escalate flagged content.
- Provide recommendations for ensuring transparency and fairness in the flagging process.
Output format Present the system as a structured guide with sections: Criteria, Implementation Steps, Training Plan, Escalation Process, and Transparency Measures. Use bullet points for clarity. Total length: 400–600 words.
Guardrails
- Do not suggest specific third-party tools unless they are widely known and free; focus on principles.
- Ensure the system respects free speech and avoids over-censorship; suggest appeals process.
- If the user mentions illegal content, remind them to consult legal counsel.
Example {{community_type}} = "A parenting blog comment section", {{existing_guidelines}} = "No personal attacks, no spam", {{specific_concerns}} = "Hate speech and promotion of violence".
Follow-up prompts
- How can I automate the initial flagging using AI without over-flagging?
- What should be the appeals process for users whose comments are flagged?
- How can I measure the effectiveness of the flagging system over time?