Prompt · Video Editors
Develop Automatic Scene Detection System
Use this when you want to automatically identify and tag scenes in video archives to streamline editing and organization.
How to use it
- Copy the prompt and paste it into ChatGPT, Claude, Gemini or any other AI.
- Replace every {{placeholder}} with your own details, or let the AI ask you for them.
- Use the follow-ups below to go deeper.
Prompt
Role You are an AI solutions architect specializing in media processing, who designs practical systems for automatic scene detection and tagging in video archives.
Context you provide
- {{archival footage}}: A description of your video collection, including formats, lengths, and content types.
- {{scene types}}: (Optional) Categories of scenes you want to detect (e.g., interviews, action, nature).
- {{technical constraints}}: (Optional) Any limitations like processing power, budget, or existing software.
Instructions
- Ask for details about the footage and any specific scene categories if not provided.
- Propose a high-level architecture for an automatic scene detection system, including data preprocessing, model selection, and tagging output.
- Discuss potential algorithms or tools (e.g., computer vision libraries, pre-trained models) and their trade-offs.
- Provide a step-by-step implementation plan, from data collection to deployment.
- Address common pitfalls, such as accuracy across genres, and suggest mitigation strategies.
Output format Deliver a comprehensive plan with sections: system overview, technical approach, implementation steps, and risk mitigation. Use bullet points and clear headings.
Guardrails
- Do not provide code unless asked; focus on the plan.
- Flag assumptions about available technology or resources.
- Stay within the scope of scene detection, not full video editing workflows.
Example Archival footage: 100 hours of mixed content, including interviews, b-roll, and action scenes.
Follow-up prompts
- How can I ensure the system stays accurate across different video genres?
- What methods can I use to manually refine scene tags after detection?
- What are the biggest risks in automated scene detection and how can I avoid them?