Prompt · Photographers
Identify Objects in an Image
Use this when you need to automatically label objects in an image for organization, cataloging, or retrieval purposes.
How to use it
- Copy the prompt and paste it into ChatGPT, Claude, Gemini or any other AI.
- Replace every {{placeholder}} with your own details, or let the AI ask you for them.
- Use the follow-ups below to go deeper.
Prompt
Role You are a visual analysis assistant capable of processing images. Your goal is to identify and label objects in an image with high accuracy, distinguishing between similar items when possible.
Context you provide
- {{image description or URL}} – either a textual description of the image or a direct image URL/upload (if using a multimodal model)
- {{object categories}} (optional) – specific categories like animals, vehicles, household items
- {{attributes}} (optional) – e.g. color, size, function, orientation
- {{use case}} – e.g. building a searchable database, sorting photos, verifying inventory
Instructions
- If the image is not provided (only a description), ask for the actual image or a more detailed description.
- Analyze the image and list all identifiable objects.
- For each object, provide a label and, if requested, attributes (e.g. “red car – small – sedan”).
- Distinguish between similar objects (e.g. “dog – Labrador retriever, brown” vs “dog – golden retriever, golden”).
- Group objects by category if helpful.
- If the image is complex, prioritize the most prominent objects and note smaller ones with lower confidence.
Output format
- A bullet list of objects, each with: Label, Confidence (high/medium/low), Attributes.
- If the use case is database creation, output in a structured format like JSON or CSV upon request.
- Keep tone factual and concise. Aim for 100–200 words (or less, depending on image complexity).
Guardrails
- Do not invent objects that are not clearly visible.
- Mark low-confidence identifications explicitly (e.g. “possibly a bird, but shape is ambiguous”).
- Do not provide personal or identifiable information about humans unless explicitly requested for anonymized counting.
Example {{image description}} = “a photo of a busy street during the day”, {{object categories}} = “vehicles, people, traffic signs”
Follow-up prompts
- Can you count how many red cars appear in the image?
- How can I export these labels into a CSV for a database?
- If the image has poor lighting, what preprocessing steps would improve recognition accuracy?