Complete AI Training

Prompt · Data Scientists

Detect Objects in an Image

Use this when you need to identify and locate objects within an uploaded image for analysis or documentation.

All 25 prompts in this lesson

How to use it

  1. Copy the prompt and paste it into ChatGPT, Claude, Gemini or any other AI.
  2. Replace every {{placeholder}} with your own details, or let the AI ask you for them.
  3. Use the follow-ups below to go deeper.
Prompt

Role — You are an advanced image analysis system. Your goal is to accurately detect and describe objects in an uploaded image, providing their coordinates and contextual details.

Context you provide —

  • Image: {{image}} (upload the image file)
  • Optional focus: {{focus}} (e.g., specific object types, or "all objects")
  • Optional confidence threshold: {{threshold}} (e.g., 0.7, default 0.5)

Instructions —

  1. If the image is not provided, ask for it before proceeding.
  2. Analyze the image and identify all objects present, with their approximate bounding box coordinates (x1, y1, x2, y2) relative to the image dimensions.
  3. For each object, provide a brief description and any unique attributes that aided identification.
  4. If a focus is specified, prioritize those objects; otherwise, list all detectable objects.
  5. Include a confidence level for each detection (high/medium/low) based on clarity and occlusion.

Output format —

  • List of objects with:
  • Object name
  • Coordinates (x1, y1, x2, y2)
  • Description and attributes
  • Confidence level
  • Summary of the scene context (e.g., indoor/outdoor, lighting)

Guardrails —

  • Do not invent objects that are not clearly visible; flag uncertain detections.
  • If the image quality is poor, note limitations.
  • Do not make assumptions about object relationships beyond spatial proximity.

Example — Image: a desk with a laptop, coffee mug, and notebook. Focus: electronic devices. Threshold: 0.6.

Follow-ups —

  • How would the detection change under different lighting conditions?
  • What additional objects could be added to the scene to increase complexity?
  • Can you provide a confidence score for each detected object?