Complete AI Training
Sign inGet my AI kit

Your job's AI kit

Get your AI kit

Tell us who you are and what you do. We show you your kit right away and email you the link: skills, prompts, AI agents, MCP servers and courses for your job.

500+ jobs ready, and we make a kit for any other job. No payment needed to look.

Share

AI agent for user experience designers

Competitor Experience Benchmark Agent

A current, comparable record of competitor task flows, with changes since the last run highlighted.

Competitor Experience Benchmark Agent: what goes in, what the agent does and what you get

What it does

A competitor changes its signup flow, and your team keeps arguing from a screenshot taken a year ago. This agent keeps the benchmark current. Every quarter it walks defined tasks on competitor sites and apps, such as signing up, finding pricing or canceling, and records the steps, friction points and screenshots. It compares step counts and patterns with your own product and flags changes since the last run. If a task fails midway, for example a login wall or a changed button, it reruns with another route or notes why it could not complete. You approve the benchmark before it is shared. Edge case: a task needing a paid account is marked not tested, not guessed.

How it works

Follow the arrows from top to bottom. The orange dashed arrow is the loop: when a check fails, the agent goes back and tries again.

Start and resultWhat it doesA check on its own workWaits for your OKGoes back and retries
Yes, continueYes, continueApprovedNoNo 1 STARTS WHEN Quarterly run 2 USES A TOOL Load tasks, competitors and the last benchmark 3 USES A TOOL Walk each task on each competitor and capture steps 4 CHECKS THE RESULT Did the task finish and capture every step? If not: Retry by another route, or mark the task nottested with the reason. Back to step 3. 5 DOES Count steps, fields and friction points 6 DOES Compare with our flow and the last benchmark 7 DOES Flag changes and new patterns 8 CHECKS THE RESULT Are the screenshots and counts consistent for eachtask? If not: Redo the capture and recount. Back to step 4. 9 USES A TOOL Build the benchmark report 10 YOU APPROVE Researcher approves the benchmark before sharing 11 RESULT Competitor benchmark
Read the steps as a list
  1. Quarterly run
  2. Load tasks, competitors and the last benchmark
  3. Walk each task on each competitor and capture steps
  4. Did the task finish and capture every step?If not: Retry by another route, or mark the task not tested with the reason. Back to step 3.
  5. Count steps, fields and friction points
  6. Compare with our flow and the last benchmark
  7. Flag changes and new patterns
  8. Are the screenshots and counts consistent for each task?If not: Redo the capture and recount. Back to step 4.
  9. Build the benchmark report
  10. Researcher approves the benchmark before sharingThe agent waits here for your OK.
  11. Competitor benchmark

How it decides

It compares step count, time to finish, number of fields and friction points for the same task, and flags a change when any differs from the last run.

  • Use the same task script every quarter
  • Mark tasks needing payment as not tested
  • Flag a change when step count changes by 1 or more
  • Include screenshots for every flagged change

Make it yours

Every agent is a starting point. You choose these settings for your own situation.

  • Competitors and tasks
  • Run schedule
  • Change threshold
  • Screenshot detail

What keeps you in control

It always asks you first

  • Researcher approves the benchmark before it is shared

Hard limits

  • Never create paid accounts or enter real personal data
  • Never sign up under false identity

It stops when

  • Done: benchmark approved
  • Stop: site blocks automated access

Set it up

We guide you through the set-up, step by step

Members get the full set-up guide for this agent. No technical skills needed: you copy, paste and upload.

10 minto set it up in your AI
5 AIsChatGPT, Claude, Copilot, Gemini, Grok
  • One set of instructions to paste into your AI, with the clicks for ChatGPT, Claude, Microsoft 365 Copilot, Gemini and Grok
  • The agent then walks you through connecting your own data, one source at a time
  • A downloadable copy with the flow chart, the rules and the full guide
Get access to this agent

An example run

What happensOn the pricing page task, Competitor A went from 3 steps to 5, adding a required demo form. The signup task failed first at a phone verification, so the agent noted it as not tested after one retry. Our own signup was 4 steps. The report highlighted A's change and the researcher approved it.

More agents for user experience designers