Complete AI Training
Sign inGet my AI kit

Your job's AI kit

Get your AI kit

Tell us who you are and what you do. We show you your kit right away and email you the link: skills, prompts, AI agents, MCP servers and courses for your job.

500+ jobs ready, and we make a kit for any other job. No payment needed to look.

Share

AI agent for economists

Model Robustness Check Agent

Show how stable the key result is across reasonable specifications and document it

Model Robustness Check Agent: what goes in, what the agent does and what you get

What it does

An economist estimates a policy effect and a referee asks whether it survives other controls. This agent takes the base model and reruns it with alternative controls, samples and estimators, such as dropping outliers, adding fixed effects or using a different standard error. It compares the key coefficient across runs and flags results that change sign, lose significance or move a lot. It logs each run with the code, data version and settings, so results can be repeated. If a run fails to converge or shows data issues, it fixes the setup and reruns. It summarizes which findings are robust and which are fragile. The economist approves what is reported. Edge case: a specification that uses a bad control is marked as invalid and left out.

How it works

Follow the arrows from top to bottom. The orange dashed arrow is the loop: when a check fails, the agent goes back and tries again.

Start and resultWhat it doesA check on its own workWaits for your OKGoes back and retries
Yes, continueYes, continueApprovedNoNo 1 STARTS WHEN Base model is ready 2 USES A TOOL Run the base model and record the key coefficient 3 DOES Define alternative controls, samples and estimators 4 USES A TOOL Run each alternative and log settings 5 CHECKS THE RESULT Did every run converge with valid data? If not: Fix the setup or drop the invalid specificationand rerun. Back to step 3. 6 DOES Compare coefficient, error and sample across runs 7 CHECKS THE RESULT Is the key result stable across valid runs? If not: Add diagnostics to find the source of the changeand rerun. Back to step 3. 8 DOES Draft the robustness table and notes 9 YOU APPROVE Economist approves what is reported 10 RESULT Robustness table and run log
Read the steps as a list
  1. Base model is ready
  2. Run the base model and record the key coefficient
  3. Define alternative controls, samples and estimators
  4. Run each alternative and log settings
  5. Did every run converge with valid data?If not: Fix the setup or drop the invalid specification and rerun. Back to step 3.
  6. Compare coefficient, error and sample across runs
  7. Is the key result stable across valid runs?If not: Add diagnostics to find the source of the change and rerun. Back to step 3.
  8. Draft the robustness table and notes
  9. Economist approves what is reportedThe agent waits here for your OK.
  10. Robustness table and run log

How it decides

It marks a result fragile when the key coefficient changes sign or moves more than 25 percent in any valid specification.

  • Mark a result fragile if the sign changes in any valid run
  • Mark it fragile if the coefficient moves more than 25 percent
  • Exclude specifications with invalid controls
  • Log code version and data version for every run

Make it yours

Every agent is a starting point. You choose these settings for your own situation.

  • Stability tolerance (default 25 percent)
  • Specifications list (default controls, samples, estimators)
  • Significance level (default 5 percent)
  • Log format (default a run table)

What keeps you in control

It always asks you first

  • Economist approves what is reported and how it is described

Hard limits

  • Never change the data or the base model
  • Report every specification that was run
  • Never describe a fragile result as robust

It stops when

  • Done: Robustness table approved
  • Stop: Data problems cannot be fixed, so hand to the economist

Set it up

We guide you through the set-up, step by step

Members get the full set-up guide for this agent. No technical skills needed: you copy, paste and upload.

10 minto set it up in your AI
5 AIsChatGPT, Claude, Copilot, Gemini, Grok
  • One set of instructions to paste into your AI, with the clicks for ChatGPT, Claude, Microsoft 365 Copilot, Gemini and Grok
  • The agent then walks you through connecting your own data, one source at a time
  • A downloadable copy with the flow chart, the rules and the full guide
Get access to this agent

An example run

What happensThe base effect is 0.18 (se 0.05). Across 14 alternatives, 12 lie between 0.15 and 0.21. Dropping the top 1 percent of firms gives 0.09, a 50 percent change, so it fails the stability check. The agent adds a diagnostic and finds 4 outliers drive it. The economist reports both results.

More agents for economists