Public work plan · Version 1.0

Research agenda

Validating
3 contributions
Next
3 studies
Queued
3 research questions
Last updated
July 21, 2026

This research agenda is the public work plan for Too Early To Say, Victoria Cholette's independent applied-economics lab. AI agents draft code, search records, and run candidate analyses. Victoria sets the economic question, identification strategy, checks, and interpretation.

The agenda is for economists, policy teams, and research partners who need evidence about where AI saves work, where it creates review costs, and which institutional checks hold up. A tool can finish a task quickly, while checking the result can erase the gain.

The research program tests whether AI shifts the cost of policy research from execution to specification, verification, and judgment.

Current and next studies

Status reflects the evidence currently available. Validating work remains provisional. Next work has an approved question or design. Queued work follows the core studies. Parked work waits for an experiment or named policy case.

Three connected programs

Evidence program

AI and evidence production

Evidence on time saved, errors introduced, and review work required.

In Difference-in-Differences in Python, a planted effect of 1.60 tests whether the estimator answers the intended economic question.

Agents: implement and rerun candidates. Researcher: defines the estimand, answer key, and failure rule.

Policy program

Policy applications

Policy findings with traceable data, methods, and limits.

In The Food Desert Myth, SNAP participation varies 4.7-fold across neighborhoods with equal store access.

Agents: assemble and check source files. Researcher: defines access, defends the comparison, and interprets the finding.

Infrastructure program

Open measurement infrastructure

Reusable checks for clean-looking failures.

In Instrumental Variables in Python, an F-statistic of 2,051 coexists with a causally invalid estimate.

Agents: run estimators and diagnostics. Researcher: defines the data-generating process and decides which claim the evidence supports.

Validating Work held to a release gate

Known-Truth Applied Econometrics Test Suite v0.1

The release standardizes the existing estimator cases into a clean-session runner, saved outputs, documented failure conditions, and a citation record.

Agent work
Rerun and reconcile cases.
Researcher decision
Approve planted truths, tolerances, and release claims.
Release gate
Fresh-session reproduction and output reconciliation.

One policy number across six surfaces

The Food Desert Myth estimate is being traced through the article, code, data, findings page, metadata, and distribution copy.

Agent work
Flag mismatches and missing context.
Researcher decision
Set the source of record and portable caveat.
Release gate
One reconciled number, unit, population, and limitation.

Special Education Spending Puzzle: held estimate

The current analysis produces a provisional spending-gap estimate. The value stays out of the public findings while source reconciliation, model checks, and interpretation review continue.

Agent work
Rebuild and reconcile outputs.
Researcher decision
Determine whether the estimate supports a policy claim.
Release gate
Source audit, model checks, and interpretation review.

Queued Questions after the core studies

AI-generated variables and valid policy inference

When an agent constructs a treatment, outcome, or covariate, how does classification error move the policy estimate?

Cross-document policy provenance

How can a policy claim retain its source, units, population, and caveat from analysis through public communication?

Identification gates after external adjudication

Which agent-written checks survive review by an economist who did not design the workflow?

Parked Work awaiting a contribution design

Generic workflow posts, model leaderboards, and standalone tutorials remain parked until an experiment or named policy case supports each claim. Completeness remains the coverage goal, but publication requires a checkable contribution.

What grounds this agenda

Every agenda item extends a published case. Each foundation links a result, a research decision, and the limit the next study must address.

Contribution and publication standard

A contribution clears the agenda only when the claim is tied to a named case, the test can fail, and the evidence package says what a reader can inspect.

  1. State the economic question and estimand before the implementation.
  2. Freeze the task, benchmark, or data-generating process.
  3. Separate agent execution from researcher judgment.
  4. Save outputs, errors, corrections, and version information.
  5. Name the limitation the check cannot resolve.
  6. Release only after a fresh-session or external reproduction.
Victoria Cholette, applied economist

External review is part of the program

Economists, policy teams, and research partners can contribute a domain case, rerun a protocol, adjudicate an error, or challenge a publication gate.

Propose a collaboration