From docs/roadmap.md.

Roadmap

This file holds the current priority, the decisions waiting on the owner and the status of each milestone. Completed work is recorded in the research log, which also keeps everything this roadmap recorded up to 2026-10-06. Registered sources are listed in source additions.

Current priority: the French Revolutionary and Napoleonic Wars (frame built)

The owner chose the Napoleonic Wars on 2026-09-25 and decided the seven scoping questions on 2026-10-06 (scoping note, record):

Done (2026-10-06): the frame is built and frozen (design).

Next:

  1. A Napoleonic evidence path: a cohort-aware dossier validator, its own directory, and a research brief applying decisions 5–7.
  2. First passes by complete campaign group, starting with third-coalition 1805 (22 entries), logging usage per campaign and giving the owner a projection after the first.

No dossiers have been drafted. The Civil War work below is complete as recorded, and its cohorts, ledgers and runs are unchanged.

Open owner decisions

The American Civil War: where things stand

Milestones

MilestoneStatus
0. Working research foundationComplete (infrastructure only)
1. First reviewed campaign dossiersCoverage complete. Cross-engagement identities, replacement boundaries and information sets remain
2. Baseline audit and expanded coverageFull-war coverage done. Arsht reproduction route, independent extraction sample and paired comparison rows remain
3. Enriched prediction and command inferenceExploratory rating runs only. Campaign estimand, causal graph and locked test set remain
4. Inspectable research interfaceStarted: a static explorer. Opportunity-adjusted estimates, disputes views and scenario toggles remain

Milestone 0: working research foundation (complete)

"Complete" applies to this infrastructure milestone only. Historical validation, AI extraction quality, command attribution and better estimates of generalship have not been established.

Milestone 1: first reviewed campaign dossiers

Deliverables:

Acceptance requires that:

Milestone 2: baseline audit and expanded coverage

Acceptance requires an immutable cohort, an exclusion ledger, an independent extraction sample, and paired comparison rows defined before model selection. Sparse evidence must not quietly become mediocre commander scores.

Milestone 3: enriched prediction and command inference

Acceptance requires:

A familiar-looking list of great generals is not evidence of correctness.

Milestone 4: inspectable research interface

Review policy

Owner policy since 2026-10-06 (AGENTS.md): there is no separate design or evidence review step. The primary agent verifies its own work against the sources and checks before committing, and that work is labelled primary-verified.

Earlier work had separate AI reviews:

Those reviews remain records of their own scope. A separate review runs only when the owner asks for one.

Deferred implementation choices

Owner ideas to design later