Research & evidence

Treat the framework as a research agenda, not as a completed proof.

A textbook argument, a classroom simulation, a technical implementation, a pilot, and an independent evaluation are different kinds of evidence. The questions below identify what should be tested rather than assumed.

Research agendaEvidence boundariesIndependent critiqueReplication

Three evidence lanes

Do not let one kind of evidence stand in for another.

01 · Scholarship

Evidence underlying the textbook

Historical, theoretical, institutional, methodological, and public-administration claims should be judged against the scholarly literature appropriate to those claims.

Publication & editorial notes →
02 · Framework

Evidence testing the Reasoning Polity

The framework can be studied independently: whether explicit framing, provenance, adversarial review, authority mapping, public justification, monitoring, and reopening improve real decision processes.

Open research agenda →
03 · Implementation

Evidence about particular implementations

A working system can show that a mechanism is implementable without proving that it improves legitimacy, decision quality, or public outcomes. Those stronger claims need their own evaluation.

See the NousPolis project ↗

Framework research agenda

Questions that should remain open until evidence answers them.

The relevant comparison is not whether a process looks more sophisticated. It is whether the architecture changes decision quality, accountability, legitimacy, or learning in measurable and normatively defensible ways.

Decision qualityDoes structured reasoning improve decisions?

Test error rates, forecast accuracy where appropriate, option quality, implementation feasibility, and the ability to detect consequences ordinary processes missed.

Evidence qualityDoes provenance improve epistemic discipline?

Test source quality, independence, lineage, uncertainty disclosure, reproducibility, correction behavior, and resistance to citation or synthesis errors.

ContestabilityDoes adversarial review change outcomes?

Measure whether challenge identifies material assumptions, hidden stakeholders, methodological weaknesses, rights conflicts, or jurisdictional errors before authorization.

Public justificationDo reasons become more inspectable?

Evaluate whether citizens, reviewers, and institutions can reconstruct the relationship between evidence, values, dissent, authority, and the authorized decision.

Corruption resistanceDoes the record reduce room for hidden steering?

Study procurement, classification, framing, routing, conflict-of-interest, override, and record-integrity mechanisms without assuming transparency alone prevents capture.

Institutional memoryDo decisions learn better over time?

Test whether versioned decision records, implementation baselines, monitoring, appeal, and reopening reduce repeated errors or institutional amnesia.

Participation qualityDoes participation become more representative and useful?

Measure access, missing voices, information quality, standing, representation claims, deliberative effects, and whether participation changes the reasoning rather than serving as decoration.

Concentrated harmAre small but severe burdens detected earlier?

Test whether stakeholder and distributional records reveal harms hidden by averages, and whether those findings actually reach authorized decision-makers.

Implementation fidelityCan the polity see the gap between decision and delivery?

Compare the authorized baseline with procurement, administrative adaptation, frontline discretion, outcomes, and the reasons for material divergence.

Correction & reopeningCan institutions reverse course without erasing history?

Study trigger quality, timeliness, rollback feasibility, reauthorization, correction costs, and whether reopening becomes either impossible or permanently destabilizing.

Evidence language

Describe what has actually been shown.

Useful distinctions include proposed, implemented, tested, externally reviewed, validated, and contested. Apply them to a specific claim or mechanism rather than using one maturity word to summarize an entire project.

Independent criticism, failed replications, negative findings, and mixed evidence belong in the record alongside successes. Future editions should be able to revise the framework in response.