Evidence underlying the textbook
Historical, theoretical, institutional, methodological, and public-administration claims should be judged against the scholarly literature appropriate to those claims.
Publication & editorial notes →Research & evidence
A textbook argument, a classroom simulation, a technical implementation, a pilot, and an independent evaluation are different kinds of evidence. The questions below identify what should be tested rather than assumed.
Three evidence lanes
Historical, theoretical, institutional, methodological, and public-administration claims should be judged against the scholarly literature appropriate to those claims.
Publication & editorial notes →The framework can be studied independently: whether explicit framing, provenance, adversarial review, authority mapping, public justification, monitoring, and reopening improve real decision processes.
Open research agenda →A working system can show that a mechanism is implementable without proving that it improves legitimacy, decision quality, or public outcomes. Those stronger claims need their own evaluation.
See the NousPolis project ↗Framework research agenda
The relevant comparison is not whether a process looks more sophisticated. It is whether the architecture changes decision quality, accountability, legitimacy, or learning in measurable and normatively defensible ways.
Test error rates, forecast accuracy where appropriate, option quality, implementation feasibility, and the ability to detect consequences ordinary processes missed.
Test source quality, independence, lineage, uncertainty disclosure, reproducibility, correction behavior, and resistance to citation or synthesis errors.
Measure whether challenge identifies material assumptions, hidden stakeholders, methodological weaknesses, rights conflicts, or jurisdictional errors before authorization.
Evaluate whether citizens, reviewers, and institutions can reconstruct the relationship between evidence, values, dissent, authority, and the authorized decision.
Study procurement, classification, framing, routing, conflict-of-interest, override, and record-integrity mechanisms without assuming transparency alone prevents capture.
Test whether versioned decision records, implementation baselines, monitoring, appeal, and reopening reduce repeated errors or institutional amnesia.
Measure access, missing voices, information quality, standing, representation claims, deliberative effects, and whether participation changes the reasoning rather than serving as decoration.
Test whether stakeholder and distributional records reveal harms hidden by averages, and whether those findings actually reach authorized decision-makers.
Compare the authorized baseline with procurement, administrative adaptation, frontline discretion, outcomes, and the reasons for material divergence.
Study trigger quality, timeliness, rollback feasibility, reauthorization, correction costs, and whether reopening becomes either impossible or permanently destabilizing.
Evidence language
Useful distinctions include proposed, implemented, tested, externally reviewed, validated, and contested. Apply them to a specific claim or mechanism rather than using one maturity word to summarize an entire project.
Independent criticism, failed replications, negative findings, and mixed evidence belong in the record alongside successes. Future editions should be able to revise the framework in response.