Tech4Humanity AtlasGround ZeroCurrent ThemesFuture ResearchGalleryLive Q&ASearch

Institutional Safety, Governance & Trust / Public Sector AI Capability

SUB-T06-062 · Evidence — SEEDED / PARTIAL — defensible non-empirical baseline; validation and replayable search outstanding

Public Value Measurement

1. Hypothesis

An explicit, testable and continuously evidenced approach to public value measurement, with clear ownership, independent review, runtime telemetry and recovery, will outperform policy-only or periodic compliance approaches.

2. Experiment design

Design: Multi-site institutional governance study focused on Public Value Measurement, combining baseline maturity assessment, controlled implementation, live telemetry review and post-incident or post-exercise evaluation. Methods: capability maturity assessment; procurement review; service-design research; workforce study; benefit evaluation; rights-impact assessment; stakeholder interviews; document and control review; fault and incident simulation; longitudinal implementation assessment; methods adapted specifically to Public Value Measurement Independent variables: control design; ownership clarity; review independence; telemetry coverage; enforcement level; organisational maturity; system risk Dependent variables: readiness; delivery quality; public value; workforce capability; rights protection; service accessibility; incident impact; recovery time; stakeholder confidence Confounders: sector; scale; legal context; legacy systems; budget; workforce capability; procurement model; incident history; political and public pressure Measures: readiness; delivery quality; public value; workforce capability; rights protection; service accessibility; control coverage; implementation fidelity; exception rate; subgroup and rights impacts; stakeholder comprehension; cost and time to evidence Success criteria: Material improvement in measured governance outcomes; clear ownership and authority; controls operate as intended; evidence is complete and reproducible; exceptions are bounded; no disproportionate rights harm; recovery and learning are demonstrated. Failure conditions: No measurable improvement; controls exist only on paper; evidence is missing or stale; authority is unclear; telemetry cannot observe execution; exceptions become routine; recovery fails; costs or harms exceed public value.

3. Seed result / current evidence

DEFENSIBLE SEED RESULT — NON-EMPIRICAL. The current evidence supports Public Value Measurement as a testable research proposition. Problem basis: Public Value Measurement is often described in policy or documentation but not consistently implemented, observed or evidenced at runtime, creating gaps between institutional claims and actual behaviour. Directional expectation: If supported, the approach should improve readiness; delivery quality; public value; workforce capability; rights protection; service accessibility, reduce control drift and incident impact, and increase justified stakeholder confidence. Proposed observations: readiness; delivery quality; public value; workforce capability; rights protection; service accessibility; control coverage; implementation fidelity; exception rate; subgroup and rights impacts; stakeholder comprehension; cost and time to evidence. Seed data profile: Evidence Strength 10/100; Confidence 25/100; Maturity 20/100; Overall Health 35/100; Novelty 75/100; Strategic Importance 95/100. Evidence boundary: No validated results yet.; experiments 0, studies 0, participants 0. This is suitable for protocol formation and baseline comparison, not as a finding of effect.

4. Seed conclusion

DEFENSIBLE SEED CONCLUSION — PROVISIONAL. Public Value Measurement warrants structured testing because the CSV identifies a defined problem, falsifiable hypothesis, measurable outcomes and relevant literature foundations. The present position is that “An explicit, testable and continuously evidenced approach to public value measurement, with clear ownership, independent review, runtime telemetry and recovery, will outperform policy-only or periodic compliance approaches.” is plausible and decision-relevant, but unvalidated. Proceed to controlled testing against the stated success and failure conditions. Confirm, narrow or reject this seed after effect sizes, uncertainty, subgroup outcomes, adverse effects, persistence and handback performance are observed.

Prior-art search performed before starting

PRIOR-ART SEED BASELINE — PARTIAL. The CSV records these literature domains: Institutional governance; assurance; audit; administrative law; public-sector management; risk and safety engineering; ethics; standards and regulation; literature specific to Public Value Measurement. It also records: Australian Government Digital Transformation Agency — https://www.dta.gov.au/; OECD AI in government — https://oecd.ai/; World Bank GovTech — https://www.worldbank.org/en/programs/govtech; UK Central Digital and Data Office — https://www.gov.uk/government/organisations/central-digital-and-data-office. Evidence register status: “Seeded; authoritative source register initiated; institutional evidence not yet ingested”. This is defensible as a starting prior-art inventory, but not as proof of a completed systematic search because search dates, databases, exact queries, reviewer, result counts, screening decisions, claim mapping and a replayable receipt are absent.

Prior-art material named: Existing literature: Institutional governance; assurance; audit; administrative law; public-sector management; risk and safety engineering; ethics; standards and regulation; literature specific to Public Value Measurement. References: Australian Government Digital Transformation Agency — https://www.dta.gov.au/; OECD AI in government — https://oecd.ai/; World Bank GovTech — https://www.worldbank.org/en/programs/govtech; UK Central Digital and Data Office — https://www.gov.uk/government/organisations/central-digital-and-data-office

Critical gap / next action

Create and attach a dated prior-art search log; lock the protocol; execute the proposed study; link raw data and analysis; then replace the results and conclusion placeholders with evidence-bounded findings.

Evidence classification: SEEDED / PARTIAL — defensible non-empirical baseline; validation and replayable search outstanding — provisional research record, not a validated finding.