Public Sector & Government · Issue 08 ·
Public Safety
Three historical sources new to the archive cover police adoption perceptions, New York court pilots and Catalan corrections explainability. Self-reported benefits, operator accounts and academic governance proposals remain distinct from measured effectiveness. One cross-source pattern addresses pilot evaluation. No fresh operational outcome since the last successful run was verified; U.S. institutional-prison outcome evidence remains a gap. Role guidance is interpretation.
- Evidence records
- 3
- Cross-source patterns
- 1
- Evidence classes
- 2 academic research1 government evaluation
- Outcomes
- 1 mixed1 emerging1 cautionary
- Source freshness
- 2 older, newly relevant1 undated
- Research completed
- 2026-09-14
Choose a role to see its takeaway beside every record in the ledger.
Synthesis · Lighthouse Advisory interpretation
Patterns across the evidence
Positive pilot feedback leaves the operational value question open
The police survey and New York court report offer useful adoption signals but do not establish net operational benefit. Evaluate user experience alongside an independently defined task outcome and total review effort. These are different institutions and study designs, so they do not support a pooled effect estimate.
Operating questionWhat evidence beyond positive feedback will determine whether the specific workflow should expand?
Supporting evidenceHunter Boehme and colleaguesNew York State Unified Court System Advisory Committee on AI and the Courts
Full record · every source keeps its link and limitations
Evidence ledger
Post-trial survey separates officer enthusiasm from verified reporting benefits
Favorable attitudes did not differ significantly between treated and control officers; reported workflow problems complicate adoption.
Why it matters, evidence and limitations
- Why it matters
- Historical workforce evidence complements the archived Manchester timing trial; this is a separate survey, not independent replication.
- Evidence and measured results
- Final analytical sample: 96 respondents. Table 2 records 19 of 40 time-question respondents perceiving savings and 19 perceiving no difference. These are self-reports, not measured labor reductions. Methods use group comparisons and descriptive analysis.
- Limitations and uncertainty
- Single agency, immediate post-intervention perceptions, uneven item denominators, limited demographic diversity. No objective quality or downstream court outcome assessment. Preprint version inspected.
New York court report distinguishes pilot feedback, task suitability and security approval
The committee reports positive pilot feedback while distinguishing technical approval from suitability for particular court tasks.
Why it matters, evidence and limitations
- Why it matters
- Historical state-court deployment evidence fills a gap between general judicial interviews and actual institutional approval workflows.
- Evidence and measured results
- The report describes a two-month Copilot Chat pilot with more than 200 participants, nearing completion at report time. Feedback is qualitative; no controlled baseline, error denominator or measured net time saving is supplied.
- Limitations and uncertainty
- Selected pilot, policy and hosting sections inspected, not all 154 pages. Present rollout and vendor security status are unverified. Broad privacy assurances are operator assertions, not independent security findings.
RisCanvi analysis examines who can understand and challenge corrections risk scores
The authors argue that restricted access to score explanations weakens meaningful challenge and oversight.
Why it matters, evidence and limitations
- Why it matters
- Historical international evidence adds a prison risk-assessment governance case; Spanish procedures and European legal claims do not establish U.S. obligations.
- Evidence and measured results
- Integrative analysis of prior audits, reporting and scholarship, with a proposed governance framework. It is not a new predictive-performance experiment or a deployed framework evaluation.
- Limitations and uncertainty
- Underlying audit datasets were not independently inspected. No new participant sample, comparator or measured benefit of the proposal is established. Do not infer that logistic regression is inherently unexplainable.
How to read this edition
Source findings, measured results and limitations come from the cited publications. Patterns, operating questions, role takeaways and implementation considerations are Lighthouse Advisory interpretation, stated as questions to validate locally rather than guaranteed outcomes. Vendor and operator claims are labeled as claims. Full research method.
- Academic research
- Research produced through an academic institution or peer-reviewed venue.
- Government evaluation
- A public body’s measured evaluation or documented pilot.