Mechanical Enforcement for LLM Governance:Evidence of Governance-Task Decoupling in Financial Decision Systems
Large language models in regulated financial workflows are governed by natural-language policies that the same model interprets, creating a principal--agent failure: outputs can appear compliant without being compliant. Existing evaluation measures task accuracy but not whether governance constrains behaviour at the decision rationale level -- where regulated decisions must be auditable. We introduce five governance metrics that quantify policy compliance at the rationale level and apply them in
Record details
Published: 14 May 2026
Source: arXiv
Category: Research
Topics: Regulation · Agents & autonomy · Transparency
Retrieved: 14 July 2026
Related evidence
These records share source-supplied organisations, an exact publisher byline, automatic topics or regions. The reason is shown on every link; related does not mean supporting, agreeing with or verifying this record.
AI Knows When It's Being Watched: Functional Strategic Action and Contextual Register Modulation in Large Language Models
arXiv · 14 May 2026
Iterative Audit Convergence in LLM-Managed Multi-Agent Systems: A Case Study in Prompt-Engineering Quality Assurance
arXiv · 12 May 2026
Going Headless? On the Boundaries of Vertical AI Firms
arXiv · 18 May 2026
From Code-Centric to Intent-Centric Software Engineering: A Reflexive Thematic Analysis of Generative AI, Agentic Systems, and Engineering Accountability
arXiv · 10 May 2026
REBAR: Reference Ethical Benchmark for Autonomy Readiness
arXiv · 18 May 2026
Governance by Construction for Generalist Agents
arXiv · 20 May 2026
How to cite this record
ethics.ai (14 May 2026), “Mechanical Enforcement for LLM Governance:Evidence of Governance-Task Decoupling in Financial Decision Systems,” evidence record 4329, https://ethics.ai/record/4329 (originally published by arXiv).
Use and limitations
This page is a stable index and citation surface for a source record. ethics.ai did not author the underlying report and has not independently verified every claim. Automatic topics may be imperfect. For consequential use, quote and cite the original publisher.