Measuring & Mitigating Over-Alignment for LLMs in Multilingual Criminal Law Courts
While the wider applicability of LLMs in the legal field is currently debated due to their reliability and the gravity of any errors, narrow uses with well-understood and mitigated risks have emerged. Notably the Swiss Federal Supreme Court uses small on-premises models for tentative translations and short-passage summarization across the four official languages. However, such usage is challenging in the context of Criminal Law. Since rulings and cases employees work on routinely can contain det
Record details
Published: 22 June 2026
Source: arXiv
Category: Research
Topics: Regulation · Safety & alignment
Retrieved: 14 July 2026
Related evidence
These records share source-supplied organisations, an exact publisher byline, automatic topics or regions. The reason is shown on every link; related does not mean supporting, agreeing with or verifying this record.
The Geometry of Refusal: Linear Instability in Safety-Aligned LLMs
arXiv · 21 June 2026
What Intermediate Layers Know: Detecting Jailbreaks from Entropy Dynamics
arXiv · 23 June 2026
Learning Action Priors for Cross-embodiment Robot Manipulation
arXiv · 24 June 2026
What Do Safety-Aligned LLMs Learn From Mixed Compliance Demonstrations?
arXiv · 18 June 2026
FinRED: An Expert-Guided Benchmark Generation and Evaluation Framework for Financial LLM Red-Teaming
arXiv · 18 June 2026
Challenges to Grassroots Organization Engagement with AI Policy
arXiv · 18 June 2026
How to cite this record
ethics.ai (22 June 2026), “Measuring & Mitigating Over-Alignment for LLMs in Multilingual Criminal Law Courts,” evidence record 698, https://ethics.ai/record/698 (originally published by arXiv).
Use and limitations
This page is a stable index and citation surface for a source record. ethics.ai did not author the underlying report and has not independently verified every claim. Automatic topics may be imperfect. For consequential use, quote and cite the original publisher.