PolicyLLM: Towards Excellent Comprehension of Public Policy for Large Language Models
Large Language Models (LLMs) are increasingly integrated into real-world decision-making, including in the domain of public policy. Yet, their ability to comprehend and reason about policy-related content remains underexplored. To fill this gap, we present \textbf{\textit{PolicyBench}}, the first large-scale cross-system benchmark (US-China) evaluating policy comprehension, comprising 21K cases across a broad spectrum of policy areas, capturing the diversity and complexity of real-world governan
Record details
Published: 14 April 2026
Source: arXiv
Category: Research
Topics: Regulation
Retrieved: 14 July 2026
Related evidence
These records share source-supplied organisations, an exact publisher byline, automatic topics or regions. The reason is shown on every link; related does not mean supporting, agreeing with or verifying this record.
Enabling and Inhibitory Pathways of Students' AI Use Concealment Intention in Higher Education: Evidence from SEM and fsQCA
arXiv · 13 April 2026
Polarization and Integration in Global AI Research
arXiv · 19 April 2026
Understanding the Role of Algorithm Registers in AI Governance Through Comparative Analysis of China and the UK
arXiv · 25 April 2026
IndustryBench: Probing the Industrial Knowledge Boundaries of LLMs
arXiv · 11 May 2026
Detection Is Cheap, Routing Is Learned: Why Refusal-Based Alignment Evaluation Fails
arXiv · 18 March 2026
AI Sovereignty as National Learning Capacity: A Human-Centered Learning Mechanics Viewpoint on France, the United States, and China
arXiv · 30 May 2026
How to cite this record
ethics.ai (14 April 2026), “PolicyLLM: Towards Excellent Comprehension of Public Policy for Large Language Models,” evidence record 5861, https://ethics.ai/record/5861 (originally published by arXiv).
Use and limitations
This page is a stable index and citation surface for a source record. ethics.ai did not author the underlying report and has not independently verified every claim. Automatic topics may be imperfect. For consequential use, quote and cite the original publisher.