Scoring Rules! Statistical and Strategic Alignment for Text Evaluation Metrics
Reference-based text evaluation metrics, which are widely used to assess natural language generation systems, score a candidate response by comparing it with a reference response. The reliability of an evaluation metric is usually judged by its statistical correlation with human ratings. However, as these metrics are increasingly used as optimization objectives, correlation alone is no longer sufficient: agents may strategically game the evaluation metric. We study this issue through two complem
Record details
Published: 2 August 2026
Source: arXiv cs.LG
Category: Research
Topics: Safety & alignment · Agents & autonomy
Retrieved: 4 August 2026
Related evidence
These records share source-supplied organisations, an exact publisher byline, automatic topics or regions. The reason is shown on every link; related does not mean supporting, agreeing with or verifying this record.
Salami Attack: Stealthy Collusive Memory Poisoning against OpenClaw
arXiv red teaming query · 3 August 2026
Constructing Executable Analytical Knowledge Representations for Meta-Analysis Synthesis Using an Agentic Harness
arXiv · 3 August 2026
MemSIF: From Structured Interactions to Dual-Track Fact Memory for LLM Agents
arXiv · 3 August 2026
From Profiling to Synthesis: Benchmarking Implicit Behavioral Alignment in Personalized LLM Agents
arXiv · 3 August 2026
OpenART: Scaling Agent Red Teaming via Open-Ended Environment Evolution
arXiv red teaming query · 1 August 2026
OpenART: Scaling Agent Red Teaming via Open-Ended Environment Evolution
HuggingFace Daily Papers · 31 July 2026
How to cite this record
ethics.ai (2 August 2026), “Scoring Rules! Statistical and Strategic Alignment for Text Evaluation Metrics,” evidence record 16154, https://ethics.ai/record/16154 (originally published by arXiv cs.LG).
Use and limitations
This page is a stable index and citation surface for a source record. ethics.ai did not author the underlying report and has not independently verified every claim. Automatic topics may be imperfect. For consequential use, quote and cite the original publisher.