Multi-Adapter Representation Interventions via Energy Calibration
Representation intervention has emerged as a promising paradigm for aligning large language models toward desired behaviors without modifying model weights. Existing methods typically apply a fixed intervention uniformly across all inputs. However, we find that the appropriate intervention direction and strength vary substantially across samples, and such indiscriminate intervention leads to degradation of general capabilities on benign inputs. To address these challenges, we propose Multi-Adapt
Record details
Published: 27 May 2026
Source: arXiv
Category: Research
Topics: Environment
Retrieved: 14 July 2026
Related evidence
These records share source-supplied organisations, an exact publisher byline, automatic topics or regions. The reason is shown on every link; related does not mean supporting, agreeing with or verifying this record.
Rethinking Memory as Continuously Evolving Connectivity
arXiv · 27 May 2026
Calibrating Conservatism for Scalable Oversight
arXiv · 27 May 2026
Anomaly as Non-Conformity via Training-Free Graph Laplacian Energy Minimization
arXiv · 27 May 2026
PIRS: Physics-Informed Reward Shaping for SAC-Based Building Energy Management
arXiv · 27 May 2026
OccuReward: LLM-Guided Occupant-Centric Reward Shaping for Demographic Equity in Grid-Interactive Buildings
arXiv · 27 May 2026
Does Distributed Training Undermine Compute Governance?
arXiv · 28 May 2026
How to cite this record
ethics.ai (27 May 2026), “Multi-Adapter Representation Interventions via Energy Calibration,” evidence record 3572, https://ethics.ai/record/3572 (originally published by arXiv).
Use and limitations
This page is a stable index and citation surface for a source record. ethics.ai did not author the underlying report and has not independently verified every claim. Automatic topics may be imperfect. For consequential use, quote and cite the original publisher.