Behavioral Reprogramming of Open-Weights Models: Cognitive Plasticity and Alignment Bounds
Large language models (LLMs) are predominantly aligned to function as passive, sycophantic assistants. We challenge this default paradigm by empirically evaluating the cognitive plasticity of open-weight architectures when subjected to rigorous behavioral reprogramming. Our objective is to induce a proactive, Socratic conversational framework, characterized by high-frequency question generation under strictly constrained high-performance computing (HPC) conditions. Through a massively paralleliz
Record details
Published: 13 August 2026
Source: arXiv
Category: Research
Topics: Safety & alignment
Retrieved: 14 August 2026
Related evidence
These records share source-supplied organisations, an exact publisher byline, automatic topics or regions. The reason is shown on every link; related does not mean supporting, agreeing with or verifying this record.
From Local Mismatch to Global Impact: Optimizing Cache Reuse Policy for Efficient Diffusion
arXiv · 13 August 2026
ProME: Prototype-Margin Environments with Repair-Aware Selection for Group-Robust Learning
arXiv cs.LG · 13 August 2026
StateBridge: Training-free Hidden-state Alignment for Latent Communication in LLM Multi-Agent Systems
arXiv · 13 August 2026
Rules or Character? Scaling Laws for AI Safety Design
arXiv · 13 August 2026
HiRoute: Hierarchical Routed Prompt Tuning for Safety Alignment of Large Language Models
arXiv red teaming query · 13 August 2026
Philosophical vertigo with artificial intelligence
arXiv cs.CY · 13 August 2026
How to cite this record
ethics.ai (13 August 2026), “Behavioral Reprogramming of Open-Weights Models: Cognitive Plasticity and Alignment Bounds,” evidence record 19183, https://ethics.ai/record/19183 (originally published by arXiv).
Use and limitations
This page is a stable index and citation surface for a source record. ethics.ai did not author the underlying report and has not independently verified every claim. Automatic topics may be imperfect. For consequential use, quote and cite the original publisher.