Governing Execution Risk in Agentic AI Systems: A Trajectory-Guided Framework for Red Teaming
arXiv:2608.04018v1 Announce Type: new Abstract: AI agents are increasingly embedded in organizational workflows, where they interact with external information sources and invoke digital tools to perform operational tasks. As organizations adopt such systems, a critical challenge is identifying and mitigating risks arising from malicious or untrusted external information that can steer agents toward unintended actions. Existing red-teaming approaches largely rely on fixed attack templates or fina
Record details
Published: 6 August 2026
Source: arXiv cs.CY
Category: Research
Topics: Safety & alignment · Agents & autonomy
Retrieved: 6 August 2026
Related evidence
These records share source-supplied organisations, an exact publisher byline, automatic topics or regions. The reason is shown on every link; related does not mean supporting, agreeing with or verifying this record.
Agent Against Agent: An Agentic System for Automatic Prompt Injection Red Teaming
arXiv red teaming query · 5 August 2026
Multi-Agent Forensic Reasoning for Generalizable Deepfake Video Detection
HuggingFace Daily Papers · 6 August 2026
IntentLint: Supporting Intent Scaffolding and Prompt-time Linting in Human-AI Collaborative Data Analysis
arXiv cs.HC · 5 August 2026
Multi-Agent Forensic Reasoning for Generalizable Deepfake Video Detection
arXiv · 7 August 2026
MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents
arXiv · 7 August 2026
When Privileged Guidance Misaligns: State-Matched Routing and Contextualized Self-Distillation for Multi-Turn Agents
HuggingFace Daily Papers · 4 August 2026
How to cite this record
ethics.ai (6 August 2026), “Governing Execution Risk in Agentic AI Systems: A Trajectory-Guided Framework for Red Teaming,” evidence record 16635, https://ethics.ai/record/16635 (originally published by arXiv cs.CY).
Use and limitations
This page is a stable index and citation surface for a source record. ethics.ai did not author the underlying report and has not independently verified every claim. Automatic topics may be imperfect. For consequential use, quote and cite the original publisher.