CivicShield: A Cross-Domain Defense-in-Depth Framework for Securing Government-Facing AI Chatbots Against Multi-Turn Adversarial Attacks
LLM-based chatbots in government services face critical security gaps. Multi-turn adversarial attacks achieve over 90% success against current defenses, and single-layer guardrails are bypassed with similar rates. We present CivicShield, a cross-domain defense-in-depth framework for government-facing AI chatbots. Drawing on network security, formal verification, biological immune systems, aviation safety, and zero-trust cryptography, CivicShield introduces seven defense layers: (1) zero-trust fo
Record details
Published: 30 March 2026
Source: arXiv
Category: Research
Topics: Military & security · Biotech
Retrieved: 14 July 2026
Related evidence
These records share source-supplied organisations, an exact publisher byline, automatic topics or regions. The reason is shown on every link; related does not mean supporting, agreeing with or verifying this record.
RAGShield: Detecting Numerical Claim Manipulation in Government RAG Systems
arXiv · 1 April 2026
SentinelAgent: Intent-Verified Delegation Chains for Securing Federal Multi-Agent AI Systems
arXiv · 3 April 2026
Harmonizing AI Safety Thresholds
arXiv · 17 July 2026
Hundreds asked ChatGPT for poison and bioweapon recipes and some got step-by-step high school level guides
The Decoder · 26 July 2026
AI can fuel biological weapons. We must harness its power for defense | Annie Jacobsen
The Guardian · 27 July 2026
Possible Is Not Plausible: How Nightmare Scenarios Hijack Unconventional Weapons Policy
War on the Rocks · 3 August 2026
How to cite this record
ethics.ai (30 March 2026), “CivicShield: A Cross-Domain Defense-in-Depth Framework for Securing Government-Facing AI Chatbots Against Multi-Turn Adversarial Attacks,” evidence record 6573, https://ethics.ai/record/6573 (originally published by arXiv).
Use and limitations
This page is a stable index and citation surface for a source record. ethics.ai did not author the underlying report and has not independently verified every claim. Automatic topics may be imperfect. For consequential use, quote and cite the original publisher.