Skip to main content

Research timeline

Related research and updates

Public articles linked to the same research event.

arXiv

TGCC uses cross-layer trust gating so a compromised LLM agent self-revokes within steps instead of escalating

The work turns a descriptive layered trust model into an operational controller: a cross-layer synergy operator propagates prerequisite-layer deficits into dependent layers, a no-regret online estimator grounds per-layer weights in observed failures, and Trust-Gated Capability Control issues short-lived, revocable grants only when composite trust and the relevant prerequisite layers clear capability-specific thresholds; the authors prove and numerically confirm that a stealthy compromise inflating behavioral trust while degrading a prerequisite layer cannot escalate privilege and instead self-revokes within a few interactions.