Literature Database

Hybrid Analysis for Secure MCP Tool Use in LLM Agents

Authors: Ping He, Yuexiang Xie, Yaliang Li, Shouling Ji | Published: 2026-07-28
Disabling Safety Mechanisms of LLM
Risk Assessment Method
自律エージェントセキュリティ

Decision-Level Hijacking: Injecting Cognitive Bias into Large Language Models via Bit-Flip Attacks

Authors: Yu Yan, Jiahao Chen, Siqi Lu, Yongjuan Wang, Ziming Zhao, Zhaoxuan Li, Tianyu Du, Qingjun Yuan, Shouling Ji | Published: 2026-07-28
Indirect Prompt Injection
攻撃成功率
記憶攻撃手法

Denial of Deadline: Network-Driven Accuracy Collapse in Distributed Inference Pipelines

Authors: Jhonatan Tavori, Gur-Eyal Sela, Ion Stoica, Gil Zussman | Published: 2026-07-27
Dataset evaluation
攻撃成功率
透かし手法

Agentic Permissions Policy Algebra for Taint Confinement in LLM Agents

Authors: Arseny Kravchenko, Vadim Liventsev, Innokentii Konstantinov, Ildar Iskhakov, Matvey Kukuy | Published: 2026-07-27
Prompt Injection
Information Flow Control
自律エージェントセキュリティ

BettiSplit: Topology-Guided Privacy-Aware Split Learning Against Feature Inversion and Gradient Leakage

Authors: Akarsh K. Nair, Muhammad Arifur Rahman, David Brown, Mufti Mahmud | Published: 2026-07-27
データセットの問題
トポロジーに基づくプライバシー評価
Model Protection Methods

When LLM Defenses Backfire: Characterizing Safety, Performance, and Cost Trade-offs

Authors: Tong Zhang, Zexin Li, Simin Chen, Yun Peng | Published: 2026-07-27
Disabling Safety Mechanisms of LLM
Prompt leaking
Model Protection Methods

DeepFaith: Evidence-Grounded LLMs for Faithful Incident Reporting in Multi-Stage APT Defense

Authors: Trung V. Phan, Tri Gia Nguyen, Thomas Bauschert | Published: 2026-07-27
Cooperative Effects with LLM
Prompt Injection
Evaluation of Explanatory Approaches

Beyond Aggregate Risk: Role-Stratified Conformal Risk Control for LLM Tool Calls

Authors: Md Ashikur Rahman, Md Arifur Rahman, Niamul Hassan Samin, Khandaker Rifah Tasnia, Sifat Rahman Ahona, Juena Ahmed Noshin | Published: 2026-07-27
Cooperative Effects with LLM
Model Protection Methods
Risk Assessment

Gubernaut: A Deterministic Homeostatic Controller for Affect-Regulated LLM Agents, Validated Across Independent Model Families

Authors: Dushyant Sharma | Published: 2026-07-27
Relationship of AI Systems
Cooperative Effects with LLM
制御システム

Just Testing, Move Along: Evasion of LLM-based System Log Interpretation by Prompt Injection

Authors: Max Landauer, Florian Skopik, Markus Wurzenberger, Franciszek Górski, Mateusz Krzysztoń | Published: 2026-07-27
Indirect Prompt Injection
攻撃手法の効果
Attacks on Explainability