AIセキュリティポータルbot

Stealing Reasoning Traces from Proprietary LLM APIs

Authors: Alexander Panfilov, David Schmotz, Ilia Shumailov, Luca Beurer-Kellner, Joachim Schaeffer, Ameya Prabhu, Jonas Geiping, Maksym Andriushchenko | Published: 2026-08-10
Prompt leaking
Model Interpretability
三角形の幾何学

Governing the KV Cache: Preventing Timing Side-Channel Leakage in Multi-Tenant LLM Inference

Authors: Tejasvi C. Addagada | Published: 2026-08-10
攻撃計画手法
Watermark Evaluation
Defense Method

Hardware Keystores for AI Agent Signing Workflows: A Zero-Trust MCP Enforcement Architecture

Authors: Leo Sambrook, Sampo Sovio | Published: 2026-08-06
Indirect Prompt Injection
Verification of Digital Signatures
Defense Method

HERALD: Counterfactual Audits and Minimal Repairs for Proof-of-Retrieval Rewards

Authors: Zhuowen Liu, Bohan Cui, YinShang Guo, Yuting Wang, Hao Li | Published: 2026-08-06
Attack Scenario Analysis
Research Methodology
Watermark Evaluation

MMAligner: Safeguarding Multimodal Large Language Models through Representation Calibration

Authors: Shenyi Zhang, Keyan Guo, Zihao Wang, Xuebin Li, Lingchen Zhao, Hongxin Hu, Chao Shen, Qian Wang | Published: 2026-08-06
Prompt Injection
Model Interpretability
Large Language Model

Tool Demo: Topology analysis with GPML for detection of cyberattacks in Water Distribution Networks

Authors: Majed Jaber, Abdul Qadir Khan, Ankush Meshram, Julien Michel, Côme Frappé - - Vialatoux, Pierre Parrend | Published: 2026-08-06
Dynamic Analysis
Attack Scenario Analysis
Industrial Control System

ChainClaw: A Layered Agent Framework for Reliable On-Chain Execution

Authors: Jiacheng Wei, Zhaoxin Fan, Xin Wen, Yuqin Lan, Dongrun Li, Wenjun Wu, Faguo Wu, Xiao Zhang | Published: 2026-08-06
Simulation Result Evaluation
Framework Support
透かし手法

GROM: Gradient-Free Rapid One-Shot Machine Unlearning

Authors: Paweł Batorski, Przemysław Spurek, Paul Swoboda | Published: 2026-08-06
Factors of Performance Degradation
透かし手法
Watermark Evaluation

Hijacking Robots with a Piece of Paper: A Systematic Study of Physical Prompt Injection in VLM-Controlled Robots

Authors: S. M . Bhagya P. Samarakoon, M. A. Viraj J. Muthugala, W. K. R. Sachinthana, Mohan Rajesh Elara | Published: 2026-08-06
Indirect Prompt Injection
Prompt leaking
Vehicle Hijacking Attack

DreamGuard: Efficient Runtime Guardrail for LLM Agents via Risk-Aware World Model

Authors: Wenhao Lin, Chenyu Yu, Xingwei Lin, Sicong Cao, Xiang Chen, Lei Xue, Le Yu, Letian Sha, Chunming Wu | Published: 2026-08-06
Indirect Prompt Injection
リスクとメカニズム
Watermark Evaluation