Literature Database

Adaptive Adversaries: A Multi-Turn, Multi-LLM Benchmark for LLM Agent Security

Authors: Devina Jain, David Hartmann, Chuan Li | Published: 2026-07-20
システムロバスト性評価
Prompt Injection
攻撃手法の説明

Self-State Attacks on Self-Hosted AI Agents: How Far Can OS Defenses Go?

Authors: Yimeng Chen, Nathanaël Denis, Roberto Di Pietro, Jürgen Schmidhuber | Published: 2026-07-20
OS防御の限界
Indirect Prompt Injection
攻撃手法評価

Zero Hallucination, by Construction: Hallucination-Aware Layered Oversight for Trustworthy Enterprise AI

Authors: Bogdan Raduta, Horia Velicu, Alexandru Preda, Serban Chiricescu | Published: 2026-07-20
Relationship of AI Systems
Model Performance Evaluation
多信号検証

Residual Observability and Attack Detectability in Encrypted OPC UA Traffic

Authors: Song Son Ha, Florian Foerster, Henry Beuster, Eduard Zeller, Dominik Merli, Gerd Scholl | Published: 2026-07-20
Performance Evaluation
Model Performance Evaluation
Attack Detection

Reasoning as a Double-Edged Sword: Architecture and Cross-Stage Robustness in Vision-Language-Action Models

Authors: Tuan Duong Trinh, Naveed Akhtar, Basim Azam | Published: 2026-07-20
Model Performance Evaluation
攻撃戦略分析
Attack Detection

Dynamic Defense Profiling Enables Cognitive Jailbreak of Text-to-Image Models

Authors: Dongdong Yang, Deyue Zhang, Zhao Liu, Zonghao Ying, Wenzhuo Xu, Jiankai Jin, Xiangzheng Zhang, Quanchen Zou | Published: 2026-07-20
Prompt leaking
Model Performance Evaluation
Adversarial Attack Methods

Detection, Attribution, Narration: An End-to-End Pipeline for Explainable Money Mule Identification

Authors: Yuge Zhang, Yuanxing Zhang, Yichao Jin, Khairul Amsyar Mohd Razis, Nicholas Qi An Choo, Kai Yin Anders Wong, Xinyan Tang, Kenneth Zhu Ke, Wee Keong Dennis Lee, Jingyuan Zhao | Published: 2026-07-20
Indirect Prompt Injection
Model Performance Evaluation
Feature Importance Analysis

(A)iSpy: Parasitic Trojans for Machine Learning Infrastructure

Authors: Habibur Rahaman, Qipan Xu, Zafaryab Haider, Prabuddha Chakraborty, Swarup Bhunia, Fnu Suya | Published: 2026-07-20
Trigger Detection
Backdoor Detection
攻撃戦略分析

DecoyFace: Beyond Obfuscation via Controllable and Imperceptible Identity Misdirection for Privacy-Preserving Face Recognition

Authors: Zhihan Ren, Lijun He, Xinyao Wang, Xinzhu Fu, Fan Li | Published: 2026-07-20
Data Privacy Management
Data Protection Method
Privacy Protection in Machine Learning

Pretraining Data Can Be Poisoned through Computational Propaganda

Authors: Victoria Graf, Hannaneh Hajishirzi, Noah A. Smith, David Kohlbrenner, Kyle Lo | Published: 2026-07-16
データセットの問題
データセットの影響
Poisoning