Literature Database

Online Shift Detection and Conformal Adaptation for Deployed Safety Classifiers

Authors: Jun Wen Leong | Published: 2026-06-10
System Observability
異常検知
Statistical Methods

Grammar-Constrained Decoding Can Jailbreak LLMs into Generating Malicious Code

Authors: Yitong Zhang, Shiteng Lu, Jia Li | Published: 2026-06-10
Prompt Injection
Large Language Model
安全性の整合性

Can Open-Source LLM Agents Replace Static Application Security Testing Tools? An Empirical Assessment

Authors: Derek Yohn, Luke Flancher, Mirajul Islam, Khaled Slhoub | Published: 2026-06-10
Indirect Prompt Injection
Data-Driven Vulnerability Assessment
静的アプリケーションセキュリティテスト

Dummy Backdoor as a Defense: Removing Unknown Backdoors via Shared Internal Mechanisms for Generative LLMs

Authors: Kazuki Iwahana, Masaru Matsubayashi, Takuma Koyama, Toshiki Shibahara, Kenichiro Omintato, Akira Ito | Published: 2026-06-10
Detection of Poison Data for Backdoor Attacks
Prompt leaking
Robustness Improvement Method

Defense Against Prompt Inversion Attacks: An Information-Theoretic Approach for LLM Collaborative Inference

Authors: Sayedeh Leila Noorbakhsh, Hossein Khalili, Nader Sehatbakhsh | Published: 2026-06-10
Indirect Prompt Injection
Privacy Enhancing Technology
Prompt validation

Hiding the Trees in the Forest: Building Network Covert Channels with Hash-Based Covert Carrier Filtering

Authors: Zexiao Zou, Zhiqiang Wang, Baoxu Liu, Yuyang Han, Yan Zhang | Published: 2026-06-10
Data-Centric Security
Robustness Improvement
Watermarking Technology

OpenPCC: Open and Confidential LLM Serving on Commodity TEEs

Authors: Haoling Zhou, Shixuan Zhao, Chao Wang, Zhiqiang Lin | Published: 2026-06-09
Data-Centric Security
Performance Evaluation
Prompt leaking

Context-Based Adversarial Attacks on AI Code Generators: Vulnerability Analysis and Implications

Authors: Walther A. Del Orbe, John D. Hastings, Varghese Vaidyan | Published: 2026-06-09
Data-Driven Vulnerability Assessment
Prompt leaking
Model Inversion

Comparative Analysis of Inference-Time Defense Methods for Multimodal Large Language Models

Authors: Bulat Nutfullin, Vladimir Evgrafov, Dmitry Namiot | Published: 2026-06-09
Data-Driven Vulnerability Assessment
Model Extraction Attack
Defensive Deception

Toward Secure LLM Agents: Threat Surfaces, Attacks, Defenses, and Evaluation

Authors: Yuchen Ling, Shengcheng Yu, Zhenyu Chen, Chunrong Fang | Published: 2026-06-09
Indirect Prompt Injection
Data-Centric Security
Backdoor Detection