Benchmark Scores Are Pipeline-Dependent: A Reliability Audit of Cybersecurity LLM Benchmarks Authors: Aymene Berriche, Cathrine Shalby, Mohannad Alhanahnah, Yazan Boshmaf | Published: 2026-09-08 Prompt InjectionPrompt leaking評価結果 2026.09.08 2026.09.10 Literature Database
Neither Adversarial Training Nor Purification: Emergent Adversarial Robustness from Oscillatory Predictive Learning Authors: Mohammed-Yassine Habibi, Klea Ziu, Martin Takáč, Makoto Yamada | Published: 2026-09-08 Model RobustnessAdversarial Learning自己監視モデル 2026.09.08 2026.09.10 Literature Database
Windows Malware Detector as a Compound AI System: Trade-Offs in Accuracy, Efficiency, and Adversarial Robustness Authors: Andrea Ponte, Luca Demetrio, Luca Oneto, Battista Biggio, Fabio Roli | Published: 2026-09-08 Backdoor DetectionDataset for Malware ClassificationAdversarial Example Detection 2026.09.08 2026.09.10 Literature Database
Do Input-Level Defenses Transfer to Observation-Level Attacks on VideoLLMs? Authors: Bangshuo Zhu, Wei Song, Yuxin Cao, Yuezhong Wu, Zhiquan Liu, Yuekang Li, Jingling Xue | Published: 2026-09-08 Model Robustness攻撃手法の効果評価結果 2026.09.08 2026.09.10 Literature Database
Revoked but Still Authoritative: An Empirical Study of Revocation Enforcement in Agent-Memory Systems Authors: Yi Ting Shen, Kentaroh Toyoda, Alex Leung | Published: 2026-09-08 攻撃手法の効果監査手法評価結果 2026.09.08 2026.09.10 Literature Database
ACEA: An Adversarial Co-Evolution Arena for Head-to-Head Red-Team and Blue-Team LLM Testing Authors: Yi Ting Shen, Kentaroh Toyoda, Alex Leung | Published: 2026-09-08 Attack Scenario Analysis攻撃手法の効果評価結果 2026.09.08 2026.09.10 Literature Database
Style Over Substance: Content-Invariant Wrappers Flip LLM Safety-Judge Verdicts Authors: Yongxi Zhou, Wenbo Ye, Yuanzhe Liu, Zihan Dong, Junwei Yao | Published: 2026-09-08 エラー解析監査手法評価結果 2026.09.08 2026.09.10 Literature Database
Does Deeper Reasoning Compromise Alignment? Revealing and Mitigating of Alignment Collapse in Large Reasoning Models Authors: Yu-Hang Wu, Yu-Jie Xiong, Henghua Zhang, Bairui Zhang, Jia-Chen Zhang, Shaohua Li | Published: 2026-09-08 Model Robustness攻撃手法の効果監査手法 2026.09.08 2026.09.10 Literature Database
LLM-Based Penetration Testing in the Presence of Honeypots Authors: Xinhong Xie, Piyush Nagasubramaniam, Neeraj Karamchandani, Sencun Zhu | Published: 2026-09-08 Honeypot TechnologyPrompt Injection評価結果 2026.09.08 2026.09.10 Literature Database
SENTINEL-RL: Offloading Topological Reasoning from LLM Agents in the Security Operations Center Authors: Uday Vallabhaneni, Cassie L. Cagwin, David J. Wild | Published: 2026-09-03 CybersecurityUser Authentication System監査手法 2026.09.03 2026.09.05 Literature Database