AIセキュリティポータルbot

Red-Teaming Auto Mode: Improving Blocking Classifiers Against Malign Coding Agents

Authors: Alex Remedios, Simon Storf, Fabien Roger, John Hughes | Published: 2026-09-17
エージェント評価手法
攻撃の評価
自動モデル監視

Securing quantum error correction against misleading advice from AI agents

Authors: A. Barış Özgüler | Published: 2026-09-16
エラー解析
評価メトリクス
量子機械学習

ASLEval: Measuring Privacy Exposure Displacement in LLM Agent Sessions

Authors: Guosen Wu, Huizhen Huang, Guoxiong Long, Tao Huang, Chen Hou | Published: 2026-09-16
データ保護
データ駆動型脆弱性評価
評価メトリクス

The Illusion of Local Privacy: Confidentiality Boundary Failures in Consumer LLM Serving Systems

Authors: Youssef Hamdi Zafan Ibrahim, Muhammad Ikram, Mohammed Khalaf Salama | Published: 2026-09-16
プロンプトインジェクション
プロンプトリーキング
評価手法

Beyond Routine Compliance: Cunning Data Cultivates Safety Vigilance in Large Language Models

Authors: Youjia Wang, Lin Xu, Yang Sun, Yuxiao Lu, Chengfang Fang, Jie Shi | Published: 2026-09-16
RAG
データセット評価
安全性調整手法

Collective Loss of Control in LLM Agent Systems: An Epidemic Account of Mutation, Contagion, and Recovery

Authors: Xiangfan Wu, Zonghao Ying, Huiyu Wu, Xing Zheng, Huangsheng Cheng, Xiaorong Shi, Jing Guo | Published: 2026-09-16
制御限界
攻撃の評価
評価メトリクス

Market Signal Injection: Adversarial Context Manipulation of LLM Pricing Agents

Authors: Dohun Lee, Hyunwoo Park | Published: 2026-09-16
インダイレクトプロンプトインジェクション
プロンプトインジェクション
攻撃の評価

A GAN-Based Framework for Robust DDoS Attack Detection

Authors: Makram Chehayeb, Walid Fahs, Amina Rizk, Rida Khatoun, Omran Berjawi | Published: 2026-09-16
サイバーセキュリティ
トリガーの検知
モデル性能評価

BENCHCOMPASS: From Scores to Signals for Training and Harness Decisions in Payment-Domain LLMs

Authors: Sijie Dong, Wei Ren, Xuanwei Hu, Jiawei Luo, Zifan Wang, Xiaoyun Feng, Hui Cai, Lyuxin Xue, Peng Lu, Jianshe Li, Xin Zhang, Wei Wu | Published: 2026-09-16
データ駆動型脆弱性評価
攻撃の評価
評価メトリクス

PentestChain: A Cost-Aware, MCP-Orchestrated Framework for Automated Penetration Testing with Free-Tier LLMs

Authors: Rushabh Vipulkumar Patel, Dipo Dunsin, Mohammed Almaiah, Mohamed Chahine Ghanem | Published: 2026-09-16
コスト意識のあるAI
サイバーセキュリティ
評価基準