文献データベース

Multi-Agent AI Control: Distributed Attacks Hamper Per-Instance Monitors

Authors: Oliver Makins, Orazio Angelini, Zohreh Shams, Mary Phuong | Published: 2026-07-08
インダイレクトプロンプトインジェクション
タスク成功率分析
攻撃計画手法

FedCVESA: Taking Away Training Data in Federated Learning via Correlation Value Encoding and Segmented Aggregation

Authors: Chongkai Li, Bang Zhang, Wenjian Luo | Published: 2026-07-08
攻撃計画手法
連合学習
連合学習システム

Measuring Intelligence Beyond Human Scale

Authors: Jerry Han, Rafael Moschopoulos, Ella Colby, Vishrut Goyal, Andrew Tu, Kia Ghods, Mark Braverman, Elad Hazan | Published: 2026-07-08
インセンティブ設計
公的検証可能性
競技チャレンジ分析

Gimitest: A Comprehensive Tool for Testing Reinforcement Learning Policies

Authors: Dennis Gross, Quentin Mazouni, Helge Spieker, Arnaud Gotlieb | Published: 2026-07-08
エージェント操作手法
シミュレーション環境
敵対的サンプルの検知

Thinking More, Harnessing Better: State Machine Guided Harness Automatic Generation with Project Digestion and Workflow Decomposition

Authors: Xing Zhang, Zikang Huang, Gang Yang, CongChong Wang, Lu Liu, Bin Yin, Mingyi Wang, Ziquan Zhao, Min Li, Zhenyu Chen, Bo Wu, Lingyun Ying | Published: 2026-07-08
ソフトウェアセキュリティ
データフロー解析
脆弱性分析

Large Language Models (LLMs) and Generative AI in Cybersecurity and Privacy: A Survey of Dual-Use Risks, AI-Generated Malware, Explainability, and Defensive Strategies

Authors: Kiarash Ahi, Saeed Valizadeh | Published: 2026-07-08
ITセキュリティの課題
インダイレクトプロンプトインジェクション
ソフトウェアセキュリティ

SA-DRL: Security-Aware Deep Reinforcement Learning for Ransomware Detection with Asymmetric Reward Design

Authors: Jannatul Ferdous, Rafiqul Islam, Md Zahidul Islam | Published: 2026-07-08
モデル選択
実験結果分析
統計的検証手法

AirflowAttack: Thermal-Airflow Adversarial Perturbations against Infrared Remote-Sensing Vision-Language Models

Authors: Cong Su, Jiaju Han, Xuemeng Sun, Chengyin Hu, Qike Zhang, Jiujiang Guo, Yiwei Wei, Jiahuan Long | Published: 2026-07-07
データセット分析
敵対的サンプルの脆弱性
脆弱性攻撃手法

Multi-Channel Spread-Spectrum Code Watermarking

Authors: Soohyeon Choi, Debin Gao, Yue Duan | Published: 2026-07-07
ソフトウェアセキュリティ
堅牢性向上手法
脆弱性攻撃手法

Beyond the Syntax: Do Security Experts Trust LLMs for NIDS Rule Engineering?

Authors: Lorenzo di Filippo, Enkeleda Bardhi, Andrea Agiollo, Alessandro Palma, Silvia Bonomi, Fernando Kuipers | Published: 2026-07-07
インダイレクトプロンプトインジェクション
ハルシネーション
ルール帰属