文献データベース

文献データベースでは、AIセキュリティに関する文献情報を分類・集約しています。詳しくは文献データベースについてをご覧ください。統計情報のページでは、収集された文献に関する統計情報を公開しています。
The Literature Database categorizes and aggregates literature related to AI security. For more details, please see About Literature Database. We provide statistical information regarding the Literature Database on the Statistics page.

Brain-SAD: A Brain-Inspired Safe Autonomous Driving Control Framework with Dynamic Fear-Oriented Constraint on Dual-Policy

Authors: Huan Rong, Chao Yin, Anouar Imel, Yijie Xia, Tinghuai Ma | Published: 2026-09-29
交通シミュレーション
動的制約
意思決定プロセスの分析

Dagger: Decoupling-based Model Stealing Attack against Graph Neural Networks

Authors: Ying Song, Xiaowei Jia, Balaji Palanisamy | Published: 2026-09-29
クラス不均衡問題
モデルインバージョン
モデル抽出攻撃の検知

Boids of a Feather Flock Together – Evolving Prey Behaviours Under Different Predator Attack Strategies

Authors: Augusta van Haren, Hanna Hoogen, Luca Pattavina | Published: 2026-09-29
交通シミュレーション
攻撃モデルの訓練
進化戦略

Where Do LLMs Decide to Break the Rules? Mechanistic Localization of Prompt Injection Compliance

Authors: Rui Wen, Jiayang Liu, Zeyu Yang, Jun Sakuma, Lu Sun | Published: 2026-09-29
インダイレクトプロンプトインジェクション
プロンプトの検証
攻撃効果の評価

Correct, Don’t Delete: Mitigating Emergent Misalignment with Corrective Supervision

Authors: Jacob Epifano | Published: 2026-09-29
データセットの問題
評価手法
透かし手法

Concealing LLM-Based Multi-Agent Topology via Phantom Structure Injection

Authors: Longzhu He, Zelang Wen, Xinfeng Li, Sen Su, XiaoFeng Wang | Published: 2026-09-29
グラフプライバシー
透かし技術
防御メカニズム

Confidence-Guided Protocol IR for LLM-Aided Security Protocol Modeling

Authors: Siqi Li, Yufan Cai, Hongshu Wang, Xinyue Zuo, Zhe Hou, Jin Song Dong | Published: 2026-09-29
セキュリティフレームワーク
モデル設計
透かし技術

Backdoor Mitigation in Decentralized LLM Fine-Tuning

Authors: Sayan Biswas, Jade Garcia Bourrée, Rachid Guerraoui, Maxime Jacovella, Anne-Marie Kermarrec, Sathwika Peechara, Martijn de Vos, Milos Vujasinovic | Published: 2026-09-29
バックドアモデルの検知
プロンプトインジェクション
攻撃効果の評価

A Sharp Transition in Data Reconstruction under Differential Privacy

Authors: Max Cairney-Leeming, Simone Bombari, Marco Mondelli | Published: 2026-09-29
データプライバシー管理
再構成攻撃
透かし手法

Beyond Semantic Narrowing: Robust and Efficient LLM Watermarking with Hamming Neighborhoods

Authors: Zewen Sun, Tongyang Zhao, Liyao Xiang, Mingxuan Ma, Lingzhe Wang, Zhiyuan Li | Published: 2026-09-29
サンプリング手法
透かしの耐久性
透かし手法