文献データベース

Towards Agentic Investigation of Security Alerts

Authors: Even Eilertsen, Vasileios Mavroeidis, Gudmund Grov | Published: 2026-04-28
LLM性能評価
RAGへのポイズニング攻撃
インダイレクトプロンプトインジェクション

From CRUD to Autonomous Agents: Formal Validation and Zero-Trust Security for Semantic Gateways in AI-Native Enterprise Systems

Authors: Ignacio Peyrano | Published: 2026-04-28
セキュアなロジスティック回帰
動的アクセス制御
脆弱性評価手法

MARD: A Multi-Agent Framework for Robust Android Malware Detection

Authors: Xueying Zeng, Youquan Xian, Sihao Liu, Xudong Mou, Yanze Li, Lei Cui, Bo Li | Published: 2026-04-28
LLM性能評価
インダイレクトプロンプトインジェクション
一般化性能

R-CoT: A Reasoning-Layer Watermark via Redundant Chain-of-Thought in Large Language Models

Authors: Ziming Zhang, Li Li, Guorui Feng, Hanzhou Wu, Xinpeng Zhang | Published: 2026-04-28
プロンプトインジェクション
報酬関数設計
検証可能な資格情報

Making AI-Assisted Grant Evaluation Auditable without Exposing the Model

Authors: Kemal Bicakci | Published: 2026-04-28
リスクシナリオ生成
検証可能な資格情報
評価手法

AgentWard: A Lifecycle Security Architecture for Autonomous AI Agents

Authors: Yixiang Zhang, Xinhao Deng, Jiaqing Wu, Yue Xiao, Ke Xu, Qi Li | Published: 2026-04-27
インダイレクトプロンプトインジェクション
リスクシナリオ生成
攻撃チェーン分析

Layerwise Convergence Fingerprints for Runtime Misbehavior Detection in Large Language Models

Authors: Nay Myat Min, Long H. Pham, Jun Sun | Published: 2026-04-27
インダイレクトプロンプトインジェクション
プロンプトインジェクション
一般化性能

GAMMAF: A Common Framework for Graph-Based Anomaly Monitoring Benchmarking in LLM Multi-Agent Systems

Authors: Pablo Mateo-Torrejón, Alfonso Sánchez-Macián | Published: 2026-04-27
LLM性能評価
インダイレクトプロンプトインジェクション
マルチエージェントシステム

A Survey on Split Learning for LLM Fine-Tuning: Models, Systems, and Privacy Optimizations

Authors: Zihan Liu, Yizhen Wang, Rui Wang, Xiu Tang, Sai Wu | Published: 2026-04-27
AIによる出力のバイアスの検出
プライバシー保護手法
連合学習

Defusing the Trigger: Plug-and-Play Defense for Backdoored LLMs via Tail-Risk Intrinsic Geometric Smoothing

Authors: Kaisheng Fan, Weizhe Zhang, Yishu Gao, Tegawendé F. Bissyandé, Xunzhu Tang | Published: 2026-04-27
バックドアモデルの検知
モデル抽出攻撃
攻撃チェーン分析