文献データベース

Inherited Circuits, Learned Semantics: How Fine-Tuning Creates Evasion Vulnerabilities Invisible to Standard Evaluation

Authors: Ryan Fetterman | Published: 2026-06-25
プロンプトリーキング
ロバスト分類
検出手法の分析

ShareLock: A Stealthy Multi-Tool Threshold Poisoning Attack Against MCP

Authors: Liwei Liu, Tianzhu Han, Zijian Liu, Zishu Dong, Na Ruan | Published: 2026-06-25
エージェント操作手法
ツール使用分析
プロンプトリーキング

A Deterministic Control Plane for LLM Coding Agents

Authors: Padmaraj Madatha | Published: 2026-06-25
インダイレクトプロンプトインジェクション
エージェント操作手法
データ保護

MIRROR: Novelty-Constrained Memory-Guided MCTS Red-Teaming for Agentic RAG

Authors: Inderjeet Singh, Andrés Murillo, Motoyoshi Sekiya, Yuki Unno, Junichi Suga | Published: 2026-06-25
RAGへのポイズニング攻撃
データセット評価
脆弱性評価手法

DroidBreaker: Practical and Functional Problem-Space Attacks on Machine-Learning Android Malware Detectors

Authors: Christian Scano, Diego Soi, Angelo Sotgiu, Luca Demetrio, Davide Maiorca, Giorgio Giacinto, Fabio Roli, Battista Biggio | Published: 2026-06-25
APK評価手法
ポイズニング
透かし設計

Agents That Know Too Much: A Data-Centric Survey of Privacy in LLM Agents

Authors: Nada Lahjouji, Ashwin Gerard Colaco | Published: 2026-06-25
インダイレクトプロンプトインジェクション
データプライバシー評価
プライバシー保護データマイニング

Empirical Software Engineering TerraProbe: A Layered-Oracle Framework for Detecting Deceptive Fixes in LLM-Assisted Terraform

Authors: Manar Alsaid, Chimdumebi Nebolisa, Faris Abbas | Published: 2026-06-25
バグ修正手法
脅威モデリング
脆弱性評価手法

Adversarial Diffusion Across Modalities: A Fusion Survey of Attacks, Defenses, and Evaluation for Text, Vision, and Vision-Language Models

Authors: Abrar Alotaibi, Moataz Ahmed | Published: 2026-06-25
LLMの安全機構の解除
データ生成手法
モデルDoS

Adaptive Evaluation of Out-of-Band Defenses Against Prompt Injection in LLM Agents

Authors: Praneeth Narisetty, Shiva Nagendra Babu Kore, Uday Kumar Reddy Kattamanchi, Jayaram Kumarapu | Published: 2026-06-25
インダイレクトプロンプトインジェクション
エージェント操作手法
透かし攻撃

Detect, Unlearn, Restore: Defending Text Summarization Models Against Data Poisoning

Authors: Poojitha Thota, Shirin Nilizadeh | Published: 2026-06-24
データセットの影響
データ毒性攻撃
ポイズニング