文献データベース

Decision-Level Hijacking: Injecting Cognitive Bias into Large Language Models via Bit-Flip Attacks

Authors: Yu Yan, Jiahao Chen, Siqi Lu, Yongjuan Wang, Ziming Zhao, Zhaoxuan Li, Tianyu Du, Qingjun Yuan, Shouling Ji | Published: 2026-07-28
インダイレクトプロンプトインジェクション
攻撃成功率
記憶攻撃手法

Denial of Deadline: Network-Driven Accuracy Collapse in Distributed Inference Pipelines

Authors: Jhonatan Tavori, Gur-Eyal Sela, Ion Stoica, Gil Zussman | Published: 2026-07-27
データセット評価
攻撃成功率
透かし手法

Agentic Permissions Policy Algebra for Taint Confinement in LLM Agents

Authors: Arseny Kravchenko, Vadim Liventsev, Innokentii Konstantinov, Ildar Iskhakov, Matvey Kukuy | Published: 2026-07-27
プロンプトインジェクション
情報フロー制御
自律エージェントセキュリティ

BettiSplit: Topology-Guided Privacy-Aware Split Learning Against Feature Inversion and Gradient Leakage

Authors: Akarsh K. Nair, Muhammad Arifur Rahman, David Brown, Mufti Mahmud | Published: 2026-07-27
データセットの問題
トポロジーに基づくプライバシー評価
モデル保護手法

When LLM Defenses Backfire: Characterizing Safety, Performance, and Cost Trade-offs

Authors: Tong Zhang, Zexin Li, Simin Chen, Yun Peng | Published: 2026-07-27
LLMの安全機構の解除
プロンプトリーキング
モデル保護手法

DeepFaith: Evidence-Grounded LLMs for Faithful Incident Reporting in Multi-Stage APT Defense

Authors: Trung V. Phan, Tri Gia Nguyen, Thomas Bauschert | Published: 2026-07-27
LLMとの協力効果
プロンプトインジェクション
説明アプローチの評価

Beyond Aggregate Risk: Role-Stratified Conformal Risk Control for LLM Tool Calls

Authors: Md Ashikur Rahman, Md Arifur Rahman, Niamul Hassan Samin, Khandaker Rifah Tasnia, Sifat Rahman Ahona, Juena Ahmed Noshin | Published: 2026-07-27
LLMとの協力効果
モデル保護手法
リスク評価

Gubernaut: A Deterministic Homeostatic Controller for Affect-Regulated LLM Agents, Validated Across Independent Model Families

Authors: Dushyant Sharma | Published: 2026-07-27
AIシステムの関係性
LLMとの協力効果
制御システム

Just Testing, Move Along: Evasion of LLM-based System Log Interpretation by Prompt Injection

Authors: Max Landauer, Florian Skopik, Markus Wurzenberger, Franciszek Górski, Mateusz Krzysztoń | Published: 2026-07-27
インダイレクトプロンプトインジェクション
攻撃手法の効果
説明可能性に対する攻撃

A Cybersecurity MLPS Large Language Model with Multi-Path Retrieval Fusion

Authors: Qian Li, Zhenyan Qi, Liang Shen, Yuan Zhang, Yifan Wan, Junyuan Ma, Yining Hu | Published: 2026-07-27
LLMとの協力効果
RAG
リスク評価