文献データベース

Detecting Functional Memorization in Code Language Models

Authors: Matthieu Meeus, Anil Ramakrishna, Matthew Grange, Zheng Xu, Luca Melis | Published: 2026-06-11
LLMの応用
データ収集
データ生成

PI-Hunter: Automated Red-Teaming for Exposing and Localizing Prompt Injections

Authors: Pengfei He, Lesly Miculicich, Vishesh Sharma, Ash Fox, George Lee, Jiliang Tang, Tomas Pfister, Long T. Le | Published: 2026-06-10
インダイレクトプロンプトインジェクション
データ駆動型脆弱性評価
自律エージェントセキュリティ

OCELOT: Inference-Leakage Budgets for Privacy-Preserving LLM Agents

Authors: Jin Xie, Songze Li | Published: 2026-06-10
プライバシー保護技術
プロンプトインジェクション

Mind your key: An Empirical Study of LLM API Credential Leakage in iOS Apps

Authors: Pinran Gao, Lingxiang Wang, Ying Zhang, Fan Yang | Published: 2026-06-10
データ流出に関する分析手法
プロンプトリーキング

Categorical Robustness Assessment for Machine Learning based Network Intrusion Detection Systems

Authors: Mayank Raj, Nathaniel D. Bastian, Lance Fiondella, Gokhan Kul | Published: 2026-06-10
モデルの頑健性保証
ロバスト性向上
敵対的学習

Online Shift Detection and Conformal Adaptation for Deployed Safety Classifiers

Authors: Jun Wen Leong | Published: 2026-06-10
システム観測性
異常検知
統計的手法

Grammar-Constrained Decoding Can Jailbreak LLMs into Generating Malicious Code

Authors: Yitong Zhang, Shiteng Lu, Jia Li | Published: 2026-06-10
プロンプトインジェクション
大規模言語モデル
安全性の整合性

Can Open-Source LLM Agents Replace Static Application Security Testing Tools? An Empirical Assessment

Authors: Derek Yohn, Luke Flancher, Mirajul Islam, Khaled Slhoub | Published: 2026-06-10
インダイレクトプロンプトインジェクション
データ駆動型脆弱性評価
静的アプリケーションセキュリティテスト

Dummy Backdoor as a Defense: Removing Unknown Backdoors via Shared Internal Mechanisms for Generative LLMs

Authors: Kazuki Iwahana, Masaru Matsubayashi, Takuma Koyama, Toshiki Shibahara, Kenichiro Omintato, Akira Ito | Published: 2026-06-10
バックドア攻撃用の毒データの検知
プロンプトリーキング
ロバスト性向上手法

Defense Against Prompt Inversion Attacks: An Information-Theoretic Approach for LLM Collaborative Inference

Authors: Sayedeh Leila Noorbakhsh, Hossein Khalili, Nader Sehatbakhsh | Published: 2026-06-10
インダイレクトプロンプトインジェクション
プライバシー保護技術
プロンプトの検証