Literature Database

LLMs Prompted for Legal Context Object More: Overrefusal from Small On-Premises LLMs in Criminal Legal Context

Authors: Anastasiia Kucherenko, François Brouchoud, Dimitri Percia David, Andrei Kucharavy | Published: 2026-06-23
Relationship of AI Systems
Prompt Injection
文献レビュー

Poster: Exploring the Limits of Audio-Based Detection of Turkish Phone Call Scams

Authors: Arda Eren, Micheal Cheung, Youqian Zhang, Grace Ngai, Eugene Yujun Fu | Published: 2026-06-23
データセットの影響
マルチモーダル安全性
Resource Scarcity Issues

Red-Teaming the Agentic Red-Team

Authors: Dario Pasquini, Michal Bazyli, Taras Fedynyshyn, Artem Sorokin | Published: 2026-06-23
Indirect Prompt Injection
エージェント操作手法
脅威モデリング自動化

PHANTOM: A Large-Scale Dataset of Multimodal Adversarial Attacks for Vision-Language Models

Authors: Simone Gallivanone, Hossein Khodadadi, Mauro Dore, Mauro Medda, Nicola Franco | Published: 2026-06-23
Dataset Applicability
Robustness
Large Language Model

AutoSpec: Safety Rule Evolution for LLM Agents via Inductive Logic Programming

Authors: Pingchuan Ma, Zhaoyu Wang, Zimo Ji, Yuguang Zhou, Zhantong Xue, Zongjie Li, Shuai Wang, Xiaoqin Zhang | Published: 2026-06-23
ILPとCEGISの統合
LLM Application
Automated Vulnerability Remediation

PixJail: Self-Evolving Paper-to-Pipeline Reproduction for Text-to-Image Jailbreak Evaluation

Authors: Leyi Sheng, Han Sun, Zhen Sun, Yuntao Yue, Jinlin Wu, Xinlei He, Jiaheng Wei | Published: 2026-06-23
システムロバスト性評価
脱獄攻撃手法
自動生成フレームワーク

An Automated Framework for Input Alphabet Construction in Stateful Protocol Implementation Learning

Authors: JiongHan Wang, WenChao Huang | Published: 2026-06-22
Mutation Testing
Vulnerability Prediction
自動評価手法

Detecting Malicious Agent Skills in the Wild using Attention

Authors: Bacem Etteib, Daniele Lunghi, Tégawendé F. Bissyandé | Published: 2026-06-22
LLM Application
Indirect Prompt Injection
データ流出に関する分析手法

FlexServe: A Fast and Secure LLM Serving System for Mobile Devices with Flexible Resource Isolation

Authors: Yinpeng Wu, Yitong Chen, Lixiang Wang, Jinyu Gu, Zhichao Hua, Yubin Xia | Published: 2026-06-22
LLM Application
セキュアリソース管理
メモリ効率化手法

Rethinking Molecular Graph Backdoors under Chemistry-aware Admission

Authors: Thinh T. H. Nguyen, Sze Jue Yang, Khoa D. Doan, Chee Seng Chan, Kok-Seng Wong | Published: 2026-06-22
データ毒性
Backdoor Detection
Backdoor Attack