AIセキュリティポータルbot

PolicyGuide: From Guarding One Action to Guiding the Whole Workflow for Policy-Compliant LLM Agents

Authors: Seongjae Kang, Taehyung Yu, Sung Ju Hwang | Published: 2026-08-20
タスク成功率分析
Formal Verification
Progress Tracking

TempJail: Temporal Jailbreak Attack against Large Vision-Language Models via Subtitle Scheduling

Authors: Ling Zhou, Yihao Huang, Jingling Sun, Zhiwen Tian, Yi Zeng, Qihe Liu, Shijie Zhou | Published: 2026-08-20
Prompt Injection
攻撃手法評価
Deep Learning Model

A Locally Tokenized Generative Model for Robust Time-Series Watermarking

Authors: Dongbin Kim, Geonwoo Shin, Yujin Choi, Soyeon Park, Jaewook Lee | Published: 2026-08-20
Reliability Assessment
Robustness of Deep Networks
Deep Learning Model

SiNMULI: Novel Signed Network Approach for Malicious URL Identification

Authors: Avijit Gayen, Sayan Mondal, Angshuman Jana | Published: 2026-08-19
Dataset Analysis
悪意のあるURL検出
Computational Complexity

Beyond the Transcript: Detecting Covert Co ordination in Latent Multi-Agent Communication

Authors: Ramneet Kaur, Pradyumna Chari, Ramesh Raskar, Jugad Singh, Sumit Kumar Jha, Anirban Roy | Published: 2026-08-19
マルチエージェントシステム
攻撃手法評価
Anomaly Detection Algorithm

Detecting Backdoors in Object Detection via Pre-NMS Prediction Distribution Shift

Authors: Longtian Wang, Zhengyu Zhao, Chenhao Lin, Le Yang, Shiwei Wang, Yuhan Zhi, Xiaofei Xie, Chao Shen | Published: 2026-08-19
Backdoor Attack
Backdoor Attack Mitigation
Deep Learning Model

From Threat Intelligence to Detection: Knowledge-driven Enrichment and Template-based Rule Grounding for Automated Sigma Rule Generation

Authors: Sepehr Ghaffarzadegan, Boubakr Nour, Makan Pourzandi, Mourad Debbabi, Chadi Assi | Published: 2026-08-19
Bias Detection in AI Output
Cyber Threat
攻撃手法評価

Breaking the weakest link to evade vision language models

Authors: Ilan Zini, Boussad Addad, Katarzyna Kapusta | Published: 2026-08-19
Model DoS
Certified Robustness
Adversarial Attack Methods

SMTrap: Cost-Effective DoS Attacks Against Large Reasoning Models via SMT Conflict Guidance

Authors: Jian Yang, Zhenqi Feng, Zhaoyang Yu, Zhaoxin Fan, Kejian Wu, Xiaofeng Wang, Zheng Zhu, Jianjun Huang, Wei You, Bin Liang | Published: 2026-08-19
SATソルバー
Prompt Injection
Model DoS

Verifiable abstention makes AI leak diagnosis accountable in water distribution networks

Authors: Tianwei Mu, Yue Wang, Mingzhe Yuan, Manhong Huang, Wenhong Wang, Xuerui Yin, Qing Luo, Min Xiao, Hui Yang, Jun Li, Dan Xue | Published: 2026-08-19
Data Collection Method
Reliability Assessment
Anomaly Detection Algorithm