AIセキュリティポータルbot

Traffic-Aware Randomized Smoothing for LLM-Based Network Intrusion Detection

Authors: Zhenpeng Li | Published: 2026-07-15
データセットの影響
Threat Model
評価基準

Towards quantum machine learning for assessing the resilience of post-quantum cryptography

Authors: Jarosław A. Miszczak | Published: 2026-07-15
Data Augmentation in Encrypted Domains
Generative Model
Quantum Computing Method

How Agents Ask for Permission: User Permissions for AI Agents, from Interfaces to Enforcement

Authors: Alexandra E. Michael, Franziska Roesner | Published: 2026-07-15
Indirect Prompt Injection
エージェント操作手法
Threat Model

Protective Capacity Hallucination: When Large Language Models Claim Nonexistent Capabilities

Authors: Eunna Lee, Jungpyo Nam, Sunjun Hwang | Published: 2026-07-15
Relationship of AI Systems
Data Collection Method
Detection of Hallucinations

UTS at ELOQUENT 2026 Voight-Kampff: structural shifts in AI writing bypass state-of-the-art detectors

Authors: Dima Galat, Marian-Andrei Rizoiu | Published: 2026-07-15
Relationship of AI Systems
データ毒性攻撃
敵対的オブジェクト生成

Adversarial Prompting Framework for AI Safety Assessment

Authors: Yash Bhatnagar, Kunal Banerjee, Anirban Chatterjee | Published: 2026-07-15
Prompt Injection
敵対的オブジェクト生成
Threat Model

DREA: Decoupled Reasoning and Exploration Agents for Repository-Level Vulnerability Detection

Authors: Mingyang Sun, Guozhu Meng | Published: 2026-07-15
Disabling Safety Mechanisms of LLM
Vulnerability Prediction
評価基準

Silent Alarm: A J-Space Protocol for Comparing Danger Recognition Across Models and Quantization Levels

Authors: Roman Prosvirnin, Victor Minchenkov, Alexey Soldatov, Vladimir Bashun | Published: 2026-07-14
Trade-off Analysis
Model evaluation methods
評価基準

Bulkhead: Automated Semantic Detection and Remediation of Container Escape Vulnerabilities

Authors: Qiyuan Fan, Zhi Li, Junjie Li, XiaoFeng Wang, Bin Yuan, Deqing Zou | Published: 2026-07-14
コンテナセキュリティ
Prompt Injection
脆弱性優先順位付け

PVDetector: Detecting Prompt Injection Attacks on Purpose-Specific LLM Agents through Policy-Violation Concept Analysis

Authors: Junhui Wang, Hangtao Zhang, Zhirun Zheng, Li Zeng, Jiejun Xiao, Xi Luo, Lihua Yin, Saiqin Long | Published: 2026-07-14
Indirect Prompt Injection
エージェント操作手法
Behavior Manipulation Attack