Literature Database

Finite-Sample Probabilistic Safety Certification for AI-Based Grid-Edge Coordination

Authors: Yihong Zhou, Hanbin Yang, Thomas Morstyn | Published: 2026-09-23
Model Robustness
Model evaluation methods
安全性調整手法

Your Model Is Leaking: Covert Information Transfer through LLM Residual Streams

Authors: Mingyuan Li, Yanna Jiang, Guangsheng Yu, Qin Wang, Xu Wang, Wei Ni, Ren Ping Liu | Published: 2026-09-23
Model Extraction Attack
Causes of Information Leakage
Attack Detection

A Bulletproof Business? Towards Detecting Infrastructure-as-a-Service Offerings on Telegram

Authors: Roy Ricaldi, Kristiyan Kyurkchiev, Irdin Pekaric | Published: 2026-09-23
Telegramにおけるサイバー犯罪
サイバー犯罪サービス
Backdoor Detection

Only Pay What You Must Spend: On-Demand Privacy Budget Payment for Differentially Private RAG

Authors: Zhonghao Sun, Zhiliang Tian, Xinyue Fang, Shuo Ma, Juhua Zhang, Yiping Song, Dongsheng Li | Published: 2026-09-23
RAG
Privacy Analysis
Differential Privacy

Psychoacoustically Aligned Latent Smoothing for Adversarial Robustness of Full-Duplex Speech-to-Speech Dialogue Models

Authors: Kian Shamsaie, Iman Modarressi | Published: 2026-09-23
Model DoS
Model Robustness
Adversarial attack

Automated Extraction of Records of Processing Activities (RoPA) Using Hybrid RAG and Locally Deployed Large Language Models

Authors: To Duy Hinh, Nguyen Le Quoc Anh, Phan Van Tri, Khuong Nguyen-An | Published: 2026-09-23
Dataset Analysis
Data Management System
Information Extraction Method

Turning Safety into Competence: Minimally Exploitable Robot Policies via Safety-Filtered Reinforcement Learning

Authors: Ruihan Wu, Rui Yang, Donggeon David Oh, Duy Nguyen, Haimin Hu | Published: 2026-09-23
Game Theory
Policy engineering
安全性と競争力

Multi-View Fusion for Encrypted C2 Detection: A Leakage-Controlled Measurement Study of Evaluation Pitfalls

Authors: Hoang-Huy Nguyen-Huu, Van-Tri Phan, Khuong Nguyen-An | Published: 2026-09-23
攻撃者のアプローチ
Feature Selection Method
通信セキュリティ

CAVEAT: Towards Robust Computer-Use Agents in Incentive-Misaligned Environments

Authors: Yuxuan Li, Will Epperson, Wesley Deng, Zezhou Huang | Published: 2026-09-23
Bias
Model Robustness
意思決定プロセスの分析

A2M: Trace-Optimized Agent Hijacking in the MCP Ecosystem

Authors: Laizhen Li, Xuan Wang, Peicheng Zhao, Juanjuan Zhao, Kejiang Ye, Cheng-zhong Xu, Xitong Gao | Published: 2026-09-22
システム設定
情報漏洩
攻撃戦略分析