Finite-Sample Probabilistic Safety Certification for AI-Based Grid-Edge Coordination Authors: Yihong Zhou, Hanbin Yang, Thomas Morstyn | Published: 2026-09-23 Model RobustnessModel evaluation methods安全性調整手法 2026.09.23 2026.09.25 Literature Database
Your Model Is Leaking: Covert Information Transfer through LLM Residual Streams Authors: Mingyuan Li, Yanna Jiang, Guangsheng Yu, Qin Wang, Xu Wang, Wei Ni, Ren Ping Liu | Published: 2026-09-23 Model Extraction AttackCauses of Information LeakageAttack Detection 2026.09.23 2026.09.25 Literature Database
A Bulletproof Business? Towards Detecting Infrastructure-as-a-Service Offerings on Telegram Authors: Roy Ricaldi, Kristiyan Kyurkchiev, Irdin Pekaric | Published: 2026-09-23 Telegramにおけるサイバー犯罪サイバー犯罪サービスBackdoor Detection 2026.09.23 2026.09.25 Literature Database
Only Pay What You Must Spend: On-Demand Privacy Budget Payment for Differentially Private RAG Authors: Zhonghao Sun, Zhiliang Tian, Xinyue Fang, Shuo Ma, Juhua Zhang, Yiping Song, Dongsheng Li | Published: 2026-09-23 RAGPrivacy AnalysisDifferential Privacy 2026.09.23 2026.09.25 Literature Database
Psychoacoustically Aligned Latent Smoothing for Adversarial Robustness of Full-Duplex Speech-to-Speech Dialogue Models Authors: Kian Shamsaie, Iman Modarressi | Published: 2026-09-23 Model DoSModel RobustnessAdversarial attack 2026.09.23 2026.09.25 Literature Database
Automated Extraction of Records of Processing Activities (RoPA) Using Hybrid RAG and Locally Deployed Large Language Models Authors: To Duy Hinh, Nguyen Le Quoc Anh, Phan Van Tri, Khuong Nguyen-An | Published: 2026-09-23 Dataset AnalysisData Management SystemInformation Extraction Method 2026.09.23 2026.09.25 Literature Database
Turning Safety into Competence: Minimally Exploitable Robot Policies via Safety-Filtered Reinforcement Learning Authors: Ruihan Wu, Rui Yang, Donggeon David Oh, Duy Nguyen, Haimin Hu | Published: 2026-09-23 Game TheoryPolicy engineering安全性と競争力 2026.09.23 2026.09.25 Literature Database
Multi-View Fusion for Encrypted C2 Detection: A Leakage-Controlled Measurement Study of Evaluation Pitfalls Authors: Hoang-Huy Nguyen-Huu, Van-Tri Phan, Khuong Nguyen-An | Published: 2026-09-23 攻撃者のアプローチFeature Selection Method通信セキュリティ 2026.09.23 2026.09.25 Literature Database
CAVEAT: Towards Robust Computer-Use Agents in Incentive-Misaligned Environments Authors: Yuxuan Li, Will Epperson, Wesley Deng, Zezhou Huang | Published: 2026-09-23 BiasModel Robustness意思決定プロセスの分析 2026.09.23 2026.09.25 Literature Database
A2M: Trace-Optimized Agent Hijacking in the MCP Ecosystem Authors: Laizhen Li, Xuan Wang, Peicheng Zhao, Juanjuan Zhao, Kejiang Ye, Cheng-zhong Xu, Xitong Gao | Published: 2026-09-22 システム設定情報漏洩攻撃戦略分析 2026.09.22 2026.09.24 Literature Database