Literature Database

Beyond Gradient-Based Attacks: Adversarial Robustness and Explainability Stability in Cybersecurity Classifiers

Authors: Mona Rajhans, Vishal Khawarey | Published: 2026-07-02
Certified Robustness
Adversarial Perturbation Techniques
脆弱性評価

Forensic-Oriented Intrusion Detection Using Synthetic Network Traffic Data and Explainable Artificial Intelligence

Authors: Jose Luis Vela Alonso, Carmen Pellicer | Published: 2026-07-01
Application of XAI
Dataset Analysis
Data Flow Analysis

HARC: Coupling Harmfulness and Refusal Directions for Robust Safety Alignment

Authors: Shei Pern Chua, Fangzhao Wu | Published: 2026-07-01
Alignment
脆弱性評価
Vulnerability Assessment Method

Cross-Domain Generalization Failure in Lightweight Intrusion Detection Models for IIoT Networks

Authors: MD Azizul Hakim, Md Shihab Uddin, Talha Ibne Anis | Published: 2026-07-01
クロスドメイン評価
Interpretability
Evaluation Method

Beyond the Prompt: Jailbreaking Function-Calling LLMs via Simulated Moderation Traces

Authors: Junlong Liu, Haobo Wang, Weiqi Luo, Xiaojun Jia | Published: 2026-07-01
Multi-Round Dialogue
Large Language Model
脱獄攻撃手法

Predicting Lethal Outcome (Cause) And Understanding Key Biomarkers Linked With Acute Myocardial Infarction Using Deep Artificial Neural Network And Ensemble Of Machine Learning Methodologies

Authors: Sagnik Ghosh | Published: 2026-07-01
Dataset Analysis
バイオマーカー分析
心疾患予測

A Penny for Your Prompts: Experiments Detecting and Mitigating LLM Usage by Survey Respondents

Authors: Zane Xu, Nathan Malkin | Published: 2026-07-01
Indirect Prompt Injection
Data Privacy Management
Large Language Model

SoK: Attack and Defense Landscape of Mobile On-device AI Systems

Authors: Yujin Huang, Xin Zheng, Xingliang Yuan, Kwok-Yan Lam | Published: 2026-07-01
Indirect Prompt Injection
Model Extraction Attack
Watermark

What’s Hidden Matters: Identifying Planning-Critical Occluded Agents using Vision-Language Models

Authors: Amirhosein Chahe, Tyler Naes, Jovin D'sa, Faizan M. Tariq, Sangjae Bae, Lifeng Zhou, David Isele | Published: 2026-07-01
Dataset Analysis
マルチモーダル安全性
評価基準

FedXDS: Leveraging Model Attribution Methods to counteract Data Heterogeneity in Federated Learning

Authors: Maximilian Andreas Hoefler, Karsten Mueller, Wojciech Samek | Published: 2026-06-30
Data Privacy Management
Training Method
Federated Learning