文献データベース

Beyond Gradient-Based Attacks: Adversarial Robustness and Explainability Stability in Cybersecurity Classifiers

Authors: Mona Rajhans, Vishal Khawarey | Published: 2026-07-02
モデルの頑健性保証
敵対的摂動手法
脆弱性評価

Forensic-Oriented Intrusion Detection Using Synthetic Network Traffic Data and Explainable Artificial Intelligence

Authors: Jose Luis Vela Alonso, Carmen Pellicer | Published: 2026-07-01
XAIの応用
データセット分析
データ流分析

HARC: Coupling Harmfulness and Refusal Directions for Robust Safety Alignment

Authors: Shei Pern Chua, Fangzhao Wu | Published: 2026-07-01
アライメント
脆弱性評価
脆弱性評価手法

Cross-Domain Generalization Failure in Lightweight Intrusion Detection Models for IIoT Networks

Authors: MD Azizul Hakim, Md Shihab Uddin, Talha Ibne Anis | Published: 2026-07-01
クロスドメイン評価
解釈可能性
評価手法

Beyond the Prompt: Jailbreaking Function-Calling LLMs via Simulated Moderation Traces

Authors: Junlong Liu, Haobo Wang, Weiqi Luo, Xiaojun Jia | Published: 2026-07-01
マルチラウンド対話
大規模言語モデル
脱獄攻撃手法

Predicting Lethal Outcome (Cause) And Understanding Key Biomarkers Linked With Acute Myocardial Infarction Using Deep Artificial Neural Network And Ensemble Of Machine Learning Methodologies

Authors: Sagnik Ghosh | Published: 2026-07-01
データセット分析
バイオマーカー分析
心疾患予測

A Penny for Your Prompts: Experiments Detecting and Mitigating LLM Usage by Survey Respondents

Authors: Zane Xu, Nathan Malkin | Published: 2026-07-01
インダイレクトプロンプトインジェクション
データプライバシー管理
大規模言語モデル

SoK: Attack and Defense Landscape of Mobile On-device AI Systems

Authors: Yujin Huang, Xin Zheng, Xingliang Yuan, Kwok-Yan Lam | Published: 2026-07-01
インダイレクトプロンプトインジェクション
モデル抽出攻撃
透かし

What’s Hidden Matters: Identifying Planning-Critical Occluded Agents using Vision-Language Models

Authors: Amirhosein Chahe, Tyler Naes, Jovin D'sa, Faizan M. Tariq, Sangjae Bae, Lifeng Zhou, David Isele | Published: 2026-07-01
データセット分析
マルチモーダル安全性
評価基準

FedXDS: Leveraging Model Attribution Methods to counteract Data Heterogeneity in Federated Learning

Authors: Maximilian Andreas Hoefler, Karsten Mueller, Wojciech Samek | Published: 2026-06-30
データプライバシー管理
トレーニング手法
連合学習