HARC: Coupling Harmfulness and Refusal Directions for Robust Safety Alignment Authors: Shei Pern Chua, Fangzhao Wu | Published: 2026-07-01 Alignment脆弱性評価Vulnerability Assessment Method 2026.07.01 2026.07.03 Literature Database
Cross-Domain Generalization Failure in Lightweight Intrusion Detection Models for IIoT Networks Authors: MD Azizul Hakim, Md Shihab Uddin, Talha Ibne Anis | Published: 2026-07-01 クロスドメイン評価InterpretabilityEvaluation Method 2026.07.01 2026.07.03 Literature Database
Beyond the Prompt: Jailbreaking Function-Calling LLMs via Simulated Moderation Traces Authors: Junlong Liu, Haobo Wang, Weiqi Luo, Xiaojun Jia | Published: 2026-07-01 Multi-Round DialogueLarge Language Model脱獄攻撃手法 2026.07.01 2026.07.03 Literature Database
Predicting Lethal Outcome (Cause) And Understanding Key Biomarkers Linked With Acute Myocardial Infarction Using Deep Artificial Neural Network And Ensemble Of Machine Learning Methodologies Authors: Sagnik Ghosh | Published: 2026-07-01 Dataset Analysisバイオマーカー分析心疾患予測 2026.07.01 2026.07.03 Literature Database
A Penny for Your Prompts: Experiments Detecting and Mitigating LLM Usage by Survey Respondents Authors: Zane Xu, Nathan Malkin | Published: 2026-07-01 Indirect Prompt InjectionData Privacy ManagementLarge Language Model 2026.07.01 2026.07.03 Literature Database
SoK: Attack and Defense Landscape of Mobile On-device AI Systems Authors: Yujin Huang, Xin Zheng, Xingliang Yuan, Kwok-Yan Lam | Published: 2026-07-01 Indirect Prompt InjectionModel Extraction AttackWatermark 2026.07.01 2026.07.03 Literature Database
What’s Hidden Matters: Identifying Planning-Critical Occluded Agents using Vision-Language Models Authors: Amirhosein Chahe, Tyler Naes, Jovin D'sa, Faizan M. Tariq, Sangjae Bae, Lifeng Zhou, David Isele | Published: 2026-07-01 Dataset Analysisマルチモーダル安全性評価基準 2026.07.01 2026.07.03 Literature Database
FedXDS: Leveraging Model Attribution Methods to counteract Data Heterogeneity in Federated Learning Authors: Maximilian Andreas Hoefler, Karsten Mueller, Wojciech Samek | Published: 2026-06-30 Data Privacy ManagementTraining MethodFederated Learning 2026.06.30 2026.07.02 Literature Database
A Lifecycle and Application-Stack Survey of Large Language Model Vulnerabilities: Attacks, Risks, Defenses, and Open Problems Authors: Seyed Bagher Hashemi Natanzi, Bo Tang | Published: 2026-06-30 Challenges in IT SecurityPoisoning attack on RAGPrompt Injection 2026.06.30 2026.07.02 Literature Database
Robust Text Watermarking for Large Language Models via Dual Semantic Embeddings Authors: Jonas Schäfer, Cezary Pilaszewicz, Gerhard Wunder | Published: 2026-06-30 Disabling Safety Mechanisms of LLMDigital Watermarking for Generative AIWatermarking Technology 2026.06.30 2026.07.02 Literature Database