DreamGuard: Efficient Runtime Guardrail for LLM Agents via Risk-Aware World Model Authors: Wenhao Lin, Chenyu Yu, Xingwei Lin, Sicong Cao, Xiang Chen, Lei Xue, Le Yu, Letian Sha, Chunming Wu | Published: 2026-08-06 Indirect Prompt InjectionリスクとメカニズムWatermark Evaluation 2026.08.06 2026.08.08 Literature Database
When Experience Becomes Instruction: Trajectory Poisoning in Self-Evolving Agent Skill Systems Authors: Jialuo Chen, Lingqi Jiang, Xinhao Deng, Xiaohu Du, Jianan Ma, Yunhao Feng, Yuqi Qing, Zhihao Yuan, Linkang Du, Jingyi Wang | Published: 2026-08-06 Relationship of AI Systems攻撃計画手法Research Methodology 2026.08.06 2026.08.08 Literature Database
PromptShield Home: Ambient Multimodal Prompt Injection Defense for Smart-Home Agents Authors: He Zhang, Feilong Li, Dingning Long, Yilin Cui, Peijun Zhang, Yuewen Zhang, Qianyao Xu, Xinyi Fu | Published: 2026-08-06 データセットの多様性Prompt InjectionMalfunction of Voice Assistants 2026.08.06 2026.08.08 Literature Database
Hardware Design and Security in the Era of Chiplets and LLMs Authors: Johann Knechtel, Ozgur Sinanoglu, Paul V. Gratz, Ramesh Karri | Published: 2026-08-05 Cooperative Effects with LLMIntellectual Property ProtectionDefense Method 2026.08.05 2026.08.07 Literature Database
Gradient Immunity: Null-Space Resistance to Malicious Fine-Tuning Authors: Yuxuan Huang, Xingyu Zeng, Tianhang Zheng, Chaochao Lu | Published: 2026-08-05 透かし手法Watermark EvaluationDefense Method 2026.08.05 2026.08.07 Literature Database
Private Direct Preference Optimization for LLM Alignment Authors: Yangfan Jiang, Fei Wei, Ergute Bao, Xiaokui Xiao, Yaliang Li, Bolin Ding | Published: 2026-08-05 AlignmentPrivacy Protection MethodPersonalization Method 2026.08.05 2026.08.07 Literature Database
LLM-Assisted Detection and Repair of Hardware Security Vulnerabilities in Verilog Designs Authors: Ethen Santana, Gabriel Gyaase, Hao Zheng | Published: 2026-08-05 Cooperative Effects with LLMPrompt leakingVulnerability Prediction 2026.08.05 2026.08.07 Literature Database
When Shared Rollouts Fail in Defensive Driving Evaluation: A NAVSIM Score Basis Audit Authors: Ziang Wei, Minjun Yu, Zheyuan Lai, Mingjie Pang, Wei Li | Published: 2026-08-05 Simulation Result Evaluation行動分析手法Watermark Evaluation 2026.08.05 2026.08.07 Literature Database
PURPOSE: Poisoning Conflict Resolution in RAG via Proxy-Fact-Grounded Updates Authors: Zijian Wang, Yubo Zhu, Muzhi Dong, Yanjun Lou, Yisheng Li, ZiLiang Zhang, Wei Tong, Yuan Zhang, Jingyu Hua, Sheng Zhong | Published: 2026-08-05 Poisoning attack on RAG攻撃計画手法Defense Method 2026.08.05 2026.08.07 Literature Database
“Allow” to Achieve, Over-Privileged Inadvertently: The Unintended Cost of Task-Completion-Driven Pop-up Decisions in Mobile GUI Agents Authors: Dongsheng Chen, Yuxuan Li, Guanhua Chen, Jiaxin Zhang, Xiangyu Zhao, Lei Ma, Xin Yao, Xuetao Wei | Published: 2026-08-05 Task DesignPrompt leakingリスクとメカニズム 2026.08.05 2026.08.07 Literature Database