文献データベース

K-ABENA: K-Adaptive Backpropagation with Error-based N-exclusion Algorithm : (Compensated Loss-Based Sample Exclusion with Unbiased Gradient Estimation)

Authors: Jean-Francois Bonbhel | Published: 2026-07-07
コストモデル
トレーニングプロトコル
モデル設計

i-EXAM: Instructable and Explainable Attack Connectivity Graph Modeler

Authors: Rakesh Podder, Wadia Ganim, Sarath Sreedharan, Indrajit Ray, Indrakshi Ray | Published: 2026-07-07
エージェント操作手法
脆弱性攻撃手法
防御手法

Code-Level Cost Function Generation for Spatial Image Steganography Using RAG-Enhanced Large Language Models

Authors: Yige Wang, Shiqi Yi, Hanzhou Wu | Published: 2026-07-07
RAG
RAGへのポイズニング攻撃
ステガノグラフィー手法

Beyond Refusal: A Same-Lineage Study of Aligned and Abliterated LLMs for Vulnerability Analysis

Authors: Mingchen Li, Meikang Qiu, Zifan Peng, Heng Fan, Song Fu, Junhua Ding, Yunhe Feng | Published: 2026-07-07
アライメント
プロンプトインジェクション
脆弱性検出

Unicode TAG-Block Concealment of Tool-Metadata Payloads in the Model Context Protocol: An Approval-View Fidelity Gap Across Three Independent Server Implementations

Authors: Mohammadreza Rashidi | Published: 2026-07-07
ツールの脆弱性
脆弱性攻撃手法
脆弱性検出

The Balkanization of Execution-Security Research for AI Coding Agents: Isolation, Access Control, and Time-of-Check-to-Time-of-Use Vulnerabilities

Authors: Mohammadreza Rashidi | Published: 2026-07-07
プロンプトインジェクション
脆弱性攻撃手法
防御手法

Selective Disclosure Watermarking for Large Language Models

Authors: Xuyang Chen, Xiang Li, Yangxinyu Xie, Qi Long | Published: 2026-07-06
モデル設計
公的検証可能性
生成AI向け電子透かし

When Claws Remember but Do Not Tell: Stealthy Memory Injection in Persistent Personal Agents

Authors: Yechao Zhang, Shiqian Zhao, Jiawen Zhang, Jie Zhang, Gelei Deng, Xiaogeng Liu, Chaowei Xiao, Tianwei Zhang | Published: 2026-07-06
インダイレクトプロンプトインジェクション
エージェント操作手法
脆弱性攻撃手法

Agent Data Injection Attacks are Realistic Threats to AI Agents

Authors: Woohyuk Choi, Juhee Kim, Taehyun Kang, Jihyeon Jeong, Luyi Xing, Byoungyoung Lee | Published: 2026-07-06
インダイレクトプロンプトインジェクション
脆弱性攻撃手法
防御手法

Your Agent’s Memories Are Not Its Own: Forged Reasoning Attacks on LLM Agent Memory and Defenses

Authors: Neeraj Karamchandani, Piyush Nagasubramaniam, Sencun Zhu, Dinghao Wu | Published: 2026-07-06
インダイレクトプロンプトインジェクション
脆弱性攻撃手法
記憶攻撃手法