文献データベース

Taxonomy of Risks on Automated Fact-Checking Systems Considering its Propagation

Authors: Jun Yajima, Tatsuya Oka, Takao Okubo | Published: 2026-06-24
リスク評価手法
社会的影響
自動化ファクトチェック

An Approach for a Supporting Multi-LLM System for Automated Certification Based on the German IT-Grundschutz

Authors: Lea Roxanne Muth, Marian Margraf | Published: 2026-06-24
RAG
RAGへのポイズニング攻撃
リソース不足の課題

CrypFormBench: Benchmarking Formal Analysis Capability of Large Language Models for Cryptographic Schemes

Authors: Zhaoxuan Li, Qionglu Zhang, Hengyuan Liu, Xiaoyan Gu, Xianhui Lu, Hongbo Liu, Bingzheng Wang, Haihui Fan, Ziming Zhao, Rui Zhang, Li Zhou | Published: 2026-06-24
パフォーマンス評価
プロンプトリーキング
透かし評価

Security and Privacy in Retrieval-Augmented Generation: Architectures, Threats, Defenses, and Future Directions for Building Trustworthy Systems

Authors: Balamurugan Palanisamy, G S S Chalapathi, Vikas Hassija, Rajkumar Buyya | Published: 2026-06-24
RAG
RAGへのポイズニング攻撃
データプライバシー評価

A Red Teaming Framework for Large Language Models: A Case Study on Faithfulness Evaluation

Authors: Abrar Alotaibi, Raed Mughus, Moataz Ahmed | Published: 2026-06-24
プロンプトインジェクション
プロンプトリーキング
脆弱性評価手法

Representation Matters: An Empirical Study of Program Representations for LLM Vulnerability Reasoning

Authors: Andrew Stoltman, Johnathan Tang, Haipeng Cai | Published: 2026-06-24
プロンプトインジェクション
プロンプトリーキング
脆弱性評価手法

Decoupling Reconnaissance and Exploitation: Measuring the Capability Boundaries of LLM-Based Web Penetration Testing

Authors: Liwei Yu, Shuo Li, Ming Zhou, Ge Chu, Yan Guo | Published: 2026-06-24
LLMの安全機構の解除
エージェント設計
自動化ペネトレーションテスト

Privacy-Preserving RAG via Multi-Agent Semantic Rewriting: Achieving Confidentiality Without Compromising Contextual Fidelity

Authors: Yuanhe Zhao, Tianyu Zhang, Huafei Xing, Derek F. Wong, Jianbin Li, Tao Fang | Published: 2026-06-23
RAG
インダイレクトプロンプトインジェクション
プライバシー保護データマイニング

Themis: An explainable AI-enabled framework for Reinforcement Learning with Human Feedback

Authors: Andreas Chouliaras, Luke Connolly, Dimitris Chatzpoulos | Published: 2026-06-23
XAI(説明可能なAI)
アライメント
エージェント設計

AdversaBench: Automated LLM Red-Teaming with Multi-Judge Confirmation and Cross-Model Transferability

Authors: Khanak Khandelwal | Published: 2026-06-23
データ生成手法
プロンプトの検証
自動化ペネトレーションテスト