AIセキュリティポータルbot

Active Learning on Adversarially Corrupted Graphs

Authors: Marco Bressan, Nicolò Cesa-Bianchi, Tommaso d`Orsi, Emmanuel Esposito, Silvio Lattanzi | Published: 2026-07-06
敵対的学習
脆弱性攻撃手法
高次元データ分析

HilEnT: Hilbert, Entropy Transformed Image Based Malware Detection

Authors: Rahul Kale, Thesath Wijayasiri, Kar Wai Fok, Vrizlynn L. L. Thing | Published: 2026-07-06
マルウェア可視化
機械学習手法
深層学習モデル

FORGE: Research-Trajectory Hijacking Attacks on Deep Research Agents

Authors: Yue Pan, Ziheng Zhang, Junxiang Lei, Changhao Jia, Qingyi Si, Hongcheng Guo | Published: 2026-07-06
敵対的サンプルの脆弱性
脆弱性攻撃手法
防御手法

Retroactive Chain-of-Thought (RetroCoT): Forensic Reconstruction Prompts as a Safety Diagnostic Across Model Generations

Authors: Samira Hajizadeh | Published: 2026-07-06
アライメント
エージェント操作手法
モデル評価

Distributed Attacks in Persistent-State AI Control

Authors: Josh Hills, Ida Caspary, Asa Cooper Stickland | Published: 2026-07-02
エージェント操作手法
バックドアモデルの検知
脆弱性検出手法

LACUNA: A Testbed for Evaluating Localization Precision for LLM Unlearning

Authors: Matteo Boglioni, Thibault Rousset, Siva Reddy, Marius Mosbach, Verna Dankers | Published: 2026-07-02
トレーニングプロトコル
モデル評価
深層学習モデル

Criticality-Based Guard Rail Validation for AI Agent Decisions in Autonomous Telecom Networks

Authors: Ravi Kant Sharma | Published: 2026-07-02
エージェント操作手法
マルチエージェントシステム
脆弱性検出

The Eticas AI Risk Taxonomy: Open Infrastructure for Operationalizing AI Audits

Authors: Gemma Galdon Clavell, Pablo Accuosto, Usman Gohar | Published: 2026-07-02
アライメント
リスクとメカニズム
公平性の確保

Privacy-Preserving and Verifiable Approximate Distributed Coded Computing

Authors: Xavier Martínez-Luaña, Alba Gude-Santos, Manuel Fernández-Veiga, Rebeca P. Díaz-Redondo | Published: 2026-07-02
ビザンチン攻撃対策
差分プライバシー
脆弱性評価

A rubric-based controlled comparison of frontier language models on expert-authored clinical reasoning tasks

Authors: Samiha A. Ismail, Fan X. Chen, Ali Merali | Published: 2026-07-02
モデル評価
レビューと調査
医療診断属性