arXiv論文メモ
新着一覧
cs.AI / cs.IR / cs.LG · 査読状況未確認

分野ごとの機密情報を再学習なしで伏せる手法

ASIRF: An Agentic Framework for Context-Dependent Sensitive Information Redaction

Sudha Priyadarshini, Mohamed Chahine Ghanem

この論文をやさしく読む

ひとことで言うと

分野に応じて何を機密とみなすかを切り替え、文章中の情報を伏せる仕組みです。

何に役立つ?

新しい業務分野の文書で、モデルを再学習せず機密情報の検出方針を適用する際に役立ちます。

この研究の面白いところ

推論時に分野固有の定義を参照し、架空の未知分野を含めて既存の学習済みフィルターと比較しています。

どこまで分かった?

示された主な数値は再現率です。誤検出率や実運用での総合的な安全性は要旨からは判断できません。

v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。

アブストラクトの日本語訳

何が機密情報かは分野と目的で決まり、普遍的な分類はない。しかしプライバシーフィルターや固有表現抽出器などの伏せ字処理は学習時に分類体系を固定するため、新しい分野ごとに再学習が必要になる。本研究ではASIRFという仕組みを提案する。入力が属する分野に合わせ、柔軟な知識ベースから分野固有の定義を推論時に取得し、再学習なしで適応する。3回のモデル呼び出しによる複数エージェントの処理系と、単一エージェントの方式を、10種類の小型の公開重みモデルと8データセットで評価した。データには、学習時の分布から外れる架空の分野も含め、学習済み分類器であるOpenAI Privacy Filter(OPF)と比較した。分野ごとに専門家が書いた数十件の定義だけを用い、学習データは使わなかった。それでも二方式の少なくとも一方は、モデルと分野の80通りの組合せのうち68通り、すなわち85%で、OPFより高い再現率を示した。下回った例の大半はOPFの学習分布に含まれる分野に集中していた。

v1の要旨から自動生成。本文の精読・人による確認は未実施。

初稿
2026-09-24(UTC)
最新改訂
2026-09-24 · v1
査読・掲載
査読状況未確認
arXivで読むPDF

更新履歴

取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。

原文の要旨

Sensitive information is defined by domain and intent, not a universal category, yet redaction systems such as privacy filters and named-entity recognizers fix a taxonomy at training time, requiring retraining for each new domain. We introduce ASIRF (Agentic Sensitive Information Redaction Framework), which retrieves domain-specific definitions based on the input's domain from a flexible knowledge base at inference time, needing no retraining to adapt. Two architectures, a three-call multi-agent pipeline and a single-agent variant, are evaluated across ten small open-weight models and eight datasets, including out-of-distribution fictional domains, against the OpenAI Privacy Filter (OPF) as a trained-classifier baseline. With only a few dozen expert-authored definitions per domain and no training data, ASIRF's recall exceeds OPF's in 68 of 80 model-domain combinations (85 percent), by at least one of the two architectures, with shortfalls confined mostly to OPF's training-distribution domains.

著者のコメント

paper accepted in NeurIPS 2026 GlobalSouthAI

arXiv ID: 2609.29191 / 要約の誤りについて