ノイズ・雨・ぼけを1つのネットワークで除く画像復元
MDIRNET: Multi-Degradation Image Restoration Network via Deep Unfolding
この論文をやさしく読む
ひとことで言うと
画像のノイズ、雨、ぼけを、種類ごとの別モデルに切り替えず1つのネットワークで取り除く方法です。画像の領域ごとに必要な表現の複雑さを変えます。
何に役立つ?
複数種類の画像復元を1つのモデルで行う用途に役立ちます。ノイズ除去、ぼけ除去、雨除去の標準ベンチマークと、人工的に組み合わせた劣化で評価しています。
この研究の面白いところ
画像の小領域にある共通構造を低ランクモデルで捉え、反復的な計算手順を学習可能なネットワークにしています。各領域の内容や劣化に合わせて部分空間の次元を割り当てます。
どこまで分かった?
ここでの統一はノイズ・雨・ぼけの3種類を共同学習する意味です。混合劣化の結果は評価した人工的な組合せに関するもので、未知の実画像のあらゆる劣化で性能を保証したものではありません。要旨に個別の数値はありません。
v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。
アブストラクトの日本語訳
実際の画像には、種類が未知の劣化や複数の劣化の混在がよく見られる。同じ観測画像の中で複数の歪みが相互作用するため、その復元は単一タスクの画像復元より大幅に難しい。そのため既存の手法は、劣化の種類についての事前知識やタスクごとの別モデルに依存することが多く、細かい構造を過度に平滑化したり、アーティファクトを残したりする場合があり、小規模なモデル駆動型の代替手法が求められる。 本研究では、モデルに基づく低ランク事前分布と端から端までの学習を組み合わせた統一的な枠組みMulti-Degradation Image Restoration Network(MDIRNET)を提案する。ここで統一的とは、ノイズ、雨、ぼけという3種類の劣化について共同学習することを指す。1つのMDIRNETモデルで3種類すべてを復元でき、推論時にタスク固有のモデル、モジュール、分岐を必要としない。低ランクの事前分布は、自然画像パッチの冗長性とコンパクトな構造を利用する。その背後にある低次元表現を特定するため、直交変分PCA(OVPCA)を通じて復元を定式化し、その反復推論を深層展開ネットワークへ変換する。 空間的に不均一な劣化や局所的な内容の違いに対応するため、学習可能なパッチ分割戦略と、領域ごとに適切な部分空間の次元を予測する軽量な動的ランク割当モジュールをさらに導入する。教師あり注意機構のモジュールを用いて、空間に適応した再構成の改善を行う。標準的なノイズ除去、ぼけ除去、雨除去のベンチマークでの広範な実験により、MDIRNETは大半の指標で強力な基準手法と同等かそれ以上の性能を達成した。さらに、統制された混合劣化の実験では、評価した人工的な劣化の組合せにわたって安定した性能が示された。コードはhttps://github.com/ScholarForge/mdirnet.gitで公開されている。
v1の要旨から自動生成。本文の精読・人による確認は未実施。
- 初稿
- 2026-10-01(UTC)
- 最新改訂
- 2026-10-01 · v1
- 査読・掲載
- 査読状況未確認
更新履歴
- v1 2026-10-01 この版を読む
取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。
原文の要旨
Real images often exhibit unknown and mixed degradations, making restoration substantially more challenging than single-task image restoration because multiple distortion types interact within the same observation. Consequently, existing methods often rely on prior knowledge of the degradation type or separate task-specific models, which may oversmooth fine structures or leave residual artifacts, motivating a compact model-driven alternative. We propose the Multi-Degradation Image Restoration Network (MDIRNET), a unified framework that combines a model-driven low-rank prior with end-to-end learning. Here, unified refers to joint training on three degradation types: noise, rain, and blur. A single MDIRNET model restores all three without requiring task-specific models, modules, or branches at inference. The low-rank prior exploits the redundancy and compact structure of natural image patches. To identify this underlying low-dimensional representation, we formalize restoration via Orthogonal Variational PCA (OVPCA) and translate its iterative inference into a deep unfolding network. To handle spatially non-uniform corruption and local content variability, we further introduce a learnable patch-partitioning strategy and a lightweight dynamic rank-allocation module that predicts the appropriate subspace dimension for each region. Spatially adaptive reconstruction refinement is performed using a supervised attention module. Extensive experiments on standard denoising, deblurring, and deraining benchmarks show that MDIRNET achieves competitive or superior performance over strong baselines across most metrics, while controlled mixed-degradation experiments demonstrate consistent performance across the evaluated synthetic degradation combinations. The code is available at https://github.com/ScholarForge/mdirnet.git.
著者のコメント
Accepted for publication in IEEE Transactions on Instrumentation and Measurement (IEEE TIM), 2026
arXiv ID: 2610.01655 / 要約の誤りについて