arXiv論文メモ
新着一覧
eess.IV / cs.IT / math.IT · 査読状況未確認

受信済みの低画質画像を使って追加配信の通信量を減らす

Opportunistic Conditional Entropy Coding with Frozen Analysis and Synthesis Transforms

Vincent Corlay, Maxime Rousselot, Andriy Enttsel

この論文をやさしく読む

ひとことで言うと

すでに持っている低画質画像を利用し、高画質版を追加で受け取る通信量を減らす仕組みです。手元に画像がない場合にも同じモデルで対応します。

何に役立つ?

画質を後から上げる画像配信で、既存の学習型符号化器の復元処理を変えずに転送量を削減する用途があります。

この研究の面白いところ

補助情報の有無で復元画像は変えず、必要な符号量だけを変えています。補助情報を利用できないときの不利も4%未満と報告しています。

どこまで分かった?

最大46%という値は、目標の1段階下の画質が受信側にある条件です。52%は追加のハイパー潜在表現を送る場合で、初回配信を含む総通信量の削減率ではありません。

v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。

アブストラクトの日本語訳

多くの配信状況では、受信側が別の送信を通じて得た、低画質または低解像度の画像表現をすでに持っている場合がある。従来の符号化器は、後から要求された高画質の表現を、この偶然利用できる補助情報を使わずに符号化する。一方、条件付き符号化器は一般に、所定の補助情報源が常に利用できると仮定する。本研究では代わりに、補助情報がある場合もない場合もある状況を考える。 復号済みの潜在表現があればそれを条件とし、なければ標準的なハイパー事前分布へ切り替える、単一のエントロピーモデルを導入する。提案するアダプターは、補助情報の潜在表現をエントロピーモデルが必要とする事前信号へ写像し、同じモデルで、目標画質と補助情報の画質の複数の組み合わせに対応できるようにする。解析変換と合成変換は固定したままで、既存の学習型符号化器に後付けでき、潜在表現と再構成の経路を保持する。 受信側が目標の1段階下の画質を持っている場合、後続の送信量を最大46%削減し、追加のハイパー潜在表現を送る場合には52%削減する。補助情報がない場合の通信量の増加は4%未満にとどまり、条件付きモードと代替モードの再構成画像はビット単位で同一である。

v1の要旨から自動生成。本文の精読・人による確認は未実施。

初稿
2026-09-18(UTC)
最新改訂
2026-09-18 · v1
査読・掲載
査読状況未確認
arXivで読むPDF

更新履歴

取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。

原文の要旨

In many delivery settings, a receiver may already hold a lower-quality or lower-resolution representation of an image, obtained through an independent transmission. Conventional codecs encode a subsequently requested higher-quality representation without exploiting this incidental side information, whereas conditional codecs generally assume a prescribed source of side information that is always available. We instead consider an opportunistic setting in which side information may or may not be present. We introduce a single entropy model that conditions on a previously decoded latent when available and falls back to a standard hyperprior otherwise. The proposed adapter maps the side-information latent to the prior signal required by the entropy model, allowing the same model to support multiple target and side-information quality combinations. The analysis and synthesis transforms remain frozen, enabling retrofitting of an existing learned codec while preserving its latent representation and reconstruction path. When the receiver holds the quality immediately below the target, the proposed method reduces the rate of the subsequent transmission by up to 46%, or by 52% when an additional hyper-latent is transmitted. In the absence of side information, the rate penalty remains below 4%, and the reconstructions are bit-identical across the conditional and fallback modes.

arXiv ID: 2609.21816 / 要約の誤りについて