量子機械学習のデータセットを少数の浅い回路へ蒸留する
Distilling Datasets into Shallow Circuits for Quantum Machine Learning
この論文をやさしく読む
ひとことで言うと
量子機械学習の大量の訓練標本を、直接読み込める少数の浅い量子回路にまとめる方法。
何に役立つ?
標本ごとの量子状態準備と繰り返し測定にかかる費用を減らす方法として役立つ可能性がある。
この研究の面白いところ
合成標本そのものをロード回路として表し、後から回路へ変換する段階を省いた。画像データセットで少数の回路と測定ショットで性能を比較した。
どこまで分かった?
精度とショット数の結果は要旨で示されたMNIST、Fashion-MNISTなどの評価条件に基づく。実機での検証も述べるが、その装置や規模の詳細は要旨にない。
v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。
アブストラクトの日本語訳
量子機械学習で量子モデルを学習するときは、各標本をロード回路によって量子状態として準備し、各学習ステップの各測定ショットでその回路を再実行する必要がある。そのため全体の負担は標本数と準備コストの両方に応じて増える。従来の方法は、個々の入力のロード費用を下げるか、データを蒸留・圧縮してから別途量子符号化を行う。この分離によって、得られた標本の準備が依然として高コストになったり、浅い回路への後続のコンパイル時に情報がさらに失われたりする可能性がある。 本研究は量子データセット蒸留(QDD)を提案し、データセット全体を少数の浅い回路へ直接蒸留する。各合成標本を、低ランクのテンソルネットワークに対応する階段状の回路としてパラメータ化することで、標本とそのロード回路を同じものにする。回路のパラメータは、ロード費用の明示的な制約の下で、分布の照合と厳密な勾配を用いて古典計算により最適化する。後から状態準備回路を合成する必要はない。 MNISTとFashion-MNISTでは、クラス当たり10回路、すなわち元データセットの0.0017倍の規模だけで、全データを使う学習に匹敵する精度に達し、標本数とロード費用をそろえた選択法の基準を上回った。測定ショット数が有限の学習では、累積ショット数を100分の1未満に抑えながら、全データでの精度の95%に到達した。さらに、実際の量子ハードウェア上でもQDDを検証し、実機での利用可能性を示した。
v1の要旨から自動生成。本文の精読・人による確認は未実施。
- 初稿
- 2026-09-23(UTC)
- 最新改訂
- 2026-09-23 · v1
- 査読・掲載
- 査読状況未確認
更新履歴
- v1 2026-09-23 この版を読む
取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。
原文の要旨
In quantum machine learning, training a quantum model requires each sample to be prepared as a quantum state by a loading circuit that must be re-executed for every shot at every training step. The total burden therefore scales with both the number of samples and the cost of preparation. Existing approaches reduce the loading cost of individual inputs or distill data and compress their representations before applying a separate quantum encoding. This separation can leave resulting samples costly to prepare or cause additional loss of the information during subsequent compilation into shallow circuits. We propose quantum dataset distillation (QDD), which distills the full dataset directly into a small set of shallow circuits. Each synthetic sample is parameterized as a staircase circuit corresponding to a low-rank tensor network, making the sample and its loading circuit the same object. The circuit parameters are optimized classically with exact gradients using distribution matching under an explicit loading budget, without post-hoc state-preparation synthesis. On MNIST and Fashion-MNIST, only 10 circuits per class ($0.0017\times$ the full dataset size) achieve accuracy comparable to full-data training and outperform selection baselines under matched sample and loading budgets. In finite-shot training, QDD reaches 95\% of the full-data accuracy with more than $100\times$ fewer cumulative shots. Additionally, we validate QDD on real quantum hardware, demonstrating its practical deployment potential.
arXiv ID: 2609.28229 / 要約の誤りについて