arXiv論文メモ
新着一覧
quant-ph / cs.DC / cs.LG / cs.PL · 査読状況未確認

量子計算のコンパイルと実行先を強化学習で選ぶ構想

MQSS-Selector: RL-Guided Pass Selection for an MLIR Compilation Pipeline

Andre Youssefi (1), Ercüment Kaya (1 and 2), Minh Chung (1), Jorge Echavarria (3), Laura B. Schulz (4), Martin Schulz (1 and 2) ((1) Leibniz Supercomputing Centre (LRZ), (2) Technical University of Munich (TUM), (3) Munich Quantum Valley (MQV), (4) Argonne National Laboratory (ANL))

この論文をやさしく読む

ひとことで言うと

量子計算で、装置選択、コンパイルの処理順、ジョブの順番を一体で選ぶ学習型の仕組みを提案しています。

何に役立つ?

量子計算と高性能計算を組み合わせた基盤で、複数の運用目標をまとめて扱う設計の参考になります。

この研究の面白いところ

忠実度だけでなく、コンパイル時間と待ち時間まで含めて、回路や装置の状態に合わせた選択を目指します。

どこまで分かった?

要旨は方式の提案と拡張可能性を述べており、実験結果や改善幅は示していません。

v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。

アブストラクトの日本語訳

高性能計算(HPC)と量子計算(QC)は、古典計算と量子計算の作業をつなぐ必要性から、統合されたHPC・QC基盤へ向かいつつある。これはハードウェアからコンパイラー、実行時環境、アプリケーションまで、システムの各層に影響する。しかし現在の量子装置は雑音の多い中規模量子計算(NISQ)の段階にあり、誤りが多く資源も限られるため、十分な忠実度を得るには専用の最適化と装置の接続構造への割当てが必要である。したがって量子ソフトウェア全体の中でも、適切なコンパイルと最適化が重要になる。現在の多くの環境では、装置選択、コンパイラーの処理段階の最適化、ジョブ待ち行列の管理が別々の部品に分かれている。本論文は、これらを一つの枠組みに統合する、学習に基づく選択器を提案する。提案する選択方式は強化学習と深層学習のモデルを使い、回路の特徴や装置の状態へ動的に適応しながら、忠実度、コンパイル時間、待ち行列での遅延など複数の目的を同時に最適化できるよう拡張可能である。

v1の要旨から自動生成。本文の精読・人による確認は未実施。

初稿
2026-09-24(UTC)
最新改訂
2026-09-24 · v1
査読・掲載
査読状況未確認
arXivで読むPDF

更新履歴

取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。

原文の要旨

High Performance Computing (HPC) and Quantum Computing (QC) systems are increasingly converging towards unified High Performance Computing-Quantum Computing (HPCQC) infrastructures, driven by a growing need to bridge classical and quantum workflows, which affects all levels of the system stack, from the hardware to compilers and runtimes, all the way to applications. However, today's QC devices are still in the Noisy Intermediate-Scale Quantum (NISQ) era, are error-prone and resource-limited, and therefore require specialized optimizations and topology mappings to achieve sufficient fidelity. This places special emphasis on proper compilation and optimization within the overall quantum software stack. Many existing stacks remain fragmented, with separate components responsible for device selection, compiler-pass optimization, and job queue scheduling. This paper proposes a unified, learning-based selector that integrates these disparate stages into a cohesive framework. Our proposed selector scheme leverages reinforcement learning and deep learning models that can be extended to simultaneously optimize multiple objectives -- such as fidelity, compilation time, and scheduling latency -- while dynamically adapting to circuit characteristics and device conditions.

著者のコメント

11 pages, 5 figures, 1 table

arXiv ID: 2609.30104 / 要約の誤りについて