arXiv論文メモ
新着一覧
hep-ex · 査読状況未確認

高エネルギー物理のデータをGPUへ高速転送する試作系

A-GHOST: High-rate streaming of trigger-level data to programmable GPU inference

I. Xiotidis, N. Clarke Hall, M. S. Larson, R. Gurunathan, C. Burdick, T. Du, N. Konstantinidis, K. Kordas, D. Leshchev, D. W. Miller, V. A. Petrovic, D. Sampsonidou, A. Thompson, T. Wengler

この論文をやさしく読む

ひとことで言うと

検出器側ではデータを集めて送り、GPU側で複雑な判断を行うための高速転送・推論システムを試作しています。

何に役立つ?

高エネルギー物理実験で、固定遅延ハードウェアでは扱いにくいニューラルネットワークを利用する構成の検討に役立ちます。

この研究の面白いところ

転送したパケットをGPUでそのまま推論入力へ組み直し、40〜100 Gbpsの入力と低い推論遅延を両立させています。複数時間の連続動作も確認しています。

どこまで分かった?

実証は開発キットとソフトウェア送信器、外部ループバックによるバックエンド試験です。FPGAデータ源との統合は今後の段階であり、実際の検出器全体での運用結果ではありません。

v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。

アブストラクトの日本語訳

A-GHOST(A Global Heterogeneous Online Scouting Trigger)は、高エネルギー物理(HEP)実験向けの高速ストリーミング読み出し構成を調べる研究開発である。中心となる考えは、検出器に近い決定論的なフロントエンドと固定遅延処理を、判断を行う基盤からデータの集約・ストリーミング供給源へ転換することにある。小型化したトリガーレベルのデータをGPU対応バックエンドへ流し、ハードウェアトリガーの資源および遅延の制約を超える、はるかに複雑なアルゴリズムを実行できるようにする。 本論文は、NVIDIA IGX Thor開発キットを使う概念実証用バックエンドを示す。ソフトウェアで送信レートを制御する送信器が、QSFPインターフェースと外部ループバックケーブルを介して第2のQSFPインターフェースへデータを送る。これにより、FPGAのデータ源と統合する前に、ネットワークからGPUへの経路を調べられる。NVIDIA DAQIRIはGPUからアクセスできるメモリへの直接受信を可能にし、独自のCUDAカーネルはパケットのペイロードを、永続的で連続したTensorRT入力ウィンドウへ再構成する。この際、データ型変換は行わず、変換はモデル側が担当する。 上位10個のカロリメータクラスタからなるHEP由来の事象表現を用いると、バックエンドは40から100 Gbpsへ拡張しながら、一定の入力レートと95パーセンタイルで0.118 msの推論遅延を維持する。システムはデータの脱落や推論エラーなしに複数時間動作する。生成型および時間的情報を扱うニューラルネットワークにより、固定遅延のFPGAトリガーシステムで通常使われるものを超える処理負荷を実演する。この結果は、A-GHOSTが将来のFPGA–GPU間ストリーミング読み出しシステムの基盤になることを示す。

v1の要旨から自動生成。本文の精読・人による確認は未実施。

初稿
2026-10-01(UTC)
最新改訂
2026-10-01 · v1
査読・掲載
査読状況未確認
arXivで読むPDF

更新履歴

取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。

原文の要旨

A-GHOST (A Global Heterogeneous Online Scouting Trigger) is an R&D effort investigating high-rate streaming readout architectures for high-energy physics (HEP) experiments. The central idea is to convert the deterministic front-end and fixed-latency processing close to the detector from a decision-making platform to an aggregation and streaming source. Compact trigger-level data are streamed to a GPU-enabled backend, where substantially more complex algorithms can run beyond the resource and latency envelope of the hardware trigger. This paper presents a proof-of-concept backend using the NVIDIA IGX Thor development kit. A software rate-controlled transmitter sends data through a QSFP interface and external loopback cable to a second QSFP interface, allowing the network-to-GPU path to be studied before integration with FPGA sources. NVIDIA DAQIRI enables reception directly into GPU-accessible memory, while a custom CUDA kernel reassembles packet payloads into persistent, contiguous TensorRT input windows without data-type conversion, which is handled by the models. Using a HEP-derived event representation consisting of the ten leading calorimeter clusters, the backend scales from 40 to 100 Gbps while sustaining a constant input rate and an inference latency of 0.118 ms at the p95 interval. The system operates for multiple hours without drops or inference errors. Generative and temporally aware neural networks demonstrate workloads beyond those normally deployed in fixed-latency FPGA trigger systems. The result establishes A-GHOST as a basis for future FPGA-to-GPU streaming readout systems.

著者のコメント

17 pages, 11 figures

arXiv ID: 2610.01761 / 要約の誤りについて