arXiv論文メモ
新着一覧
cs.AR · 査読状況未確認

公開ツールでStreamNTTのハードウェア性能を比較

Benchmarking StreamNTT with a Verilog-to-Routing Toolchain

Wei He and Young-kyu Choi and Hyunwoo Park and Sunwoong Kim

この論文をやさしく読む

ひとことで言うと

耐量子暗号で使う数論変換の高速化器を公開ツールで構築し比較した研究。

何に役立つ?

商用ツールがない環境でハードウェア設計を再現・比較する際に役立つ。

この研究の面白いところ

演算資源の使用は近かったが、内部メモリの使用に差が出た点。

どこまで分かった?

要旨は資源使用の比較を報告し、処理量や電力についての新たな数値は示していない。

v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。

アブストラクトの日本語訳

耐量子暗号アルゴリズムの大規模なデータセンター導入が進むにつれ、計算上のボトルネックである数論変換NTTのハードウェア高速化に関心が集まっている。高位合成とFPGAを使う高速化器StreamNTTは、さまざまな最適化によって最高水準の処理量を得ている。しかし商用ツールと特定の装置に依存するため、それらを利用できない研究者には直接比較が難しい。本研究は、公開されたVerilog-to-Routingのツール群でStreamNTTを構築することでこの問題に対応した。デジタル信号処理資源と乗算器の使用量は同程度だった。一方、内部メモリの使用量には大きな差があり、商用ツールの性能に近づくにはメモリの段階でさらに最適化する必要があることを示した。

v1の要旨から自動生成。本文の精読・人による確認は未実施。

初稿
2026-09-19(UTC)
最新改訂
2026-09-19 · v1
査読・掲載
査読状況未確認
arXivで読むPDF

更新履歴

取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。

原文の要旨

As post-quantum cryptography algorithms move toward large-scale data center deployment, hardware acceleration of their computational bottleneck, which is the number theoretic transform (NTT), has gained increasing attention. StreamNTT, a high-level synthesis- and field-programmable gate array-based accelerator, achieves state-of-the-art throughput through various optimization techniques. However, its reliance on a commercial tool and a device makes direct comparisons difficult for researchers without access. We address this by building StreamNTT on an open-source Verilog-to-Routing toolchain, which achieves similar digital signal processing and multiplier usage. Significant differences in internal memory utilization indicate that further memory-level optimization is needed to approach commercial tool performance.

著者のコメント

Accepted to the 2nd Workshop on Domain-Specialized FPGAs (WDSFPGA), co-located with ISFPGA 2026

arXiv ID: 2609.23116 / 要約の誤りについて