arXiv論文メモ
新着一覧
cs.LG · 査読状況未確認

局所形状と全体尺度を分けて文章から時系列を生成

ShapeLex: Decoupling Local Shape Symbolization and Global Scale Modeling for Text-Controlled Time Series Generation

Subo Wei, Jianqi Gao, Mingyan Fan, Shaorong Xie, Xinzhi Wang, Yongpeng Dong

この論文をやさしく読む

ひとことで言うと

「途中で急上昇する」といった文章から時系列を作る際、上昇や急落の形と、全体の値の大きさ・変動幅を別々に作る方法です。局所的な特徴がぼやけたり位置を外したりする問題を狙っています。

何に役立つ?

文章で指定した特徴を持つ時系列データの生成や、そのデータを使った予測課題の検討に役立ちます。形状の語彙から教師データも作ることで、手作業の注釈負担を抑える設計です。

この研究の面白いところ

生成器が直接すべての数値を決めるのではなく、形状の単語を並べて骨格を作り、最後に全体の尺度を与えます。局所構造を記号として明示する点が特徴です。

どこまで分かった?

評価は12の公開データセットなどで行われ、実データ分布との適合改善が報告されています。改善量や文章の各指示への追従率の具体値は要旨にありません。

v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。

アブストラクトの日本語訳

文章で制御する時系列生成は、実データの分布に忠実でありながら、自然言語の説明に従う系列を合成することを目指す。既存の方式は、意味の理解と系列のモデル化を一つの連続潜在空間に結び付けることが多く、局所的な意味を明示的に対応付ける要素や、全体的な連続属性と局所的な離散形状の分離が欠けている。その結果、重要な局所構造が平滑化されたり、欠落したり、誤った位置に置かれたりしうる。 本研究は、文章から系列への生成を、局所形状の離散的な記号化と、全体属性の連続的なモデル化に分離するShape Lexicon(ShapeLex)を提案する。ShapeLexはまず、上昇、急な突出、急落といった離散的な形状単位の再利用可能な語彙を学習データから導き、解釈可能な記号空間を形成する。次に自己回帰型の生成器が文章の説明に従って形状を選び、位置や持続時間などの属性を調整し、時間順に組み合わせて形状の骨格を作る。最後に混合密度型のスケールヘッドが全体の水準と変動性をモデル化してサンプリングし、現実的な全体尺度を復元する。 12の公開データセット、実際の利用者が書いた文章、および下流の予測課題を用いた実験により、ShapeLexは既存手法より実データの分布に適合する系列を生成することが示された。また、学習した語彙から対となる教師データを自動合成するため、データセットの規模に伴って増える注釈付け費用を避け、拡張性を高める。

v1の要旨から自動生成。本文の精読・人による確認は未実施。

初稿
2026-09-21(UTC)
最新改訂
2026-09-21 · v1
査読・掲載
査読状況未確認
arXivで読むPDF

更新履歴

取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。

原文の要旨

Text-controlled time series generation aims to synthesize sequences that follow natural-language descriptions while remaining faithful to real data distributions. Existing paradigms often couple semantic understanding and sequence modeling in a single continuous latent space, lacking explicit local semantic anchors and separation between global continuous attributes and local discrete shapes. As a result, key local structures may be smoothed, missed, or misplaced. We propose Shape Lexicon (ShapeLex), which decouples text-to-sequence generation into discrete symbolization of local shapes and continuous modeling of global attributes. ShapeLex first induces a reusable vocabulary of discrete shape units, such as rises, spikes, and sharp drops, from training data, forming an interpretable symbolic space. An autoregressive generator then selects shapes according to the textual description, adjusts attributes such as position and duration, and composes them in temporal order into a shape skeleton. Finally, a mixture-density scale head models and samples the overall level and volatility to restore realistic global scale. Experiments on twelve public datasets, real user-written text, and downstream forecasting tasks show that ShapeLex generates series that better match real data distributions than existing methods. In addition, paired supervision is automatically synthesized from the learned vocabulary, avoiding annotation costs that grow with dataset size and improving scalability.

arXiv ID: 2609.24003 / 要約の誤りについて