拡散Transformerで動く人体上の衣服変形を生成する
DiT-Garment: Garment Dynamics with Diffusion Transformers
この論文をやさしく読む
ひとことで言うと
人が動いたときの服の形を、2次元の展開座標上で学習する生成モデルです。未学習の服のデザインや素材への対応を調べています。
何に役立つ?
考えられる用途は、人体モデルに着せる服のアニメーション生成です。目標姿勢から変形を直接推定できる構成を提案しています。
この研究の面白いところ
3次元の服をUV空間の位置マップにして、2次元の拡散Transformerで扱います。人工シミュレーションで学び、実測や制作された服にも評価を広げています。
どこまで分かった?
要旨には変形誤差や計算速度の数値はありません。物理パラメータを条件にすることと、すべての出力が厳密な物理法則を満たす保証は別です。
v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。
アブストラクトの日本語訳
任意の動きをする人体モデル上で、動的な3次元衣服をモデル化するDiT-Garmentを提示する。既存手法と異なり、学習時に見ていないデザインや物理的な素材の衣服を動かせると同時に、任意の目標姿勢に対する変形を直接推定できる。そのために、2次元拡散Transformerの構造を利用して、2次元UV空間で3次元変形を学習する。結果は決定論的ではないため、生成モデルは生じ得る結果の分布を学ぶ。 テンプレート衣服は、標準姿勢の3次元人体モデルに空間的に整列させた3次元三角形メッシュとして表す。共通テンプレートや複雑なグラフ畳み込みを必要とせず異なる衣服デザインを扱うため、UV空間で表したテンプレートの3次元位置マップを拡散Transformerの条件とする。これにより、標準姿勢の体の周囲にある3次元空間の変形を暗黙的に学べる。さらに、体の動きと物理パラメータを条件に加えて、モデルを物理に根差したものにする。 人工データと実データの両方で、DiT-Garmentを定量的・定性的に評価する。自動生成した衣服デザインの人工シミュレーションだけで学習したにもかかわらず、実測された衣服やアーティストが制作した衣服のデザインへ一般化する。コードとデータは、研究目的でhttps://dumoulina.github.io/dit-garment/から利用できる。
v1の要旨から自動生成。本文の精読・人による確認は未実施。
- 初稿
- 2026-09-16(UTC)
- 最新改訂
- 2026-09-16 · v1
- 査読・掲載
- 査読状況未確認
更新履歴
- v1 2026-09-16 この版を読む
取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。
原文の要旨
We present DiT-Garment to model dynamic 3D clothing over human body models in arbitrary motion. Unlike existing methods, DiT-Garment can animate garments with unseen designs and physical materials, while allowing for direct inference of deformations for any target pose. To achieve this, we leverage a 2D diffusion transformer architecture to learn 3D deformations in a 2D UV-space. As the result is non-deterministic, our generative model learns the distribution of possible outcomes. The template garment is represented as a 3D triangle mesh spatially aligned with a 3D human body model in a standardized pose. To work with different garment designs without the need of a common template or complex graph convolution operations, the diffusion transformer is conditioned on a 3D position map of the template, represented in UV-space, which allows to implicitly learn a deformation of the 3D space around the body in standard pose. Further conditioning on body motion and physical parameters allows to physically ground the model. We quantitatively and qualitatively evaluate DiT-Garment on both synthetic and real data. While only trained on synthetic simulations of automatically generated cloth designs, our method generalizes to captured and artist-made garment designs. Code and data are available for research purposes at https://dumoulina.github.io/dit-garment/.
arXiv ID: 2609.18510 / 要約の誤りについて