arXiv論文メモ
新着一覧
cs.CV · 査読状況未確認

一枚の画像から衣服の縫製図上に一貫した3Dテクスチャを作る

OmniFabric: Coherent UV Space Texture Synthesis for 3D Garment Reconstruction

Ding-Jiun Huang, Yuanhao Wang, Cheng Zhang, Hugo Bertiche, Alexandru-Eugen Ichim, Thabo Beeler, Fernando De la Torre

この論文をやさしく読む

ひとことで言うと

一枚の衣服画像から、縫製パターンに沿って一貫した3D衣服用テクスチャを作る方法。

何に役立つ?

考えられる用途は3D衣服データの制作と、後から照明を変える表現。要旨の実証は比較実験でのテクスチャ品質と見た目の改善。

この研究の面白いところ

縫製パターン全体にまず粗いテクスチャを置き、3D位置情報を使う拡散モデルでUV空間内から洗練する。影が焼き込まれる問題にも取り組む。

どこまで分かった?

要旨は比較手法を大きく上回ると述べるが、具体的な評価値や、物理シミュレーションでの使用結果は示していない。

v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。

アブストラクトの日本語訳

一枚の画像から制作に使える3D衣服データを自動生成することは、デジタルコンテンツ制作の中心的な課題である。近年の生成モデルは3D形状の再構成を大きく進めたが、高品質なテクスチャの合成はなお障害となっている。既存の方法は、環境照明や影をテクスチャマップに焼き込むことが多く、また全体の構造的一貫性を保てない場合がある。その結果、得られたデータを物理シミュレーションや再照明に利用しにくい。本研究は、2次元の縫製パターン空間内で、全体として一貫したテクスチャマップを直接合成するOmniFabricを導入する。一枚の参照画像から、推定した3Dメッシュと高性能な視覚言語モデルの生成的な事前知識を使い、展開した縫製パターン全体に、粗いながらも欠けのない初期テクスチャを作る。次に、自動生成された合成データで学習し、3D位置特徴を条件とする専用の拡散トランスフォーマーによって、標準化されたUV領域でこの初期値を直接洗練する。これにより、ゆがみや焼き込まれた不要な模様を取り除き、元の衣服デザインを保つ、きれいで正規化されたテクスチャマップを抽出する。広範な実験では、OmniFabricは最先端の比較手法を大きく上回り、高品質なテクスチャを備えた写実的な3D衣服を生成した。

v1の要旨から自動生成。本文の精読・人による確認は未実施。

初稿
2026-09-24(UTC)
最新改訂
2026-09-24 · v1
査読・掲載
査読状況未確認
arXivで読むPDF

更新履歴

取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。

原文の要旨

Automated generation of production-ready 3D garment assets from a single image is a central challenge in digital content creation. While recent generative models have significantly advanced 3D geometry reconstruction, synthesizing high-quality textures remains a bottleneck. Existing methods often bake environmental illumination and shadows directly into the texture map, or they fail to maintain global structural coherence, making the resulting assets unusable for physical simulation and relighting. In this work, we introduce OmniFabric, a novel approach that synthesizes globally coherent texture maps directly within the 2D sewing pattern space. Given a single reference image, our pipeline utilizes an estimated 3D mesh and generative priors of powerful Vision-Language Models (VLM) to establish a complete but coarse texture initialization across the unwrapped sewing patterns. We then leverage a specialized diffusion transformer, trained via an automated synthetic data engine and conditioned on 3D positional features, to refine this initialization directly in the canonical UV domain. This effectively removes distortion and baked-in artifacts to extract a clean and normalized texture map that preserves the original garment design. Extensive experiments demonstrate that OmniFabric significantly outperforms state-of-the-art baselines, yielding photorealistic 3D garments with high-quality textures.

著者のコメント

Accepted to SIGGRAPH Asia 2026. Project Page: https://humansensinglab.github.io/OmniFabric/

arXiv ID: 2609.30234 / 要約の誤りについて