arXiv論文メモ
新着一覧
cs.CV / cs.GR · 査読状況未確認

複数視点の実写動画から飛び散る液体を再構成

SplashSplat: Reconstructing Splashing Liquids from Real-World Multi-View Videos

Peiyu Liu, Dingxi Zhang, Federico Tombari, Marc Pollefeys, Christina Tsalicoglou, Daniel Barath

この論文をやさしく読む

ひとことで言うと

短時間で形状が変わる水しぶきを、7台の同期多視点カメラ映像から再構成するデータセットとSplashSplatを提案した研究です。

何に役立つ?

飛沫や液滴の3次元表示、時間補間、スタイル転送など、動きの速い液体を実写から扱う評価に役立つ可能性があります。

この研究の面白いところ

20実シーンを4K・毎秒60フレームで撮影し、SDF、レベルセット輸送、ラグランジュ的担体、局所ガウスを組み合わせています。既存手法より物理的に妥当な運動と低い学習コストを報告しました。

どこまで分かった?

データセットは20シーン、7台のカメラという条件です。あらゆる液体や撮影条件での性能、物理的妥当性の定量指標は要旨にありません。

v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。

アブストラクトの日本語訳

水しぶきの寿命は1秒にも満たない。液膜は裂けて細い液糸や液滴になり、見え方は視点に依存してほとんど模様を持たず、追跡できるほど長く残るものは少ない。そのため、再構成の研究は煙、合成液体、または穏やかに変形する表面に集中してきた。著者らの知る限り、飛び散る液体を同期撮影した多視点データセットは存在しない。 そこで、まとまった流れから激しい飛沫まで20の実シーンを、同期・較正済みの7台の4Kカメラで毎秒60フレームで撮影したベンチマークを導入する。各視点の液体と容器のマスクを手作業で修正し、固定した評価用の分割を提供する。さらに、「観測によって制約できる箇所にだけ物理構造を課す」という一つの原則に基づくSplashSplatを提案する。マスクを融合して得たフレームごとの液体の符号付き距離関数(SDF)が形状を与え、連続するSDF間のレベルセット輸送から粗い速度場を求める。この流れに沿って移流するラグランジュ的な担体を、新たな観測ごとに補正し、被覆が失われた箇所では再配置する。担体から局所的なガウス分布を復号し、微分可能なレンダリングに用いる。 SplashSplatは、実写データと合成ベンチマークの両方で、最先端の動的ガウシアンスプラッティング手法を上回り、より物理的に妥当な運動と低い学習コストを実現する。同じ表現を用いて、再最適化せずに時間補間とスタイル転送も行える。

v1の要旨から自動生成。本文の精読・人による確認は未実施。

初稿
2026-09-17(UTC)
最新改訂
2026-09-17 · v1
査読・掲載
査読状況未確認
arXivで読むPDF

更新履歴

取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。

原文の要旨

A splash lives for a fraction of a second: sheets tear into ligaments and droplets, appearance is view-dependent and nearly textureless, and little persists long enough to track. Reconstruction research has consequently focused on smoke, synthetic liquids, or gently deforming surfaces. To our knowledge, no synchronized multi-view dataset of splashing liquids exists. We therefore introduce a benchmark of 20 real scenes, from coherent streams to violent splashes, captured by seven synchronized, calibrated 4K cameras at 60 fps, with manually refined per-view liquid and container masks and fixed evaluation splits. We further present SplashSplat, built on a single principle: impose physical structure only where the observations can constrain it. Per-frame liquid SDFs fused from the masks provide the geometry, level-set transport between consecutive SDFs yields a coarse velocity field, and Lagrangian carriers advected along this flow, corrected against each new observation and reseeded where coverage is lost, decode local Gaussians for differentiable rendering. SplashSplat outperforms state-of-the-art dynamic Gaussian splatting methods on our real captures and on a synthetic benchmark, with physically more plausible motion and a lower training cost. The same representation supports temporal interpolation and style transfer without re-optimization.

著者のコメント

18 pages (11 main + 7 supplementary), 14 figures, 12 tables. Project page: https://niko-creater.github.io/splashsplat-web/

arXiv ID: 2609.20818 / 要約の誤りについて