都市建物のインスタンス分割と細粒度分類
Instance Segmentation and Fine-grained Classification for Urban Buildings with Adaptive Region Dividing and Spatially-Supervised Contrastive Learning
この論文をやさしく読む
ひとことで言うと
都市の大きな3次元点群から建物ごとに切り分け、その機能も細かく分類する方法です。
何に役立つ?
都市モデルや都市分析で、固定ブロックの境界が建物を分断してしまう問題を減らす設計になります。
この研究の面白いところ
鳥瞰図で建物を検出して元の点群へ戻し、構造に沿った領域を作ります。形・色・周囲の情報と、近い同種建物を重視する対照学習を組み合わせます。
どこまで分かった?
UrbanBISとSTPLS3Dで既存手法との比較を行っています。要旨には精度の差や都市全体の処理時間の数値がありません。
v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。
アブストラクトの日本語訳
大規模点群に含まれる都市建物をインスタンス単位で正確に理解し、機能別に細かく分類することは、デジタル都市モデルと都市分析に不可欠である。しかし都市シーンは空間範囲が広いため、既存手法の多くは、訓練と評価にあらかじめ定義したブロックへ依存している。このような分割は現実の応用では利用できないことが多く、追加の前処理を必要とし、完全な建物構造を分断する。 この課題に対し、統一されたシーンレベル評価を可能にする適応的領域分割戦略を提案する。まず3次元点群を鳥瞰図(BEV)平面へ投影し、事前学習済みの分割モデルで建物領域を検出する。検出した境界ボックスを元の点群へ逆投影し、構造に沿った適応的訓練ブロックを構築する。これにより、手動設計なしに意味情報に基づく動的分割を実現する。 さらに、インスタンスレベルの理解を超えて、都市建物の細粒度分類を扱う研究は少ないため、空間教師付き対照学習損失を備えた細粒度分類モデルも提案する。分割された各建物について、点トランスフォーマー分類器が、幾何、色、中心周辺の文脈情報を使って建物本体と局所文脈を共同で符号化する。次に、クラス間の深刻な不均衡を緩和するため、クラス均衡重み付き交差エントロピーを用いる。提案する空間教師付き対照損失は、空間的に近く同じカテゴリーに属する建物へ大きな重みを与え、機能表現をコンパクトにし、混同しやすいカテゴリーを分離することで、クラス間の識別性を高める。 UrbanBISとSTPLS3Dでの広範な実験は、建物インスタンス分割と細粒度分類の両方で、既存の最先端手法に対する提案手法の優位性を示した。
v1の要旨から自動生成。本文の精読・人による確認は未実施。
- 初稿
- 2026-09-17(UTC)
- 最新改訂
- 2026-09-17 · v1
- 査読・掲載
- 査読状況未確認
更新履歴
- v1 2026-09-17 この版を読む
取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。
原文の要旨
Accurate instance-level and functional understanding of urban buildings in large-scale point clouds is essential for digital city modeling and urban analysis. However, the extensive spatial coverage of urban scenes leads most existing methods to rely on predefined blocks for training and evaluation, although such partitions are rarely available in real-world applications and introduce additional preprocessing while fragmenting complete building structures. To address this issue, we propose an adaptive region-dividing strategy with unified scene-level evaluation. Specifically, the 3D point cloud is projected onto a bird's-eye-view (BEV) plane, where a pretrained segmentation model is used to detect building regions. The detected bounding boxes are then back-projected to the original point cloud to construct structure-aligned adaptive training blocks, enabling semantically guided dynamic partitioning without manual design. Furthermore, beyond instance-level understanding, few methods have explored fine-grained classification for urban buildings, and thus we also put forward a fine-grained classification model for urban buildings with a spatially-supervised contrastive loss. First, for each segmented building, a point transformer classifier jointly encodes its body and local context using geometric, color, and core-context information. Then, the class-balanced weighted cross-entropy is used to alleviate severe class imbalance. The proposed spatially-supervised contrastive loss further enhances inter-class discriminability by assigning greater weight to spatially proximate, same-category buildings, encouraging compact functional representations while separating easily confused categories. Extensive experiments on UrbanBIS and STPLS3D demonstrate the advantages of the proposed method in building instance segmentation and fine-grained classification compared to existing SOTA methods.
著者のコメント
10 pages, 4 figures
arXiv ID: 2609.19631 / 要約の誤りについて