一部の場面しか見えない環境で行動モデルを学ぶ
Neurosymbolic Action Model Learning under Partial Observability
この論文をやさしく読む
ひとことで言うと
画像や操作履歴が欠けていても、行動が可能になる条件とその結果を学ぶ方法。
何に役立つ?
観測が不完全な画面操作や視覚的な計画で、行動モデルを人手で作る負担を減らす方法の検討に役立つ。
この研究の面白いところ
既存手法が完全な画像履歴を仮定する問題を、変分的な理論分析と六領域・三条件の実験の両方で扱った。
どこまで分かった?
要旨は真の行動モデルの関連する部分の復元を報告しており、すべての状態や行動の完全な復元を示したとは述べていない。
v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。
アブストラクトの日本語訳
AIの計画問題では、エージェントが行動を順に実行して目標へ到達する方法を考える。正しく計画するには、各行動をいつ実行でき、世界をどう変えるかを記す行動モデルが必要である。これを人手で作るには分野の専門知識が必要で、費用もかかり誤りも生じやすい。既存のニューロシンボリック手法ならデータから行動モデルを学べるが、画像によって状態を完全に観測できる、欠けのない行動履歴を利用できると仮定する。画像の一部が存在しないか現在の状態を十分に伝えない部分観測下では、既存手法はモデルを学べない。そこで本論文は、部分観測下で行動モデルを学ぶ新たなニューロシンボリックの枠組みNeSyAMを提案する。さらに、提案手法と比べた既存手法の制限を理論的に分析する統一的な変分枠組みを示す。六つの視覚的計画領域と三つの観測条件で幅広く試験した結果、NeSyAMは部分観測下でも真の行動モデルの関連する部分を一貫して復元した。
v1の要旨から自動生成。本文の精読・人による確認は未実施。
- 初稿
- 2026-09-22(UTC)
- 最新改訂
- 2026-09-22 · v1
- 査読・掲載
- 査読状況未確認
更新履歴
- v1 2026-09-22 この版を読む
取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。
原文の要旨
AI planning studies how an agent can reach a goal by executing a sequence of actions. To plan correctly, the agent needs an action model describing when each action can be executed and how it changes the world. Constructing such models by hand requires domain expertise, and can be costly and error-prone. Action models can instead be learned from available data using existing neurosymbolic approaches, but they currently assume access to complete traces of fully observable images . These approaches fail to learn action models under partial observability where some of the images might not be present or are not fully informative of the current state of the world. Hence, this paper proposes NeSyAM, a novel neurosymbolic modeling paradigm for action model learning under partial observability. In addition, the paper presents a unified variational framework for theoretically analysing the limitations of existing methods compared to our proposed approach. NeSyAM is then tested extensively on six visual planning domains and three observation regimes to show it consistently recovers relevant parts of the true action model under partial observability.
arXiv ID: 2609.25766 / 要約の誤りについて