arXiv論文メモ
新着一覧
cs.CL / cs.AI · 査読状況未確認

宇宙力学の問題を解く言語モデルに検索と推論を組み込む

Taramandal-GPT: Enhancing Astrodynamics Problem-Solving with Knowledge Retrieval and Structured Thinking

Akhil Sharma, Jatin Gupta and Ali Imam Abidi

この論文をやさしく読む

ひとことで言うと

宇宙力学の専門知識を検索して使う仕組みを言語モデルに加え、計算や段階的な推論が必要な問題に答えやすくする研究です。

何に役立つ?

考えられる用途は、宇宙科学や宇宙機工学の問題を検討する際の補助です。今回評価したのは299問のベンチマークへの回答です。

この研究の面白いところ

数値が許容範囲に入るかと、意味が近いかの両面から回答を評価しています。検索に加えてフォールバック機構を備える構成です。

どこまで分かった?

要旨は競争力のある性能と述べますが、具体的な得点や誤答率、各機構の寄与は示していません。実際の宇宙機運用での信頼性を実証した結果ではありません。

v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。

アブストラクトの日本語訳

大規模言語モデル(LLM)は自然言語理解で著しい進歩を示しているが、天文学や宇宙力学などの専門分野では、多段階推論、記号操作、分野固有の用語の課題により、有効性が依然として限られている。これに対処するため、Qwen3-8bを基盤とし、検索拡張生成(RAG)のパイプラインと、文脈の精度を高めるフォールバック機構を備えた分野適応型の枠組みTaramandal-GPT(Constellation-GPT)を提示する。 宇宙科学の基礎から高度な水準までを扱う299問のデータセット、Astrodynamics Problems Benchmark(APBench)で評価する。数値の許容誤差に基づく採点と意味的類似性評価という二重の評価法を使い、Taramandal-GPTは最先端のオープンソースおよびクローズドソースのモデルに対して競争力のある性能を達成し、とくに集中的な思考を要する課題で強みを示す。これらの結果は、正確さと解釈可能性を要求する分野における専門化したLLMの価値を示し、Taramandal-GPTを、天体物理学、宇宙機工学、宇宙探査のための信頼できる人工知能(AI)アシスタントに向けた一歩として位置付ける。

v1の要旨から自動生成。本文の精読・人による確認は未実施。

初稿
2026-09-21(UTC)
最新改訂
2026-09-21 · v1
査読・掲載
査読状況未確認
arXivで読むPDF

更新履歴

取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。

原文の要旨

Large language models (LLMs) have shown remarkable progress in natural language understanding, yet their effectiveness in specialized fields like astronomy and astrodynamics remains limited due to challenges in multi-step reasoning, symbolic manipulation, and domain-specific terminology. To address this, we present Taramandal-GPT (Constellation-GPT), a domain-adapted framework built on the Qwen3-8b backbone, enhanced with a Retrieval-Augmented Generation (RAG) pipeline and a fallback mechanism for improved contextual precision. We evaluate it on the Astrodynamics Problems Benchmark (APBench), a dataset of 299 questions covering foundational to advanced levels of space science. Using a dual evaluation method - numeric margin-based scoring and semantic similarity assessment - Taramandal-GPT achieves competitive performance against state-of-the-art open- and closed-source models, with notable strength in thinking-intensive tasks. These results highlight the value of specialized LLMs for domains demanding accuracy and interpretability, positioning Taramandal-GPT as a step toward reliable Artificial Intelligence (AI) assistants for astrophysics, spacecraft engineering, and space exploration.

著者のコメント

Proceedings of All India Hindi Technical Conference, 05-06 February 2026

arXiv ID: 2609.24246 / 要約の誤りについて