arXiv論文メモ
新着一覧
cs.AI / cs.CL · 査読状況未確認

通信できない環境で航空業務の文書を参照するAI

Offline Multimodal Large Language Models for Decision Support in Air Operations

Joao P. A. Dantas, Jelton A. Cunha, Gabriel Dietzsch

この論文をやさしく読む

ひとことで言うと

ネットにつながらない環境で、文書と画像を検索して原典付きの回答を出すAIを検討した研究です。空軍の分析担当者4人の予備研究を報告しています。

何に役立つ?

外部接続が使えない組織で、技術文書の知識に自然言語でアクセスする設計の参考になります。今回の試験は、将来の支援付き作業と比較するための基準を作っています。

この研究の面白いところ

知識試験の得点・時間と、AIなしで報告書を作る作業負荷を分けて測っています。同じ8点でも試験時間はシステム7.1分、人の平均26.5分でした。

どこまで分かった?

担当者4人の予備研究で、AI支援による報告書作成の効果比較は今後の評価です。知識試験の短時間化を、そのまま実作業の時間短縮や負荷軽減の実証と読むことはできません。

v1のアブストラクトに基づくAI解説。日本語訳とは別に、用途の解釈を含みます。

アブストラクトの日本語訳

航空作戦は、接続性が限られ厳格なセキュリティ制約がある状況で、複雑な規則、確立された手順、時間的制約の厳しい分析に依存する。このような環境では、分析担当者は外部の計算資源にアクセスできないことも多い中、文書化された教範と画像を組み合わせなければならない。本論文は、隔離・制限された環境に配置され、自然言語でのやり取りを通して、原典まで追跡できる教範知識を分析担当者に提供する意思決定支援ツールとして、オフラインの大規模言語モデルを研究する。インターネット接続なしで動作するのに適したモジュール式の検索拡張アーキテクチャを記述し、技術マニュアルのテキストと画像の両方の入力に対応する。この構成の評価に向けた第一歩として、ブラジル空軍の画像分析担当者4人による予備研究を報告する。研究は、(i)電子目標の識別教範に基づく知識評価で、人と提案システムを同じ試験で比較することと、(ii)AI支援なしで偵察目標報告書(Relatório de Missão de Reconhecimento、REMIR)を手作業で作成する際の認知的作業負荷の測定を組み合わせる。結果は、手作業が特に精神的要求(7点中6.0)と努力(7点中5.0)の面で負荷の高い仕事であることを示した。一方、提案システムは人と同じ得点(10点中8点)を獲得し、評価を7.1分で終えた。人の平均は26.5分だった。これにより、将来のAI支援評価の基準が得られる。最後に、手作業とAI支援のワークフローを体系的に比較するための今後の評価手順を記述する。

v1の要旨から自動生成。本文の精読・人による確認は未実施。

初稿
2026-09-18(UTC)
最新改訂
2026-09-18 · v1
査読・掲載
査読状況未確認
arXivで読むPDF

更新履歴

取得できた版を表示。版の更新は査読済みを意味しません。過去版の本文差分は未解析です。

原文の要旨

Air operations rely on complex rules, established procedures, and time-critical analysis under limited connectivity and strict security constraints. In such environments, analysts must combine written doctrine with images, often without access to external computing resources. This paper studies offline large language models as decision support tools, deployed in isolated and restricted environments to give analysts access to doctrinal knowledge that remains traceable to its original sources through natural language interaction. We describe a modular retrieval-augmented architecture suitable for operation without Internet connectivity, supporting both text and image input from technical manuals. As a first step toward evaluating this architecture, we report a pilot study with four image analysts of the Brazilian Air Force, combining (i) a doctrinal knowledge assessment based on their electronic-target identification doctrine, comparing human and proposed system performance on the same test, and (ii) a measurement of the cognitive workload involved in manually producing a reconnaissance target report (Relatório de Missão de Reconhecimento - REMIR) without AI assistance. The results show a demanding manual task, especially in terms of mental demand (6.0/7) and effort (5.0/7), while the proposed system matches the human score (8/10) and completes the assessment in 7.1 minutes (compared to a human average of 26.5 minutes), establishing a baseline for future AI-assisted evaluation. Finally, we describe a future evaluation protocol to systematically compare manual and AI-assisted workflows.

arXiv ID: 2609.21390 / 要約の誤りについて