Skip to main content
QUICK REVIEW

[論文レビュー] Explicability? Legibility? Predictability? Transparency? Privacy? Security? The Emerging Landscape of Interpretable Agent Behavior

Tathagata Chakraborti, Anagha Kulkarni|arXiv (Cornell University)|Nov 23, 2018
Logic, Reasoning, and Knowledge被引用数 50
ひとこと要約

本論文は、エージェントの振る舞いに対する解釈可能性の概念(explicability、legibility、predictability、transparency)の分類体系を整理し、協調と対抗の設定の両方におけるプライバシー/セキュリティへと拡張し、観察者モデルが計画の解釈に与える影響を明らかにする。

ABSTRACT

There has been significant interest of late in generating behavior of agents that is interpretable to the human (observer) in the loop. However, the work in this area has typically lacked coherence on the topic, with proposed solutions for "explicable", "legible", "predictable" and "transparent" planning with overlapping, and sometimes conflicting, semantics all aimed at some notion of understanding what intentions the observer will ascribe to an agent by observing its behavior. This is also true for the recent works on "security" and "privacy" of plans which are also trying to answer the same question, but from the opposite point of view -- i.e. when the agent is trying to hide instead of revealing its intentions. This paper attempts to provide a workable taxonomy of relevant concepts in this exciting and emerging field of inquiry.

研究の動機と目的

  • エージェントの振る舞いにおける explicability、legibility、predictability、transparency の整合的な定義と関係を明確にする。
  • 解釈可能な planning における協調的設定と対抗的設定を区別する。
  • 観察者モデルと計算制約が計画の解釈に与える影響を説明する。
  • オンライン対オフラインの相互作用と、それらが解釈可能性指標に与える影響を強調する。

提案手法

  • 計画問題、計画、計算モデル、および観測モデルを含む、エージェントと観察者 ( ϕPi^A, Pi^Theta) をモデル化する一般的な枠組みを提示する。
  • この枠組みの中で explicability、predictability、legibility、transparency を定義し、区別する。
  • モーション計画ドメインとタスク計画ドメインを論じ、観察者の計算能力の役割を論じる。
  • 関連研究を要約し、協調的設定と対抗的設定に跨る概念の表を提供する。
  • 観察者モデルの学習や長期的相互作用を含む、未解決の課題と潜在的拡張について論じる。

実験結果

リサーチクエスチョン

  • RQ1計画における explicability、legibility、predictability、transparency の正確な定義とそれらの関係は何か。
  • RQ2観察者モデルと計算制約は、エージェントが解釈可能な振る舞いを産み出す能力をどのように形作るか。
  • RQ3協調的設定と対抗的設定は、解釈可能または隠蔽された計画を達成する目標と方法をどのように変えるか。
  • RQ4オンライン対オフラインの解釈可能性と観察者モデルの学習における主なギャップと今後の方向性は何か。

主な発見

  • explicability、legibility、predictability、transparency を観察者モデルと目標/計画完了に結び付ける、一貫した分類法が提案されている。
  • explicability と predictability は単調ではなく、オンラインとオフライン設定で計画の前方部(prefix)または後方部(suffix)に依存することがある。
  • Legibility または Transparency はゴール推定を扱い、観察者の可能なゴール間の曖昧性を最小化することを要求する。
  • フレームワークはモーション計画とタスク計画を区別し、観察者の計算能力が解釈可能性指標にどう影響するかを論じる。
  • 対抗的設定(プライバシー、偽装、欺瞞、セキュリティ)についての議論があり、それらが協調的な解釈可能性の概念とどう関連するかを扱う。

より良い研究を、今すぐ始めましょう

論文の読解から最終レビューまで、研究時間を劇的に削減しましょう。

クレジットカード登録不要

このレビューはAIが作成し、人間の編集者が確認しました。