Skip to main content
QUICK REVIEW

[论文解读] Decentralized Detection with Signaling

Ashutosh Nayyar, Demosthenis Teneketzis|arXiv (Cornell University)|May 17, 2010
Distributed Sensor Networks and Detection Algorithms参考文献 7被引用 4
一句话总结

本文研究了一个去中心化的序列检测问题,其中两名观察者可以通过其停止决策相互传递信号。与经典的双阈值策略不同,由于信号传递的存在,最优策略需要一种新的参数化表征,这从根本上改变了去中心化环境中最优停止规则的结构。

ABSTRACT

We consider a sequential problem in decentralized detection. Two observers can make repeated noisy observations of a binary hypothesis on the state of the environment. At any time, any of the two observers can stop and send a final message to the other observer or it may continue to take more measurements. After an observer has sent its final message, it stops operating. The other observer is then faced with a different stopping problem. At each time instant, it can decide either to stop and declare a final decision on the hypothesis or take another measurement. At each time, the system incurs an operating cost depending on the number of observers that are active at that time. A terminal cost that measures the accuracy of the final decision is incurred at the end. We show that, unlike in other sequential detection problems, stopping rules characterized by two thresholds on an observer's posterior belief no longer guarantee optimality in this problem. Thus the potential for signaling among observers alters the nature of optimal policies. We obtain a new parametric characterization of optimal policies for this problem.

研究动机与目标

  • 分析一个两名观察者可通过停止决策相互传递信息的序列去中心化检测问题。
  • 研究观察者之间的信号传递如何影响最优停止策略的结构。
  • 证明当观察者通过决策共享信息时,经典双阈值策略不再是最优的。
  • 推导一种新的参数化表征方法,以反映信号传递的影响。
  • 建立价值函数的凹性,并为该问题提供动态规划框架。

提出的方法

  • 构建一个具有信号传递的双观察者序列检测问题,其中每个观察者的停止决策均传递信息。
  • 使用对假设的后验信念作为每个观察者足够的统计量进行建模。
  • 使用动态规划刻画价值函数,证明其在信念状态下的凹性。
  • 基于仿射函数的下确界,推导出最优策略的参数化表示,取代经典的双阈值结构。
  • 通过分析信号传递的影响,证明在信号传递下,价值函数无法被双阈值策略保持不变。
  • 证明最优策略依赖于未来价值函数的凹包的参数化阈值族。

实验结果

研究问题

  • RQ1观察者之间的信号传递如何影响去中心化序列检测中经典双阈值策略的最优性?
  • RQ2基于信号的去中心化检测中的最优停止规则是否可由参数化阈值族表征?
  • RQ3当观察者使用决策传递信号时,价值函数和策略空间会发生何种结构性变化?
  • RQ4当观察者能观察到彼此的停止决策时,经典双阈值策略是否仍是最优的?
  • RQ5如何调整动态规划解法以考虑通过停止动作传递的信息?

主要发现

  • 当观察者可通过停止决策传递信号时,经典双阈值策略在去中心化检测中不再是最优的。
  • 最优策略由仿射函数下确界的参数化阈值族表征,取代了双阈值结构。
  • 在每个时间步,价值函数在后验信念上是凹的,从而可通过支撑超平面实现参数化表示。
  • 信号传递改变了信息结构,即使其他观察者的策略固定,每个观察者的问题也不再是经典序列检测问题。
  • 最优策略依赖于决策与观测的联合历史,信念状态通过融合信号的贝叶斯更新而演化。
  • 通过动态规划,价值函数的凹性得以保持,从而可使用下确界表示进行最优动作选择。

更好的研究,从现在开始

从阅读论文到最终审阅,大幅缩短您的研究时间。

无需绑定信用卡

本解读由 AI 生成,并经人工编辑审核。