This paper considers the problem of co-synthesis in $k$-player games over a finite graph where each player has an individual $\omega$-regular specification $\phi_i$. In this context, a secure equilibrium (SE) is a Nash equilibrium w.r.t. the lexicographically ordered objectives of each player to first satisfy their own specification, and second, to falsify other players' specifications. A winning secure equilibrium (WSE) is an SE strategy profile $(\pi_i)_{i\in[1;k]}$ that ensures the specification $\phi:=\bigwedge_{i\in[1;k]}\phi_i$ if no player deviates from their strategy $\pi_i$. Distributed implementations generated from a WSE make components act rationally by ensuring that a deviation from the WSE strategy profile is immediately punished by a retaliating strategy that makes the involved players lose. In this paper, we move from deviation punishment in WSE-based implementations to a distributed, assume-guarantee based realization of WSE. This shift is obtained by generalizing WSE from strategy profiles to specification profiles $(\varphi_i)_{i\in[1;k]}$ with $\bigwedge_{i\in[1;k]}\varphi_i = \phi$, which we call most general winning secure equilibria (GWSE). Such GWSE have the property that each player can individually pick a strategy $\pi_i$ winning for $\varphi_i$ (against all other players) and all resulting strategy profiles $(\pi_i)_{i\in[1;k]}$ are guaranteed to be a WSE. The obtained flexibility in players' strategy choices can be utilized for robustness and adaptability of local implementations. Concretely, our contribution is three-fold: (1) we formalize GWSE for $k$-player games over finite graphs, where each player has an $\omega$-regular specification $\phi_i$; (2) we devise an iterative semi-algorithm for GWSE synthesis in such games, and (3) obtain an exponential-time algorithm for GWSE synthesis with parity specifications $\phi_i$.


翻译:本文研究有限图上$k$玩家博弈中的协同综合问题,其中每个玩家具有独立的$\omega$-正则规范$\phi_i$。在此背景下,安全均衡(SE)是指关于每个玩家词典序目标(首要满足自身规范,其次破坏其他玩家的规范)的纳什均衡。获胜安全均衡(WSE)是一种SE策略组合$(\pi_i)_{i\in[1;k]}$,使得当无玩家偏离其策略$\pi_i$时,能确保规范$\phi:=\bigwedge_{i\in[1;k]}\phi_i$成立。基于WSE生成的分布式实现通过确保偏离WSE策略组合的行为会立即遭到报复策略的惩罚(使相关玩家失败),从而促使组件理性运作。本文从基于WSE实现的偏离惩罚转向基于分布式假设-保证的WSE实现。这一转变通过将WSE从策略组合推广为规范组合$(\varphi_i)_{i\in[1;k]}$(满足$\bigwedge_{i\in[1;k]}\varphi_i = \phi$)来实现,我们称之为最大通用获胜安全均衡(GWSE)。此类GWSE具有如下性质:每个玩家可独立选择对$\varphi_i$获胜(对抗所有其他玩家)的策略$\pi_i$,且由此产生的所有策略组合$(\pi_i)_{i\in[1;k]}$必然构成WSE。策略选择空间的可获得灵活性可用于增强局部实现的鲁棒性与自适应性。具体而言,本文贡献有三点:(1) 形式化定义了有限图上$k$玩家博弈(每个玩家具有$\omega$-正则规范$\phi_i$)中的GWSE;(2) 设计了此类博弈中GWSE综合的迭代半算法;(3) 针对奇偶性规范$\phi_i$实现了指数时间复杂度的GWSE综合算法。

0
下载
关闭预览

相关内容

FlowQA: Grasping Flow in History for Conversational Machine Comprehension
专知会员服务
35+阅读 · 2019年10月18日
Stabilizing Transformers for Reinforcement Learning
专知会员服务
61+阅读 · 2019年10月17日
《DeepGCNs: Making GCNs Go as Deep as CNNs》
专知会员服务
32+阅读 · 2019年10月17日
Keras François Chollet 《Deep Learning with Python 》, 386页pdf
专知会员服务
164+阅读 · 2019年10月12日
【SIGGRAPH2019】TensorFlow 2.0深度学习计算机图形学应用
专知会员服务
41+阅读 · 2019年10月9日
Hierarchically Structured Meta-learning
CreateAMind
27+阅读 · 2019年5月22日
Transferring Knowledge across Learning Processes
CreateAMind
29+阅读 · 2019年5月18日
强化学习的Unsupervised Meta-Learning
CreateAMind
18+阅读 · 2019年1月7日
Unsupervised Learning via Meta-Learning
CreateAMind
44+阅读 · 2019年1月3日
meta learning 17年:MAML SNAIL
CreateAMind
11+阅读 · 2019年1月2日
A Technical Overview of AI & ML in 2018 & Trends for 2019
待字闺中
18+阅读 · 2018年12月24日
STRCF for Visual Object Tracking
统计学习与视觉计算组
15+阅读 · 2018年5月29日
Focal Loss for Dense Object Detection
统计学习与视觉计算组
12+阅读 · 2018年3月15日
IJCAI | Cascade Dynamics Modeling with Attention-based RNN
KingsGarden
13+阅读 · 2017年7月16日
From Softmax to Sparsemax-ICML16(1)
KingsGarden
74+阅读 · 2016年11月26日
国家自然科学基金
13+阅读 · 2017年12月31日
国家自然科学基金
2+阅读 · 2015年12月31日
国家自然科学基金
3+阅读 · 2015年12月31日
国家自然科学基金
2+阅读 · 2015年12月31日
国家自然科学基金
0+阅读 · 2014年12月31日
国家自然科学基金
1+阅读 · 2014年12月31日
国家自然科学基金
0+阅读 · 2014年12月31日
国家自然科学基金
0+阅读 · 2014年12月31日
国家自然科学基金
2+阅读 · 2014年12月31日
Arxiv
0+阅读 · 2024年2月27日
Arxiv
12+阅读 · 2023年5月22日
Arxiv
34+阅读 · 2022年12月20日
Arxiv
45+阅读 · 2022年9月19日
The Matrix Calculus You Need For Deep Learning
Arxiv
12+阅读 · 2018年7月2日
VIP会员
最新内容
面向2027年及未来的海军情报改革
专知会员服务
3+阅读 · 8月5日
《无人机蜂群:释放人类-蜂群编队的潜能》
专知会员服务
6+阅读 · 8月5日
《战略战术化:一项综合性述评》
专知会员服务
5+阅读 · 8月5日
相关资讯
Hierarchically Structured Meta-learning
CreateAMind
27+阅读 · 2019年5月22日
Transferring Knowledge across Learning Processes
CreateAMind
29+阅读 · 2019年5月18日
强化学习的Unsupervised Meta-Learning
CreateAMind
18+阅读 · 2019年1月7日
Unsupervised Learning via Meta-Learning
CreateAMind
44+阅读 · 2019年1月3日
meta learning 17年:MAML SNAIL
CreateAMind
11+阅读 · 2019年1月2日
A Technical Overview of AI & ML in 2018 & Trends for 2019
待字闺中
18+阅读 · 2018年12月24日
STRCF for Visual Object Tracking
统计学习与视觉计算组
15+阅读 · 2018年5月29日
Focal Loss for Dense Object Detection
统计学习与视觉计算组
12+阅读 · 2018年3月15日
IJCAI | Cascade Dynamics Modeling with Attention-based RNN
KingsGarden
13+阅读 · 2017年7月16日
From Softmax to Sparsemax-ICML16(1)
KingsGarden
74+阅读 · 2016年11月26日
相关论文
Arxiv
0+阅读 · 2024年2月27日
Arxiv
12+阅读 · 2023年5月22日
Arxiv
34+阅读 · 2022年12月20日
Arxiv
45+阅读 · 2022年9月19日
The Matrix Calculus You Need For Deep Learning
Arxiv
12+阅读 · 2018年7月2日
相关基金
国家自然科学基金
13+阅读 · 2017年12月31日
国家自然科学基金
2+阅读 · 2015年12月31日
国家自然科学基金
3+阅读 · 2015年12月31日
国家自然科学基金
2+阅读 · 2015年12月31日
国家自然科学基金
0+阅读 · 2014年12月31日
国家自然科学基金
1+阅读 · 2014年12月31日
国家自然科学基金
0+阅读 · 2014年12月31日
国家自然科学基金
0+阅读 · 2014年12月31日
国家自然科学基金
2+阅读 · 2014年12月31日
Top
微信扫码咨询专知VIP会员