Bridge Bidding via Deep Reinforcement Learning and Belief Monte Carlo Search

下载PDF

导出

摘要 Contract Bridge,a four-player imperfect information game,comprises two phases:bidding and playing.While computer programs excel at playing,bidding presents a challenging aspect due to the need for information exchange with partners and interference with communication of opponents.In this work,we introduce a Bridge bidding agent that combines supervised learning,deep reinforcement learning via self-play,and a test-time search approach.Our experiments demonstrate that our agent outperforms WBridge5,a highly regarded computer Bridge software that has won multiple world championships,by a performance of 0.98 IMPs(international match points)per deal over 10000 deals,with a much cost-effective approach.The performance significantly surpasses previous state-of-the-art(0.85 IMPs per deal).Note 0.1 IMPs per deal is a significant improvement in Bridge bidding.

作者 Zizhang Qiu Shouguang Wang Dan You MengChu Zhou

机构地区 School of Information and Electronic Engineering IEEE

出处《IEEE/CAA Journal of Automatica Sinica》 SCIE EI CSCD 2024年第10期2111-2122,共12页 自动化学报（英文版）

关键词 Contract Bridge reinforcement learning SEARCH

分类号 TP18 [自动化与计算机技术—控制理论与控制工程]

IEEE/CAA Journal of Automatica Sinica

2024年第10期

浏览历史

内容加载中请稍等...

Bridge Bidding via Deep Reinforcement Learning and Belief Monte Carlo Search

相关作者

相关机构

相关主题

浏览历史