混合行动空间交通信号控制强化学习 (Reinforcement learning for traffic signal control in hybrid action space) - 专知论文

会员服务 ·

0

控制器 · Learning · 可约的 · Continuity · 离散化 ·

2022 年 11 月 25 日

Reinforcement learning for traffic signal control in hybrid action space

翻译：混合行动空间交通信号控制强化学习

Haoqing Luo,sheng jin

from arxiv, There are serious problems with the innovation of the paper

The prevailing reinforcement-learning-based traffic signal control methods are typically staging-optimizable or duration-optimizable, depending on the action spaces. In this paper, we propose a novel control architecture, TBO, which is based on hybrid proximal policy optimization. To the best of our knowledge, TBO is the first RL-based algorithm to implement synchronous optimization of the staging and duration. Compared to discrete and continuous action spaces, hybrid action space is a merged search space, in which TBO better implements the trade-off between frequent switching and unsaturated release. Experiments are given to demonstrate that TBO reduces the queue length and delay by 13.78% and 14.08% on average, respectively, compared to the existing baselines. Furthermore, we calculate the Gini coefficients of the right-of-way to indicate TBO does not harm fairness while improving efficiency.

翻译：现有的基于强化学习的交通信号控制方法通常视行动空间而定,可以中转或延长时间限制。在本文中,我们提出一个新的控制结构TBO,它以混合近似政策优化为基础。据我们所知,TBO是第一个基于RL的算法,可以同步优化中转和持续操作空间。与离散和连续操作空间相比,混合行动空间是一个合并的搜索空间,TBO可以更好地在频繁切换和不饱和释放之间实现平衡。我们进行了实验,以证明TBO与现有基线相比,平均将排队长度和延迟分别减少13.78%和14.08%。此外,我们计算了路权基尼系数,以表明TBO在提高效率的同时不会损害公平。

0

相关内容

控制器

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

专知会员服务

50+阅读 · 2019年10月17日

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

专知会员服务

36+阅读 · 2019年10月17日

Stabilizing Transformers for Reinforcement Learning

Stabilizing Transformers for Reinforcement Learning

专知会员服务

60+阅读 · 2019年10月17日

强化学习最新教程，17页pdf

强化学习最新教程，17页pdf

专知会员服务

182+阅读 · 2019年10月11日

【人工智能在2019：一年回顾】反人工智能，AI in 2019: A Year in Review

【人工智能在2019：一年回顾】反人工智能，AI in 2019: A Year in Review

专知会员服务

79+阅读 · 2019年10月10日

VCIP 2022 Call for Special Session Proposals

VCIP 2022 Call for Special Session Proposals

CCF多媒体专委会

1+阅读 · 2022年4月1日

AIART 2022 Call for Papers

AIART 2022 Call for Papers

CCF多媒体专委会

1+阅读 · 2022年2月13日

Transferring Knowledge across Learning Processes

Transferring Knowledge across Learning Processes

CreateAMind

29+阅读 · 2019年5月18日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

43+阅读 · 2019年1月3日

Hierarchical Imitation - Reinforcement Learning

Hierarchical Imitation - Reinforcement Learning

CreateAMind

19+阅读 · 2018年5月25日

Vaspin在胰岛β细胞炎症、胰岛素抵抗及氧化应激中的作用及机制研究

国家自然科学基金

0+阅读 · 2014年12月31日

SCI后铁超载及其致Ferroptosis在白质继发损伤中的作用研究

国家自然科学基金

0+阅读 · 2014年12月31日

Caspr在维系脑微血管内皮细胞屏障功能中的作用和机制

国家自然科学基金

0+阅读 · 2011年12月31日

受干扰信号的自适应滤波及信号检测

国家自然科学基金

0+阅读 · 2011年12月31日

蛋白磷酸化酶(Calcineurin)在钙调节蛋白及钙离子诱导下的构象调控机理研究

国家自然科学基金

0+阅读 · 2009年12月31日

Single-Trajectory Distributionally Robust Reinforcement Learning

Arxiv

0+阅读 · 2023年1月27日

Discriminative Experience Replay for Efficient Multi-agent Reinforcement Learning

Arxiv

0+阅读 · 2023年1月25日

An Incremental Inverse Reinforcement Learning Approach for Motion Planning with Human Preferences

Arxiv

0+阅读 · 2023年1月25日

A Survey on Transformers in Reinforcement Learning

Arxiv

31+阅读 · 2023年1月8日

A Survey on Deep Reinforcement Learning for Data Processing and Analytics

Arxiv

24+阅读 · 2022年2月4日

VIP会员

文章信息

相关主题

相关VIP内容

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

Auto-Sizing the Transformer Network: Improving Speed, Efficiency, and Performance for Low-Resource Machine Translation

专知会员服务

50+阅读 · 2019年10月17日

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

Connections between Support Vector Machines, Wasserstein distance and gradient-penalty GANs

专知会员服务

36+阅读 · 2019年10月17日

Stabilizing Transformers for Reinforcement Learning

Stabilizing Transformers for Reinforcement Learning

专知会员服务

60+阅读 · 2019年10月17日

强化学习最新教程，17页pdf

强化学习最新教程，17页pdf

专知会员服务

182+阅读 · 2019年10月11日

【人工智能在2019：一年回顾】反人工智能，AI in 2019: A Year in Review

【人工智能在2019：一年回顾】反人工智能，AI in 2019: A Year in Review

专知会员服务

79+阅读 · 2019年10月10日

热门VIP内容

开通专知VIP会员享更多权益服务

大模型推理时代的知识编辑

《利用人工智能对军事行动进行建模》

【MIT博士论文】加速科学发现的因果建模实践算法

机器人、无人机与实时影像：应对城市爆炸威胁的三大技术方案

相关资讯

VCIP 2022 Call for Special Session Proposals

VCIP 2022 Call for Special Session Proposals

CCF多媒体专委会

1+阅读 · 2022年4月1日

AIART 2022 Call for Papers

AIART 2022 Call for Papers

CCF多媒体专委会

1+阅读 · 2022年2月13日

Transferring Knowledge across Learning Processes

Transferring Knowledge across Learning Processes

CreateAMind

29+阅读 · 2019年5月18日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

43+阅读 · 2019年1月3日

Hierarchical Imitation - Reinforcement Learning

Hierarchical Imitation - Reinforcement Learning

CreateAMind

19+阅读 · 2018年5月25日

相关论文

Single-Trajectory Distributionally Robust Reinforcement Learning

Arxiv

0+阅读 · 2023年1月27日

Discriminative Experience Replay for Efficient Multi-agent Reinforcement Learning

Arxiv

0+阅读 · 2023年1月25日

An Incremental Inverse Reinforcement Learning Approach for Motion Planning with Human Preferences

Arxiv

0+阅读 · 2023年1月25日

A Survey on Transformers in Reinforcement Learning

Arxiv

31+阅读 · 2023年1月8日

A Survey on Deep Reinforcement Learning for Data Processing and Analytics

Arxiv

24+阅读 · 2022年2月4日

相关基金

Vaspin在胰岛β细胞炎症、胰岛素抵抗及氧化应激中的作用及机制研究

国家自然科学基金

0+阅读 · 2014年12月31日

SCI后铁超载及其致Ferroptosis在白质继发损伤中的作用研究

国家自然科学基金

0+阅读 · 2014年12月31日

Caspr在维系脑微血管内皮细胞屏障功能中的作用和机制

国家自然科学基金

0+阅读 · 2011年12月31日

受干扰信号的自适应滤波及信号检测

国家自然科学基金

0+阅读 · 2011年12月31日

蛋白磷酸化酶(Calcineurin)在钙调节蛋白及钙离子诱导下的构象调控机理研究

国家自然科学基金

0+阅读 · 2009年12月31日

微信扫码咨询专知VIP会员