引用本文:张雁冰,杭大明,马正新,曹志刚.基于再励学习的主动队列管理算法.软件学报,2004,15(7):1090-1098
【打印本页】   【下载PDF全文】   查看/发表评论  【EndNote】   【RefMan】   【BibTex】
←前一篇|后一篇→ 过刊浏览    高级检索
本文已被:浏览 5263次   下载 6832 本文二维码信息
码上扫一扫!
分享到: 微信 更多
基于再励学习的主动队列管理算法
张雁冰1, 杭大明1, 马正新1, 曹志刚1
清华大学,电子工程系,微波与数字通信国家重点实验室,北京,100084
摘要:
从最优决策的角度出发,将人工智能中的再励学习方法引入主动队列管理的研究中,提出了一种基于再励学习的主动队列管理算法RLGD(reinforcement learning gradient-descent).RLGD以速率匹配和队列稳定为优化目标,根据网络状态自适应地调节更新步长,使得队列长度能够很快收敛到目标值,并且抖动很小.此外,RLGD不需要知道源端的速率调整算法,因而具有很好的可扩展性.通过不同网络环境下的仿真显示,RLGD与REM,PI等AQM算法相比,具有更好的性能和鲁棒性.
关键词:  拥塞控制  主动队列管理  再励学习
DOI:
分类号:
基金项目:Supported bythe National High-Tech Research and Development Plan of Chinaunder Grant No.2001AA121062(国家高技术研究发展计划(863))
A Robust Active Queue Management Algorithm Based on Reinforcement Learning
ZHANG Yan-Bing,HANG Da-Ming,MA Zheng-Xin,CAO Zhi-Gang
Abstract:
From the viewpoint of decision theory, AQM (active queue management) can be considered as an optimal decision problem. In this paper, a new AQM scheme, Reinforcement Learning Gradient-Descent (RLGD), is described based on the optimal decision theory of reinforcement learning. Aiming to maximize the throughput and stabilize the queue length, RLGD adjusts the update step adaptively, without the demand of knowing the rate adjustment scheme of the source sender. Simulation demonstrates that RLGD can lead to the convergence of the queue length to the desired value quickly and maintain the oscillation small. The results also show that the RLGD scheme is very robust to disturbance under various network conditions and outperforms the traditional REM and PI controllers significantly.
Key words:  congestion control  active queue management  reinforcement learning

引用本文:
【打印本页】   【下载PDF全文】   查看/发表评论  【EndNote】   【RefMan】   【BibTex】
←前一篇|后一篇→ 过刊浏览    高级检索
本文已被:浏览次   下载  
分享到: 微信 更多
摘要:
关键词:  
DOI:
分类号:
基金项目:
Abstract:
Key words: