节点文献

优化加权平均流程时间的平行机调度

Parallel machine scheduling minimizing the mean weighted flow time

  • 推荐 CAJ下载
  • PDF下载
  • 不支持迅雷等下载工具,请取消加速工具后下载。

【作者】 张智聪郑力翁小华

【Author】 Zhang Zhi-cong1,2,Zheng Li2,Michael X.Weng3(1 Dongguan University of Technology,Dongguan 523808,Guangdong,CHN;2 Department of Industrial Engineering,Tsinghua University,Beijing 100084,CHN;3 Department of Industrial & Management Systems Engineering,University of South Florida,Tampa 33620,Florida,USA)

【机构】 广东东莞理工学院机电工程系清华大学工业工程系南佛罗里达大学工业与管理系统工程系 东莞523808北京100084坦帕33620

【摘要】 平行机调度问题在工业界有着广泛应用,实际生产中瓶颈工序的调度很多属于这类问题。运用增强学习算法来研究以最小化作业的加权平均滞留时间为目标的动态平行机调度问题Qm|rj,sjk,Mj|∑wjfj,考虑与作业顺序相关的转换时间和机器-作业资格约束。为了把调度问题转化为增强学习问题,定义了系统状态的表示方式,利用加权最短加工时间优先(WSPT)规则、Weng算法、排名(RA)算法和LFJ-RA(Least Flexible Job-Ranking Algorithm)算法构造行为,定义了与调度目标函数等价的报酬函数,并采用结合函数泛化器的Q学习算法来解决。实验表明Q学习算法对每个测试问题的调度结果都优于WSPT规则、排名算法、LFJ-RA算法和Weng算法。

【Abstract】 Parallel machine scheduling problem is common in industry.A Reinforcement Learning(RL)algorithm,Q-learning,was used to solve unrelated parallel machine scheduling problem Qm|rj,sjk,Mj|∑wjfj.The sequence-dependent conversion times and machine eligibility constraint were considered.To convert the scheduling problem into an RL problem,the problem was formulated as Semi-Markov Decision Process by defining system state,actions and the reward function.Four heuristics,WSPT,Weng’s Algorithm,Ranking Algorithm(RA)and LFJ-RA,were defined as actions.Q-Learning combining linear gradient-descent function approximation was used to minimize the mean weighted flow time.Q-Learning learned to select optimal or sub-optimal actions at different states through simulation.Experiment results show that Q-Learning is superior to the four heuristics in all test problems.

【关键词】 调度平行机Q学习
【Key words】 SchedulingParallel machineQ-Learning
【基金】 国家自然科学基金资助项目(50375082)
  • 【文献出处】 现代制造工程 ,Modern Manufacturing Engineering , 编辑部邮箱 ,2007年09期
  • 【分类号】O223
  • 【被引频次】2
  • 【下载频次】198
节点文献中: 

本文链接的文献网络图示:

本文的引文网络