To achieve the ability for fulfilling the new task based on the learning form demonstration(LfD),a framework was proposed to represent the policy,imitation learning and optimization for robot intelligent trajectory planning was conducted,based on Gaussian mixture regression and policy improvement with path integral strategy(GMR-PI2).Meanwhile,self-reconstruction of the basis function and search of the weight coefficient of the policy expression were performed in two spaces alternately to conquer above probl...