Recognizing Environmental Change through Multiplex Reinforcement Learning in Group Robot System.

The high performance and efficiency of multiple unmanned surface vehicles (multi-USV) promote the further civilian and military applications of coordinated USV. As the basis of multiple USVs’ cooperative work, considerable attention has been spent on developing the decentralized formation control of the USV swarm. Formation control of multiple USV belongs to the geometric problems of a multi-robot system. The main challenge is the way to generate and maintain the formation of a multi-robot system. The rapid development of reinforcement learning provides us with a new solution to deal with these problems. In this paper, we introduce a decentralized structure of the multi-USV system and employ reinforcement learning to deal with the formation control of a multi-USV system in a leader–follower topology. Therefore, we propose an asynchronous decentralized formation control scheme based on reinforcement learning for multiple USVs. First, a simplified USV model is established. Simultaneously, the formation shape model is built to provide formation parameters and to describe the physical relationship between USVs. Second, the advantage deep deterministic policy gradient algorithm (ADDPG) is proposed. Third, formation generation policies and formation maintenance policies based on the ADDPG are proposed to form and maintain the given geometry structure of the team of USVs during movement. Moreover, three new reward functions are designed and utilized to promote policy learning. Finally, various experiments are conducted to validate the performance of the proposed formation control scheme. Simulation results and contrast experiments demonstrate the efficiency and stability of the formation control scheme.

Download Full-text

Analysis and solution of a predator–protector–prey multi-robot system by a high-level reinforcement learning architecture and the adaptive systems theory

Robotics and Autonomous Systems ◽

10.1016/j.robot.2010.08.005 ◽

2010 ◽

Vol 58 (12) ◽

pp. 1266-1272 ◽

Cited By ~ 3

Author(s):

José Antonio Martín H. ◽

Javier de Lope ◽

Darío Maravall

Keyword(s):

Reinforcement Learning ◽

Systems Theory ◽

Adaptive Systems ◽

Robot System ◽

High Level ◽

Multi Robot

Download Full-text

Improving the Robustness of Reinforcement Learning for a Multi-Robot System Environment

Advances in Soft Computing - Soft Computing as Transdisciplinary Science and Technology ◽

10.1007/3-540-32391-0_34 ◽

2007 ◽

pp. 263-272 ◽

Cited By ~ 1

Author(s):

Toshiyuki Yasuda ◽

Kazuhiro Ohkura

Keyword(s):

Reinforcement Learning ◽

Robot System ◽

System Environment ◽

Multi Robot

Download Full-text

A Reinforcement Learning Technique with an Adaptive Action Generator for a Multi-Robot System

Multi-Robot Systems, Trends and Development ◽

10.5772/13337 ◽

2011 ◽

Author(s):

Kazuhiro Ohkura ◽

Toshiyuki Yasu

Keyword(s):

Reinforcement Learning ◽

Robot System ◽

Learning Technique ◽

Adaptive Action ◽

Multi Robot

Download Full-text

Sharing Experience for Behavior Generation of Real Swarm Robot Systems Using Deep Reinforcement Learning

Journal of Robotics and Mechatronics ◽

10.20965/jrm.2019.p0520 ◽

2019 ◽

Vol 31 (4) ◽

pp. 520-525 ◽

Cited By ~ 1

Author(s):

Toshiyuki Yasuda ◽

Kazuhiro Ohkura ◽

◽

Keyword(s):

Reinforcement Learning ◽

Collective Behavior ◽

Design Methodology ◽

Learning Performance ◽

Centralized Control ◽

Robot System ◽

Swarm Robots ◽

Robot Systems ◽

Swarm Robot ◽

Typical Design

Swarm robotic systems (SRSs) are a type of multi-robot system in which robots operate without any form of centralized control. The typical design methodology for SRSs comprises a behavior-based approach, where the desired collective behavior is obtained manually by designing the behavior of individual robots in advance. In contrast, in an automatic design approach, a certain general methodology is adopted. This paper presents a deep reinforcement learning approach for collective behavior acquisition of SRSs. The swarm robots are expected to collect information in parallel and share their experience for accelerating their learning. We conducted real swarm robot experiments and evaluated the learning performance of the swarm in a scenario where the robots consecutively traveled between two landmarks.

Download Full-text

DQN as an alternative to Market-based approaches for Multi-Robot processing Task Allocation (MRpTA)

International Journal of Robotic Computing ◽

10.35708/rc1870-126266 ◽

2021 ◽

Vol 3 (1) ◽

pp. 69-98

Author(s):

Paul Gautier ◽

Johann Laurent

Keyword(s):

Reinforcement Learning ◽

Task Allocation ◽

Computing System ◽

Processing Load ◽

Robot System ◽

Research Challenges ◽

Local Solutions ◽

Overall Efficiency ◽

Robot Task ◽

Multi Robot

Multi-robot task allocation (MRTA) problems require that robots make complex choices based on their understanding of a dynamic and uncertain environment. As a distributed computing system, the Multi-Robot System (MRS) must handle and distribute processing tasks (MRpTA). Each robot must contribute to the overall efficiency of the system based solely on a limited knowledge of its environment. Market-based methods are a natural candidate to deal processing tasks over a MRS but recent and numerous developments in reinforcement learning and especially Deep Q-Networks (DQN) provide new opportunities to solve the problem. In this paper we propose a new DQN-based method so that robots can learn directly from experience, and compare it with Market-based approaches as well with centralized and purely local solutions. Our study shows the relevancy of learning-based methods and also highlight research challenges to solve the processing load-balancing problem in MRS.

Download Full-text

Alignment Method of Combined Perception for Peg-in-Hole Assembly with Deep Reinforcement Learning

Journal of Sensors ◽

10.1155/2021/5073689 ◽

2021 ◽

Vol 2021 ◽

pp. 1-12

Author(s):

Yongzhi Wang ◽

Lei Zhao ◽

Qian Zhang ◽

Ran Zhou ◽

Liping Wu ◽

...

Keyword(s):

Reinforcement Learning ◽

Visual Perception ◽

Simulation Training ◽

Tactile Perception ◽

Visual Feature ◽

Alignment Method ◽

Torque Sensor ◽

Robot System ◽

Contact State ◽

Simulation Results

The method of tactile perception can accurately reflect the contact state by collecting force and torque information, but it is not sensitive to the changes in position and posture between assembly objects. The method of visual perception is very sensitive to changes in pose and posture between assembled objects, but they cannot accurately reflect the contact state, especially since the objects are occluded from each other. The robot will perceive the environment more accurately if visual and tactile perception can be combined. Therefore, this paper proposes the alignment method of combined perception for the peg-in-hole assembly with self-supervised deep reinforcement learning. The agent first observes the environment through visual sensors and then predicts the action of the alignment adjustment based on the visual feature of the contact state. Subsequently, the agent judges the contact state based on the force and torque information collected by the force/torque sensor. And the action of the alignment adjustment is selected according to the contact state and used as a visual prediction label. Whereafter, the network of visual perception performs backpropagation to correct the network weights according to the visual prediction label. Finally, the agent will have learned the alignment skill of combined perception with the increase of iterative training. The robot system is built based on CoppeliaSim for simulation training and testing. The simulation results show that the method of combined perception has higher assembly efficiency than single perception.

Download Full-text