Optimal Stroke Learning with Policy Gradient Approach for Robotic Table Tennis

Gao, Yapeng; Tebbe, Jonas; Zell, Andreas

doi:10.1007/s10489-022-04131-w

Computer Science > Robotics

arXiv:2109.03100 (cs)

[Submitted on 7 Sep 2021 (v1), last revised 2 Nov 2021 (this version, v2)]

Title:Optimal Stroke Learning with Policy Gradient Approach for Robotic Table Tennis

Authors:Yapeng Gao, Jonas Tebbe, Andreas Zell

View PDF

Abstract:Learning to play table tennis is a challenging task for robots, as a wide variety of strokes required. Recent advances have shown that deep Reinforcement Learning (RL) is able to successfully learn the optimal actions in a simulated environment. However, the applicability of RL in real scenarios remains limited due to the high exploration effort. In this work, we propose a realistic simulation environment in which multiple models are built for the dynamics of the ball and the kinematics of the robot. Instead of training an end-to-end RL model, a novel policy gradient approach with TD3 backbone is proposed to learn the racket strokes based on the predicted state of the ball at the hitting time. In the experiments, we show that the proposed approach significantly outperforms the existing RL methods in simulation. Furthermore, to cross the domain from simulation to reality, we adopt an efficient retraining method and test it in three real scenarios. The resulting success rate is 98% and the distance error is around 24.9 cm. The total training time is about 1.5 hours.

Subjects:	Robotics (cs.RO)
Cite as:	arXiv:2109.03100 [cs.RO]
	(or arXiv:2109.03100v2 [cs.RO] for this version)
	https://doi.org/10.48550/arXiv.2109.03100
Related DOI:	https://doi.org/10.1007/s10489-022-04131-w

Submission history

From: Yapeng Gao [view email]
[v1] Tue, 7 Sep 2021 14:00:13 UTC (4,576 KB)
[v2] Tue, 2 Nov 2021 14:29:01 UTC (7,534 KB)

Computer Science > Robotics

Title:Optimal Stroke Learning with Policy Gradient Approach for Robotic Table Tennis

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Robotics

Title:Optimal Stroke Learning with Policy Gradient Approach for Robotic Table Tennis

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators