Replacing the deep critic with a broad learning system trained by ridge regression makes DDPG, SAC, and TD3 train faster on MuJoCo continuous-control benchmarks, with roughly similar or slightly better final rewards.
Deep reinforcement learning on variable stiffness compliant control for programming -free robotic assembly in smart manufacturing,
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.LG 1years
2024 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
Broad Critic Deep Actor Reinforcement Learning for Continuous Control
Replacing the deep critic with a broad learning system trained by ridge regression makes DDPG, SAC, and TD3 train faster on MuJoCo continuous-control benchmarks, with roughly similar or slightly better final rewards.