Replacing the deep critic with a broad learning system trained by ridge regression makes DDPG, SAC, and TD3 train faster on MuJoCo continuous-control benchmarks, with roughly similar or slightly better final rewards.
Augmenting reinforcement learning with Transformer -based scene representation learning for decision-making of autonomous driving,
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.LG 1years
2024 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
Broad Critic Deep Actor Reinforcement Learning for Continuous Control
Replacing the deep critic with a broad learning system trained by ridge regression makes DDPG, SAC, and TD3 train faster on MuJoCo continuous-control benchmarks, with roughly similar or slightly better final rewards.