Addressing Function Approximation Error in Actor-Critic Methods

E1274356 UNEXPLORED

"Addressing Function Approximation Error in Actor-Critic Methods" is a research paper that introduces the Twin Delayed Deep Deterministic Policy Gradient (TD3) algorithm to improve stability and performance in continuous control reinforcement learning.

All labels observed (1)

How this entity was disambiguated

Referenced by (1)

Full triples — surface form annotated when it differs from this entity's canonical label.

TD3 introducedInPaper Addressing Function Approximation Error in Actor-Critic Methods