Implementation of Twin Delayed Deep Deterministic Policy Gradient (TD3) for humanoid continuous control in a Unity physics simulation environment using Keras and TensorFlow.
-
Updated
Mar 8, 2026 - C#
Implementation of Twin Delayed Deep Deterministic Policy Gradient (TD3) for humanoid continuous control in a Unity physics simulation environment using Keras and TensorFlow.
Unity racing with raycast perception, continuous control, and a custom PyTorch SAC agent. Includes selected policy, real gameplay, and reproducible evaluation evidence.
To associate your repository with the continuous-control topic, visit your repo's landing page and select "manage topics."