Replicating SD3 Code and Results from Paper: Replicating Softmax Deep Double Deterministic Policy Gradients See PDF for more information