NeurIPS · 2025 · Journal article · Top venue

Scaling Off-Policy Reinforcement Learning with Batch and Weight Normalization

Daniel Palenicek, Florian Vogt, Joe Watson, Jan Peters

Rendered by arXiv from the LaTeX source. If anything looks wrong, switch to the PDF.HTML · PDF · Open in a new tab ↗ · Close
Published in
Neural Information Processing Systems
Date
2025-02-11
Citations
20
arXiv
2502.07523
Cite
@article{palenicek2025scaling,
  title = {Scaling Off-Policy Reinforcement Learning with Batch and Weight Normalization},
  author = {Daniel Palenicek and Florian Vogt and Joe Watson and Jan Peters},
  year = {2025},
  journal = {Neural Information Processing Systems},
  eprint = {2502.07523},
  archivePrefix = {arXiv},
  url = {https://arxiv.org/abs/2502.07523},
}