Machine learningML & AI Methods
Budgeting Counterfactual for Offline RL
The paper suggests a method to improve offline reinforcement learning by limiting out-of-distribution actions during training, showing improved performance on D4RL benchmarks.
Featured in No. 51 on 28 May 2024 · · 6 citations today · published in Neural Information Processing Systems
- Released
- 12 Jul 2023
- First featured
- No. 51 · 28 May 2024
- Citations (Semantic Scholar)
- 6
- Influential citations
- 0
- Published in
- Neural Information Processing Systems
- Shares when featured
- 14
- Identifier
- arXiv:2307.06328
Citations and venue from Semantic Scholar (ODC-BY), refreshed weekly. Summary: Quant Letter (CC BY 4.0).