ML-QuantSubscribe

Machine learningML & AI Methods

Budgeting Counterfactual for Offline RL

The paper suggests a method to improve offline reinforcement learning by limiting out-of-distribution actions during training, showing improved performance on D4RL benchmarks.

Featured in No. 51 on 28 May 2024 · · 6 citations today · published in Neural Information Processing Systems

Released
12 Jul 2023
First featured
No. 51 · 28 May 2024
Citations (Semantic Scholar)
6
Influential citations
0
Published in
Neural Information Processing Systems
Shares when featured
14
Identifier
arXiv:2307.06328

Citations and venue from Semantic Scholar (ODC-BY), refreshed weekly. Summary: Quant Letter (CC BY 4.0).

    Type to search. Try rough volatility, LLM agents or FinGPT.

    ↑↓ move↵ openesc closeFull search page