ML-QuantSubscribe

Machine learningML & AI Methods

Survival Instinct in Offline Reinforcement Learning

Survival: Offline reinforcement learning algorithms can still create effective policies even with incorrect reward labels due to their inherent pessimism and biases in data collection.

Featured in No. 26 on 15 Nov 2023 · · 26 citations today · published in Neural Information Processing Systems

Released
5 Jun 2023
First featured
No. 26 · 15 Nov 2023
Citations (Semantic Scholar)
26
Influential citations
4
Published in
Neural Information Processing Systems
Shares when featured
110
Identifier
arXiv:2306.03286

Citations and venue from Semantic Scholar (ODC-BY), refreshed weekly. Summary: Quant Letter (CC BY 4.0).

    Type to search. Try rough volatility, LLM agents or FinGPT.

    ↑↓ move↵ openesc closeFull search page