ML-QuantSubscribe

Machine learningML & AI Methods

MA-RLHF: Reinforcement Learning from Human Feedback with Macro Actions

The MA-RLHF framework integrates macro actions into the learning process of large language models, enhancing learning efficiency and performance in tasks like text summarization and dialogue generation.

Featured in No. 69 on 9 Oct 2024 · 6 days after release · 15 citations today · published in International Conference on Learning Representations

Released
3 Oct 2024
First featured
No. 69 · 9 Oct 2024
Citations (Semantic Scholar)
15
Influential citations
1
Published in
International Conference on Learning Representations
Shares when featured
10
Identifier
arXiv:2410.02743

Citations and venue from Semantic Scholar (ODC-BY), refreshed weekly. Summary: Quant Letter (CC BY 4.0).

    Type to search. Try rough volatility, LLM agents or FinGPT.

    ↑↓ move↵ openesc closeFull search page