Machine learningML & AI Methods
MA-RLHF: Reinforcement Learning from Human Feedback with Macro Actions
The MA-RLHF framework integrates macro actions into the learning process of large language models, enhancing learning efficiency and performance in tasks like text summarization and dialogue generation.
Featured in No. 69 on 9 Oct 2024 · 6 days after release · 15 citations today · published in International Conference on Learning Representations
- Released
- 3 Oct 2024
- First featured
- No. 69 · 9 Oct 2024
- Citations (Semantic Scholar)
- 15
- Influential citations
- 1
- Published in
- International Conference on Learning Representations
- Shares when featured
- 10
- Identifier
- arXiv:2410.02743
Citations and venue from Semantic Scholar (ODC-BY), refreshed weekly. Summary: Quant Letter (CC BY 4.0).