ML-QuantSubscribe

Machine learningLLMs & Text

Efficient Adversarial Training in LLMs with Continuous Attacks

CAdvUL, a new adversarial training algorithm, enhances the resilience of large language models against adversarial attacks by efficiently calculating attacks in the continuous embedding space.

Featured in No. 73 on 6 Nov 2024 · · 148 citations today · published in Neural Information Processing Systems

Released
24 May 2024
First featured
No. 73 · 6 Nov 2024
Citations (Semantic Scholar)
148
Influential citations
14
Published in
Neural Information Processing Systems
Shares when featured
82
Identifier
arXiv:2405.15589

Citations and venue from Semantic Scholar (ODC-BY), refreshed weekly. Summary: Quant Letter (CC BY 4.0).

    Type to search. Try rough volatility, LLM agents or FinGPT.

    ↑↓ move↵ openesc closeFull search page