Machine learningML & AI Methods
Taming Data and Transformers for Audio Generation
AutoCap and GenAu, two new models for generating ambient sounds and effects, are introduced, improving the quality of audio captions and generated audio.
Featured in No. 55 on 3 Jul 2024 · 6 days after release · 40 citations today · published in International Journal of Computer Vision
- Released
- 27 Jun 2024
- First featured
- No. 55 · 3 Jul 2024
- Citations (Semantic Scholar)
- 40
- Influential citations
- 0
- Published in
- International Journal of Computer Vision
- Shares when featured
- 6
- Identifier
- arXiv:2406.19388
Citations and venue from Semantic Scholar (ODC-BY), refreshed weekly. Summary: Quant Letter (CC BY 4.0).