ML-QuantSubscribe

Machine learningML & AI Methods

Taming Data and Transformers for Audio Generation

AutoCap and GenAu, two new models for generating ambient sounds and effects, are introduced, improving the quality of audio captions and generated audio.

Featured in No. 55 on 3 Jul 2024 · 6 days after release · 40 citations today · published in International Journal of Computer Vision

Released
27 Jun 2024
First featured
No. 55 · 3 Jul 2024
Citations (Semantic Scholar)
40
Influential citations
0
Published in
International Journal of Computer Vision
Shares when featured
6
Identifier
arXiv:2406.19388

Citations and venue from Semantic Scholar (ODC-BY), refreshed weekly. Summary: Quant Letter (CC BY 4.0).

    Type to search. Try rough volatility, LLM agents or FinGPT.

    ↑↓ move↵ openesc closeFull search page