Audio Retrieval with Natural Language Queries

Andreea-Maria Oncescu,Zeynep Akata,Samuel Albanie,João F Henriques,A Sophia Koepke

doi:10.21437/interspeech.2021-2227

Audio Retrieval with Natural Language Queries

Andreea-Maria Oncescu, Zeynep Akata + Show 3 more

Open Access

https://doi.org/10.21437/interspeech.2021-2227

Copy DOI

Publication Date: Aug 30, 2021

Citations: 24

#Audio Retrieval #Cross-modal Retrieval + Show 8 more

Abstract
Full-Text PDF
Similar Papers

Abstract

We consider the task of retrieving audio using free-form natural language queries. To study this problem, which has received limited attention in the existing literature, we introduce challenging new benchmarks for text-based audio retrieval using text annotations sourced from the Audiocaps and Clotho datasets. We then employ these benchmarks to establish baselines for cross-modal audio retrieval, where we demonstrate the benefits of pre-training on diverse audio tasks. We hope that our benchmarks will inspire further research into cross-modal text-based audio retrieval with free-form text queries.

Full Text

Paper version not known

Open DOI Link

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

Similar Papers

Paper Title

Journal

Date

Author

View more papers

Disclaimer: All third-party content on this website/platform is and will remain the property of their respective owners and is provided on "as is" basis without any warranties, express or implied. Use of third-party content does not indicate any affiliation, sponsorship with or endorsement by them. Any references to third-party content is to identify the corresponding services and shall be considered fair use under The CopyrightLaw.