Skip to main navigation Skip to search Skip to main content

An encoder-decoder based audio captioning system with transfer and reinforcement learning

  • Xinhao Mei
  • , Qiushi Huang
  • , Xubo Liu
  • , Gengyun Chen
  • , Jingqian Wu
  • , Yusong Wu
  • , Jinzheng Zhao
  • , Shengchen Li
  • , Tom Ko
  • , Lilian Tang
  • , Xi Shao
  • , Mark D. Plumbley
  • , Wenwu Wang
  • University of Surrey
  • Southern University of Science and Technology
  • Nanjing University of Posts and Telecommunications
  • Beijing University of Posts and Telecommunications
  • Tencent
  • ByteDance Ltd.

Research output: Chapter in Book or Report/Conference proceedingConference Proceedingpeer-review

Original languageEnglish
Title of host publicationProceedings of the 6th Detection and Classification of Acoustic Scenes and Events 2021 Workshop (DCASE2021)
DOIs
Publication statusPublished - 15 Nov 2021

Cite this