Fetching the paper…
Reading the bibliography…
Automated audio captioning (AAC) is the task of automatically generating textual descriptions for general audio signals.
Y. Koizumi, R. Masumura,
1981
Earlier work this paper cites.
T. Mikolov, I. Sutskever,
2013
Earlier work this paper cites.
F. Font, G. Roma,
2013
Earlier work this paper cites.
I. Sutskever, O. Vinyals,
2014
Earlier work this paper cites.
J. Pennington, R. Socher,
2014
Earlier work this paper cites.
K. Simonyan and A. Zisserman, “Very deep convolutional networks for large-scale image recognition,”
2015
Earlier work this paper cites.
Y. Zhu, R. Kiros,
2015
Earlier work this paper cites.
2016
Earlier work this paper cites.
K. Drossos, S. Adavanne,
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
S. Hershey, S. Chaudhuri,
2017
Earlier work this paper cites.
S. Liu, Z. Zhu,
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
J. F. Gemmeke, D. P. W. Ellis,
2017
Cited alongside, same era.
R. Arandjelovic and A. Zisserman, “Look, listen and learn,” in
2017
Cited alongside, same era.
T. Mikolov, E. Grave,
2018
Cited alongside, same era.
S. Lipping, K. Drossos,
2019
Cited alongside, same era.
M. Wu, H. Dinkel,
2019
Cited alongside, same era.
S. Ikawa and K. Kashino, “Neural audio captioning based on conditional sequence-to-sequence model,” in
2019
Cited alongside, same era.
C. D. Kim, B. Kim,
2019
Cited alongside, same era.
2020
Later among the works it cites.
2020
Later among the works it cites.
X. Xu, H. Dinkel,
2020
Later among the works it cites.
D. Takeuchi, Y. Koizumi,
2020
Later among the works it cites.
Y. Wu, K. Chen,
2020
Later among the works it cites.
K. Drossos, S. Lipping,
2020
Later among the works it cites.
X. Favory, K. Drossos,
2020
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
J. Cramer, H.-H. Wu,
2019
Cited alongside, same era.
J. Devlin, M.-W. Chang,
2019
Cited alongside, same era.
A. Ö. Eren and M. Sert, “Audio captioning based on combined audio and semantic embeddings,” in
2020
Cited alongside, same era.
2020
Cited alongside, same era.
A. Ö. Eren and M. Sert, “Audio captioning using gated recurrent units,”
2020
Cited alongside, same era.
Later among the works it cites.
X. Xu, H. Dinkel,
2021
Closest in time.
A. Tran, K. Drossos,
2021
Closest in time.
X. Xu, H. Dinkel,
2021
Closest in time.
W. Yuan, Q. Han,
2021
Closest in time.
2021
Closest in time.
2021
Closest in time.