Fetching the paper…
Reading the bibliography…
Simultaneous translation models play a crucial role in facilitating communication.
Bleu: a method for automatic evaluation of machine translation
Kishore Papineni, Salim Roukos, Todd Ward, and Wei-Jing Zhu. 2002 · 2002
Earlier work this paper cites.
Glancing transformer for non-autoregressive neural machine translation
Lihua Qian, Hao Zhou, Yu Bao, Mingxuan Wang, Lin Qiu, Weinan Zhang, Yong Yu, and Lei Li. 2021 · 2003
Earlier work this paper cites.
Connectionist temporal classification: Labelling unsegmented sequence data with recurrent neural networks
Alex Graves, Santiago Fernández, Faustino Gomez, and Jürgen Schmidhuber. 2006 · 2006
Earlier work this paper cites.
Sequence transduction with recurrent neural networks
Alex Graves. 2012 · 2012
Earlier work this paper cites.
Adam: A method for stochastic optimization
Diederik P. Kingma and Jimmy Ba. 2015 · 2015
Earlier work this paper cites.
Can neural machine translation do simultaneous translation?
Kyunghyun Cho and Masha Esipova. 2016 · 2016
Earlier work this paper cites.
Sequence-level knowledge distillation
Yoon Kim and Alexander M. Rush. 2016 · 2016
Earlier work this paper cites.
MuST-C: a Multilingual Speech Translation Corpus
Mattia A. Di Gangi, Roldano Cattoni, Luisa Bentivogli, Matteo Negri, and Marco Turchi. 2019 · 2017
Earlier work this paper cites.
Learning to translate in real-time with neural machine translation
Jiatao Gu, Graham Neubig, Kyunghyun Cho, and Victor O.K. Li. 2017 · 2017
Earlier work this paper cites.
Online and linear-time attention by enforcing monotonic alignments
Colin Raffel, Minh-Thang Luong, Peter J. Liu, Ron J. Weiss, and Douglas Eck. 2017 · 2017
Earlier work this paper cites.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Ł ukasz Kaiser, and Illia Polosukhin. 2017 · 2017
Earlier work this paper cites.
Monotonic chunkwise attention
Chung-Cheng Chiu and Colin Raffel. 2018 · 2018
Earlier work this paper cites.
SentencePiece: A simple and language independent subword tokenizer and detokenizer for neural text processing
Taku Kudo and John Richardson. 2018 · 2018
Earlier work this paper cites.
A call for clarity in reporting BLEU scores
Matt Post. 2018 · 2018
Earlier work this paper cites.
Self-attention with relative position representations
Peter Shaw, Jakob Uszkoreit, and Ashish Vaswani. 2018 · 2018
Earlier work this paper cites.
Monotonic infinite lookback attention for simultaneous machine translation
Naveen Arivazhagan, Colin Cherry, Wolfgang Macherey, Chung-Cheng Chiu, Semih Yavuz, Ruoming Pang, Wei Li, and Colin Raffel. 2019 · 2019
Earlier work this paper cites.
Direct speech-to-speech translation with a sequence-to-sequence model
Ye Jia, Ron J. Weiss, Fadi Biadsy, Wolfgang Macherey, Melvin Johnson, Zhifeng Chen, and Yonghui Wu. 2019 · 2019
Earlier work this paper cites.
STACL: Simultaneous translation with implicit anticipation and controllable latency using prefix-to-prefix framework
Mingbo Ma, Liang Huang, Hao Xiong, Renjie Zheng, Kaibo Liu, Baigong Zheng, Chuanqiang Zhang, Zhongjun He, Hairong Liu, Xing Li, Hua Wu, and Haifeng Wang. 2019 · 2019
Cited alongside, same era.
Specaugment: A simple data augmentation method for automatic speech recognition
Daniel S Park, William Chan, Yu Zhang, Chung-Cheng Chiu, Barret Zoph, Ekin D Cubuk, and Quoc V Le. 2019 · 2019
Cited alongside, same era.
Conformer: Convolution-augmented Transformer for Speech Recognition
Anmol Gulati, James Qin, Chung-Cheng Chiu, Niki Parmar, Yu Zhang, Jiahui Yu, Wei Han, Shibo Wang, Zhengdong Zhang, Yonghui Wu, and Ruoming Pang. 2020 · 2020
Cited alongside, same era.
Direct segmentation models for streaming speech translation
Javier Iranzo-Sánchez, Adrià Giménez Pastor, Joan Albert Silvestre-Cerdà, Pau Baquero-Arnal, Jorge Civera Saiz, and Alfons Juan. 2020 · 2020
Cited alongside, same era.
Low-Latency Sequence-to-Sequence Speech Recognition and Translation by Partial Hypothesis Selection
Danni Liu, Gerasimos Spanakis, and Jan Niehues. 2020 · 2020
Speech Resynthesis from Discrete Disentangled Self-Supervised Representations
Adam Polyak, Yossi Adi, Jade Copet, Eugene Kharitonov, Kushal Lakhotia, Wei-Ning Hsu, Abdelrahman Mohamed, and Emmanuel Dupoux. 2021 · 2021
Later among the works it cites.
Fastspeech 2: Fast and high-quality end-to-end text to speech
Yi Ren, Chenxu Hu, Xu Tan, Tao Qin, Sheng Zhao, Zhou Zhao, and Tie-Yan Liu. 2021 · 2021
Later among the works it cites.
RealTranS: End-to-end simultaneous speech translation with convolutional weighted-shrinking transformer
Xingshan Zeng, Liangyou Li, and Qun Liu. 2021 · 2021
Later among the works it cites.
Directed acyclic transformer for non-autoregressive machine translation
Fei Huang, Hao Zhou, Yang Liu, Hang Li, and Minlie Huang. 2022b · 2022
Later among the works it cites.
Direct speech-to-speech translation with discrete units
Ann Lee, Peng-Jen Chen, Changhan Wang, Jiatao Gu, Sravya Popuri, Xutai Ma, Adam Polyak, Yossi Adi, Qing He, Yun Tang, Juan Pino, and Wei-Ning Hsu. 2022 · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Incremental text-to-speech synthesis with prefix-to-prefix framework
Mingbo Ma, Baigong Zheng, Kaibo Liu, Renjie Zheng, Hairong Liu, Kainan Peng, Kenneth Church, and Liang Huang. 2020a · 2020
Cited alongside, same era.
SIMULEVAL: An evaluation toolkit for simultaneous translation
Xutai Ma, Mohammad Javad Dousti, Changhan Wang, Jiatao Gu, and Juan Pino. 2020b · 2020
Cited alongside, same era.
SimulSpeech: End-to-end simultaneous speech to text translation
Yi Ren, Jinglin Liu, Xu Tan, Chen Zhang, Tao Qin, Zhou Zhao, and Tie-Yan Liu. 2020 · 2020
Cited alongside, same era.
Simultaneous speech-to-speech translation system with neural incremental asr, mt, and tts
Katsuhito Sudoh, Takatomo Kano, Sashi Novitasari, Tomoya Yanagita, Sakriani Sakti, and Satoshi Nakamura. 2020 · 2020
Cited alongside, same era.
Transformer transducer: A streamable speech recognition model with transformer encoders and rnn-t loss
Qian Zhang, Han Lu, Hasim Sak, Anshuman Tripathi, Erik McDermott, Stephen Koo, and Shankar Kumar. 2020 · 2020
Cited alongside, same era.
Fluent and low-latency simultaneous speech-to-speech translation with self-adaptive training
Renjie Zheng, Mingbo Ma, Baigong Zheng, Kaibo Liu, Jiahong Yuan, Kenneth Church, and Liang Huang. 2020 · 2020
Cited alongside, same era.
Direct simultaneous speech-to-text translation assisted by synchronized streaming ASR
Junkun Chen, Mingbo Ma, Renjie Zheng, and Liang Huang. 2021 · 2021
Cited alongside, same era.
Xutai Ma, Hongyu Gong, Danni Liu, Ann Lee, Yun Tang, Peng-Jen Chen, Wei-Ning Hsu, Phillip Koehn, and Juan Pino. 2022 · 2022
Later among the works it cites.
Over-generation cannot be rewarded: Length-adaptive average lagging for simultaneous speech translation
Sara Papi, Marco Gaido, Matteo Negri, and Marco Turchi. 2022 · 2022
Later among the works it cites.
Non-monotonic latent alignments for ctc-based non-autoregressive machine translation
Chenze Shao and Yang Feng. 2022 · 2022
Later among the works it cites.
Learning adaptive segmentation policy for end-to-end simultaneous translation
Ruiqing Zhang, Zhongjun He, Hua Wu, and Haifeng Wang. 2022 · 2022
Later among the works it cites.
Daspeech: Directed acyclic transformer for fast and high-quality speech-to-speech translation
Qingkai Fang, Yan Zhou, and Yang Feng. 2023 · 2023
Later among the works it cites.
UnitY: Two-pass direct speech-to-speech translation with discrete units
Hirofumi Inaguma, Sravya Popuri, Ilia Kulikov, Peng-Jen Chen, Changhan Wang, Yu-An Chung, Yun Tang, Ann Lee, Shinji Watanabe, and Juan Pino. 2023 · 2023
Later among the works it cites.
Non-autoregressive streaming transformer for simultaneous translation
Zhengrui Ma, Shaolei Zhang, Shoutao Guo, Chenze Shao, Min Zhang, and Yang Feng. 2023 · 2023
Later among the works it cites.
AlignAtt: Using Attention-based Audio-Translation Alignments as a Guide for Simultaneous Speech Translation
Sara Papi, Marco Turchi, and Matteo Negri. 2023c · 2023
Later among the works it cites.
Beyond mle: Convex learning for text generation
Chenze Shao, Zhengrui Ma, Min Zhang, and Yang Feng. 2023 · 2023
Later among the works it cites.
Hybrid transducer and attention based encoder-decoder modeling for speech-to-text tasks
Yun Tang, Anna Sun, Hirofumi Inaguma, Xinyue Chen, Ning Dong, Xutai Ma, Paden Tomasello, and Juan Pino. 2023 · 2023
Later among the works it cites.
End-to-end simultaneous speech translation with differentiable segmentation
Shaolei Zhang and Yang Feng. 2023a · 2023
Later among the works it cites.