Fetching the paper…
Reading the bibliography…
Data scarcity is a problem that occurs in languages and tasks where we do not have large amounts of labeled data but want to use state-of-the-art models.
1905
Earlier work this paper cites.
1905
Earlier work this paper cites.
1908
Earlier work this paper cites.
1909
Earlier work this paper cites.
G. A. Miller, “Wordnet: A lexical database for english,” Commun. ACM , vol. 38, no. 11, p. 39–41, nov 1995. [Online]. Available: https://doi.org/10.1145/219717.219748
1995
Earlier work this paper cites.
N. V. Chawla, K. W. Bowyer et al. , “Smote: Synthetic minority over-sampling technique,” J. Artif. Int. Res. , vol. 16, no. 1, p. 321–357, jun 2002
2002
Earlier work this paper cites.
P. Y. Simard, D. Steinkraus, and J. C. Platt, “Best practices for convolutional neural networks applied to visual document analysis,” Proceedings of the International Conference on Document Analysis and Recognition, ICDAR , vol. 2003-January, pp. 958–963, 2003
2003
Earlier work this paper cites.
2005
Earlier work this paper cites.
2005
Earlier work this paper cites.
J. Ebrahimi, A. Rao et al. , “HotFlip: White-box adversarial examples for text classification,” in Proceedings of the 56th Annual Meeting of the Association for Computational Linguistics (Volume 2: Short Papers) . Melbourne, Australia: Association for Computational Linguistics, Jul. 2018, pp. 31–36. [Online]. Available: https://aclanthology.org/P18-2006
2006
Earlier work this paper cites.
2011
Earlier work this paper cites.
R. Bhagat and E. Hovy, “What Is a Paraphrase?” Computational Linguistics , vol. 39, no. 3, pp. 463–472, 09 2013. [Online]. Available: https://doi.org/10.1162/COLI_a_00166
2013
Earlier work this paper cites.
T. Ko, V. Peddinti et al. , “Audio augmentation for speech recognition,” Proceedings of the Annual Conference of the International Speech Communication Association, INTERSPEECH , vol. 2015-January, pp. 3586–3589, 2015
2015
Earlier work this paper cites.
W. Y. Wang and D. Yang, “That’s so annoying!!!: A lexical and frame-semantic embedding based data augmentation approach to automatic categorization of annoying behaviors using #petpeeve tweets,” in Proceedings of the 2015 Conference on Empirical Methods in Natural Language Processing . Lisbon, Portugal: Association for Computational Linguistics, Sep. 2015, pp. 2557–2563. [Online]. Available: https://aclanthology.org/D15-1306
2015
Earlier work this paper cites.
W. Ling, C. Dyer et al. , “Finding function in form: Compositional character models for open vocabulary word representation,” in Proceedings of the 2015 Conference on Empirical Methods in Natural Language Processing . Lisbon, Portugal: Association for Computational Linguistics, Sep. 2015, pp. 1520–1530. [Online]. Available: https://aclanthology.org/D15-1176
2015
Earlier work this paper cites.
R. Sennrich, B. Haddow, and A. Birch, “Neural machine translation of rare words with subword units,” in Proceedings of the 54th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) . Berlin, Germany: Association for Computational Linguistics, Aug. 2016, pp. 1715–1725. [Online]. Available: https://aclanthology.org/P16-1162
2016
Earlier work this paper cites.
2016
Earlier work this paper cites.
2017
Earlier work this paper cites.
2018
Earlier work this paper cites.
J. H. Park, J. Shin, and P. Fung, “Reducing gender bias in abusive language detection,” in Proceedings of the 2018 Conference on Empirical Methods in Natural Language Processing . Brussels, Belgium: Association for Computational Linguistics, Oct.-Nov. 2018, pp. 2799–2804. [Online]. Available: https://aclanthology.org/D18-1302
2018
Earlier work this paper cites.
M. Alzantot, Y. Sharma et al. , “Generating natural language adversarial examples,” Proceedings of the 2018 Conference on Empirical Methods in Natural Language Processing, EMNLP 2018 , pp. 2890–2896, 2018. [Online]. Available: https://aclanthology.org/D18-1316
2018
Earlier work this paper cites.
2018
Earlier work this paper cites.
G. G. Şahin and M. Steedman, “Data augmentation via dependency tree morphing for low-resource languages,” in Proceedings of the 2018 Conference on Empirical Methods in Natural Language Processing . Brussels, Belgium: Association for Computational Linguistics, Oct.-Nov. 2018, pp. 5004–5009. [Online]. Available: https://aclanthology.org/D18-1545
2018
Earlier work this paper cites.
2018
Earlier work this paper cites.
2018
Earlier work this paper cites.
2018
Earlier work this paper cites.
V. Yadav and S. Bethard, “A survey on recent advances in named entity recognition from deep learning models,” in Proceedings of the 27th International Conference on Computational Linguistics . Santa Fe, New Mexico, USA: Association for Computational Linguistics, Aug. 2018, pp. 2145–2158. [Online]. Available: https://aclanthology.org/C18-1182
2018
Cited alongside, same era.
J. Devlin, M.-W. Chang et al. , “BERT: Pre-training of deep bidirectional transformers for language understanding,” in Proceedings of the 2019 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, Volume 1 (Long and Short Papers) . Minneapolis, Minnesota: Association for Computational Linguistics, Jun. 2019, pp. 4171–4186. [Online]. Available: https://aclanthology.org/N19-1423
2019
Cited alongside, same era.
C. Shorten and T. M. Khoshgoftaar, “A survey on image data augmentation for deep learning,” Journal of Big Data , vol. 6, pp. 1–48, 12 2019. [Online]. Available: https://journalofbigdata.springeropen.com/articles/10.1186/s40537-019-0197-0
2019
Cited alongside, same era.
S. Y. Feng, V. Gangal et al. , “A survey of data augmentation approaches for NLP,” Findings of the Association for Computational Linguistics: ACL-IJCNLP 2021 , pp. 968–988, 2021. [Online]. Available: https://aclanthology.org/2021.findings-acl.84
2021
Later among the works it cites.
B. Markus, K. Marc-André, and R. Christian, “A survey on data augmentation for text classification,” ACM Computing Surveys , 7 2021. [Online]. Available: https://dl.acm.org/doi/10.1145/3544558
2021
Later among the works it cites.
M. A. Hedderich, L. Lange et al. , “A survey on recent approaches for natural language processing in low-resource scenarios,” in Proceedings of the 2021 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies . Online: Association for Computational Linguistics, Jun. 2021, pp. 2545–2568. [Online]. Available: https://aclanthology.org/2021.naacl-main.201
2021
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
J. Wei and K. Zou, “Eda: Easy data augmentation techniques for boosting performance on text classification tasks,” EMNLP-IJCNLP 2019 - 2019 Conference on Empirical Methods in Natural Language Processing and 9th International Joint Conference on Natural Language Processing, Proceedings of the Conference , pp. 6382–6388, 2019. [Online]. Available: https://aclanthology.org/D19-1670
2019
Cited alongside, same era.
X. Wu, S. Lv et al. , “Conditional bert contextual augmentation,” Lecture Notes in Computer Science (including subseries Lecture Notes in Artificial Intelligence and Lecture Notes in Bioinformatics) , vol. 11539 LNCS, pp. 84–95, 2019. [Online]. Available: https://link.springer.com/chapter/10.1007/978-3-030-22747-0_7
2019
Cited alongside, same era.
V. Karpukhin, O. Levy et al. , “Training on synthetic noise improves robustness to natural noise in machine translation,” in Proceedings of the 5th Workshop on Noisy User-generated Text (W-NUT 2019) . Hong Kong, China: Association for Computational Linguistics, Nov. 2019, pp. 42–47. [Online]. Available: https://aclanthology.org/D19-5506
2019
Cited alongside, same era.
D. Berthelot, N. Carlini et al. , MixMatch: A Holistic Approach to Semi-Supervised Learning . Red Hook, NY, USA: Curran Associates Inc., 2019
2019
Cited alongside, same era.
G. Yan, Y. Li et al. , “Data augmentation for deep learning of judgment documents,” in Intelligence Science and Big Data Engineering. Big Data and Machine Learning , Z. Cui, J. Pan et al. , Eds. Cham: Springer International Publishing, 2019, pp. 232–242
2019
Cited alongside, same era.
G. Rizos, K. Hemker, and B. Schuller, “Augment to prevent: Short-text data augmentation in deep learning for hate-speech classification,” in Proceedings of the 28th ACM International Conference on Information and Knowledge Management , ser. CIKM ’19. New York, NY, USA: Association for Computing Machinery, 2019, p. 991–1000. [Online]. Available: https://doi.org/10.1145/3357384.3358040
2019
Cited alongside, same era.
Y. Cheng, L. Jiang, and W. Macherey, “Robust neural machine translation with doubly adversarial inputs,” in Proceedings of the 57th Annual Meeting of the Association for Computational Linguistics . Florence, Italy: Association for Computational Linguistics, Jul. 2019, pp. 4324–4333. [Online]. Available: https://aclanthology.org/P19-1425
2019
Cited alongside, same era.
M. Jungiewicz and A. Smywinski-Pohl, “Towards textual data augmentation for neural networks: synonyms and maximum loss,” Computer Science , vol. 20, no. 1, Mar. 2019. [Online]. Available: https://journals.agh.edu.pl/csci/article/view/3023
2019
Cited alongside, same era.
2019
Cited alongside, same era.
T.-H. Cheung and D.-Y. Yeung, “{MODALS}: Modality-agnostic automated data augmentation in the latent space,” in International Conference on Learning Representations , 2021. [Online]. Available: https://openreview.net/forum?id=XjYgR6gbCEc
2021
Later among the works it cites.
C. Shorten, T. M. Khoshgoftaar, and B. Furht, “Text data augmentation for deep learning,” Journal of Big Data 2021 8:1 , vol. 8, pp. 1–34, 7 2021. [Online]. Available: https://journalofbigdata.springeropen.com/articles/10.1186/s40537-021-00492-0
2021
Later among the works it cites.
A. Karimi, L. Rossi, and A. Prati, “Aeda: An easier data augmentation technique for text classification,” Findings of the Association for Computational Linguistics, Findings of ACL: EMNLP 2021 , pp. 2748–2754, 2021. [Online]. Available: https://aclanthology.org/2021.findings-emnlp.234
2021
Later among the works it cites.
H. Shi, K. Livescu, and K. Gimpel, “Substructure substitution: Structured data augmentation for NLP,” in Findings of the Association for Computational Linguistics: ACL-IJCNLP 2021 . Online: Association for Computational Linguistics, Aug. 2021, pp. 3494–3508. [Online]. Available: https://aclanthology.org/2021.findings-acl.307
2021
Later among the works it cites.
K. M. Yoo, D. Park et al. , “GPT3Mix: Leveraging large-scale language models for text augmentation,” in Findings of the Association for Computational Linguistics: EMNLP 2021 . Punta Cana, Dominican Republic: Association for Computational Linguistics, Nov. 2021, pp. 2225–2239. [Online]. Available: https://aclanthology.org/2021.findings-emnlp.192
2021
Later among the works it cites.
N. Thakur, N. Reimers et al. , “Augmented SBERT: Data augmentation method for improving bi-encoders for pairwise sentence scoring tasks,” in Proceedings of the 2021 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies . Online: Association for Computational Linguistics, Jun. 2021, pp. 296–310. [Online]. Available: https://aclanthology.org/2021.naacl-main.28
2021
Later among the works it cites.
M. S. Bari, T. Mohiuddin, and S. Joty, “UXLA: A robust unsupervised data augmentation framework for zero-resource cross-lingual NLP,” in Proceedings of the 59th Annual Meeting of the Association for Computational Linguistics and the 11th International Joint Conference on Natural Language Processing (Volume 1: Long Papers) . Online: Association for Computational Linguistics, Aug. 2021, pp. 1978–1992. [Online]. Available: https://aclanthology.org/2021.acl-long.154
2021
Later among the works it cites.
J. Wei, C. Huang et al. , “Few-shot text classification with triplet networks, data augmentation, and curriculum learning,” in Proceedings of the 2021 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies . Online: Association for Computational Linguistics, Jun. 2021, pp. 5493–5500. [Online]. Available: https://aclanthology.org/2021.naacl-main.434
2021
Later among the works it cites.
J. Wei, C. Huang et al. , “Text augmentation in a multi-task view,” in Proceedings of the 16th Conference of the European Chapter of the Association for Computational Linguistics: Main Volume . Online: Association for Computational Linguistics, Apr. 2021, pp. 2888–2894. [Online]. Available: https://aclanthology.org/2021.eacl-main.252
2021
Later among the works it cites.
S. Ren, J. Zhang et al. , “Text AutoAugment: Learning compositional augmentation policy for text classification,” in Proceedings of the 2021 Conference on Empirical Methods in Natural Language Processing . Online and Punta Cana, Dominican Republic: Association for Computational Linguistics, Nov. 2021, pp. 9029–9043. [Online]. Available: https://aclanthology.org/2021.emnlp-main.711
2021
Later among the works it cites.
Q. Li, Z. Huang et al. , “A framework of data augmentation while active learning for chinese named entity recognition,” in Knowledge Science, Engineering and Management: 14th International Conference, KSEM 2021, Tokyo, Japan, August 14–16, 2021, Proceedings, Part II . Berlin, Heidelberg: Springer-Verlag, 2021, p. 88–100. [Online]. Available: https://doi.org/10.1007/978-3-030-82147-0_8
2021
Later among the works it cites.
G. Zeng, F. Qi et al. , “OpenAttack: An open-source textual adversarial attack toolkit,” in Proceedings of the 59th Annual Meeting of the Association for Computational Linguistics and the 11th International Joint Conference on Natural Language Processing: System Demonstrations . Online: Association for Computational Linguistics, Aug. 2021, pp. 363–371. [Online]. Available: https://aclanthology.org/2021.acl-demo.43
2021
Later among the works it cites.
2021
Later among the works it cites.
X. Wang, Q. Liu et al. , “TextFlint: Unified multilingual robustness evaluation toolkit for natural language processing,” in Proceedings of the 59th Annual Meeting of the Association for Computational Linguistics and the 11th International Joint Conference on Natural Language Processing: System Demonstrations . Online: Association for Computational Linguistics, Aug. 2021, pp. 347–355. [Online]. Available: https://aclanthology.org/2021.acl-demo.41
2021
Later among the works it cites.
G. G. Şahin, “To augment or not to augment? a comparative study on text augmentation techniques for low-resource NLP,” Computational Linguistics , vol. 48, pp. 5–42, 4 2022. [Online]. Available: https://aclanthology.org/2022.cl-1.2
2022
Later among the works it cites.
B. Li, Y. Hou, and W. Che, “Data augmentation approaches in natural language processing: A survey,” AI Open , 2022. [Online]. Available: https://www.sciencedirect.com/science/article/pii/S2666651022000080
2022
Later among the works it cites.
2022
Later among the works it cites.
M.-R. Amini, V. Feofanov et al. , “Self-training: A survey,” ArXiv , vol. abs/2202.12040, 2022
2022
Later among the works it cites.
S. Y. Park and C. Caragea, “A data cartography based MixUp for pre-trained language models,” in Proceedings of the 2022 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies . Seattle, United States: Association for Computational Linguistics, Jul. 2022, pp. 4244–4250. [Online]. Available: https://aclanthology.org/2022.naacl-main.314
2022
Later among the works it cites.
2022
Later among the works it cites.
Y. Zheng, Z. Zhang et al. , “Deep autoaugment,” in International Conference on Learning Representations , 2022. [Online]. Available: https://openreview.net/forum?id=St-53J9ZARf
2022
Later among the works it cites.
2022
Later among the works it cites.
E. Pavlick, P. Rastogi et al. , “PPDB 2.0: Better paraphrase ranking, fine-grained entailment relations, word embeddings, and style classification,” in Proceedings of the 53rd Annual Meeting of the Association for Computational Linguistics and the 7th International Joint Conference on Natural Language Processing (Volume 2: Short Papers) . Beijing, China: Association for Computational Linguistics, Jul. 2015, pp. 425–430. [Online]. Available: https://aclanthology.org/P15-2070
2070
Closest in time.