Fetching the paper…
Reading the bibliography…
We propose task-adaptive tokenization as a way to adapt the generation pipeline to the specifics of a downstream task and enhance long-form generation in mental health.
Mike Lewis, Yinhan Liu, Naman Goyal, Marjan Ghazvininejad, Abdelrahman Mohamed, Omer Levy, Veselin Stoyanov, and Luke Zettlemoyer. 2019 · 1910
Earlier work this paper cites.
Huggingface’s transformers: State-of-the-art natural language processing
Thomas Wolf, Lysandre Debut, Victor Sanh, Julien Chaumond, Clement Delangue, Anthony Moi, Pierric Cistac, Tim Rault, Rémi Louf, Morgan Funtowicz, et al. 2019 · 1910
Earlier work this paper cites.
Immediate constituents
Rulon S Wells. 1947 · 1947
Earlier work this paper cites.
Cognitive structures in comprehension and memory of narrative discourse
Perry W Thorndyke. 1977 · 1977
Earlier work this paper cites.
Empirical evidence for narrative structure
James Paul Gee and Francois Grosjean. 1984 · 1984
Earlier work this paper cites.
Parsing by chunks
Steven P Abney. 1992 · 1992
Earlier work this paper cites.
The relationship of lexical proficiency to the quality of esl compositions
Cheryl A Engber. 1995 · 1995
Earlier work this paper cites.
Bleu: a method for automatic evaluation of machine translation
Kishore Papineni, Salim Roukos, Todd Ward, and Wei-Jing Zhu. 2002 · 2002
Earlier work this paper cites.
Statistical phrase-based translation
Philipp Koehn, Franz Josef Och, and Daniel Marcu. 2003 · 2003
Earlier work this paper cites.
Longformer: The long-document transformer
Iz Beltagy, Matthew E Peters, and Arman Cohan. 2020 · 2004
Earlier work this paper cites.
Vocabulary in a second language: Selection, acquisition, and testing , volume 10
Paul Bogaards and Batia Laufer. 2004 · 2004
Earlier work this paper cites.
ROUGE: A package for automatic evaluation of summaries
Chin-Yew Lin. 2004 · 2004
Earlier work this paper cites.
Morfessor in the morpho challenge
Mathias Creutz and Krista Lagus. 2006 · 2006
Earlier work this paper cites.
Moses: Open source toolkit for statistical machine translation
Philipp Koehn, Hieu Hoang, Alexandra Birch, Chris Callison-Burch, Marcello Federico, Nicola Bertoldi, Brooke Cowan, Wade Shen, Christine Moran, Richard Zens, Chris Dyer, Ondřej Bojar, Alexandra Constantin, and Evan Herbst. 2007 · 2007
Earlier work this paper cites.
Writing and Cognition
Mark Torrance, Luuk Van Waes, and David Galbraith. 2007 · 2007
Earlier work this paper cites.
Expressive interviewing: A conversational system for coping with covid-19
Charles F Welch, Allison Lahnala, Verónica Pérez-Rosas, Siqi Shen, Sarah Seraj, Lawrence An, Kenneth Resnicow, James W. Pennebaker, and Rada Mihalcea. 2020 · 2007
Earlier work this paper cites.
Vocabulary size and the skills of listening, reading and writing
Lars Stenius Stæhr. 2008 · 2008
Earlier work this paper cites.
Natural language processing with Python: analyzing text with the natural language toolkit
Steven Bird, Ewan Klein, and Edward Loper. 2009 · 2009
Earlier work this paper cites.
Improving access to psychological treatments: lessons from developing countries
Vikram Patel, Neerja Chowdhary, Atif Rahman, and Helen Verdeli. 2011 · 2011
Earlier work this paper cites.
A new resolution for global mental health
Rebecca S Hock, Flora Or, Kavitha Kolappa, Matthew D Burkey, Pamela J Surkan, and William W Eaton. 2012 · 2012
Earlier work this paper cites.
Better word representations with recursive neural networks for morphology
Thang Luong, Richard Socher, and Christopher Manning. 2013 · 2013
Earlier work this paper cites.
Demand and access to mental health services: a qualitative formative study in nepal
Natassia F Brenman, Nagendra P Luitel, Sumaya Mall, and Mark JD Jordans. 2014 · 2014
Earlier work this paper cites.
Effective strategies for turning receptive vocabulary into productive vocabulary in efl context
Avan Kamal Aziz Faraj. 2015 · 2015
Earlier work this paper cites.
On using very large target vocabulary for neural machine translation
Sébastien Jean, Kyunghyun Cho, Roland Memisevic, and Yoshua Bengio. 2015 · 2015
Earlier work this paper cites.
Neural machine translation of rare words with subword units
Rico Sennrich, Barry Haddow, and Alexandra Birch. 2015 · 2015
Cited alongside, same era.
Receptive vocabulary knowledge or productive vocabulary knowledge in writing skill, which one important
Zunita Mohamad Maskor, Harun Baharudin, et al. 2016 · 2016
Cited alongside, same era.
Neural machine translation of rare words with subword units
Rico Sennrich, Barry Haddow, and Alexandra Birch. 2016 · 2016
Cited alongside, same era.
Google’s neural machine translation system: Bridging the gap between human and machine translation
Yonghui Wu, Mike Schuster, Zhifeng Chen, Quoc V Le, Mohammad Norouzi, Wolfgang Macherey, Maxim Krikun, Yuan Cao, Qin Gao, Klaus Macherey, et al. 2016 · 2016
Cited alongside, same era.
Subword regularization: Improving neural network translation models with multiple subword candidates
Taku Kudo. 2018 · 2018
PsyQA: A Chinese dataset for generating long counseling text for mental health support
Hao Sun, Zhenru Lin, Chujie Zheng, Siyang Liu, and Minlie Huang. 2021 · 2021
Later among the works it cites.
Fastseq: Make sequence generation faster
Yu Yan, Fei Hu, Jiusheng Chen, Nikhil Bhendawade, Ting Ye, Yeyun Gong, Nan Duan, Desheng Cui, Bingyu Chi, and Ruofei Zhang. 2021 · 2021
Later among the works it cites.
Sangwon Yu, Jongyoon Song, Heeseung Kim, Seong-min Lee, Woo-Jong Ryu, and Sungroh Yoon. 2021 · 2021
Later among the works it cites.
Cliniqg4qa: Generating diverse questions for domain adaptation of clinical question answering
Xiang Yue, Xinliang Frederick Zhang, Ziyu Yao, Simon Lin, and Huan Sun. 2021 · 2021
Later among the works it cites.
Training a tokenizer for free with private federated learning
Eugene Bagdasaryan, Congzheng Song, Rogier van Dalen, Matt Seigel, and Áine Cahill. 2022 · 2022
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
SentencePiece: A simple and language independent subword tokenizer and detokenizer for neural text processing
Taku Kudo and John Richardson. 2018 · 2018
Cited alongside, same era.
Language models are unsupervised multitask learners
Alec Radford, Jeffrey Wu, Rewon Child, David Luan, Dario Amodei, Ilya Sutskever, et al. 2019 · 2019
Cited alongside, same era.
Q8bert: Quantized 8bit bert
Ofir Zafrir, Guy Boudoukh, Peter Izsak, and Moshe Wasserblat. 2019 · 2019
Cited alongside, same era.
A multi-persona chatbot for hotline counselor training
Orianna Demasi, Yu Li, and Zhou Yu. 2020 · 2020
Cited alongside, same era.
Stolen probability: A structural weakness of neural language models
David Demeter, Gregory Kimmel, and Doug Downey. 2020 · 2020
Cited alongside, same era.
Deepspeed: System optimizations enable training deep learning models with over 100 billion parameters
Jeff Rasley, Samyam Rajbhandari, Olatunji Ruwase, and Yuxiong He. 2020 · 2020
Cited alongside, same era.
Counseling-style reflection generation using generative pretrained transformers with augmented context
Siqi Shen, Charles Welch, Rada Mihalcea, and Verónica Pérez-Rosas. 2020 · 2020
Cited alongside, same era.
Later among the works it cites.
Time-aware prompting for text generation
Shuyang Cao and Lu Wang. 2022 · 2022
Later among the works it cites.
Hugging face optimum
Hugging Face Optimum developers. 2022 · 2022
Later among the works it cites.
Bridging fairness and environmental sustainability in natural language processing
Marius Hessenthaler, Emma Strubell, Dirk Hovy, and Anne Lauscher. 2022 · 2022
Later among the works it cites.
LoRA: Low-rank adaptation of large language models
Edward J Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu, Yuanzhi Li, Shean Wang, Lu Wang, and Weizhu Chen. 2022 · 2022
Later among the works it cites.
MentalBERT: Publicly available pretrained language models for mental healthcare
Shaoxiong Ji, Tianlin Zhang, Luna Ansari, Jie Fu, Prayag Tiwari, and Erik Cambria. 2022 · 2022
Later among the works it cites.
Elmer: A non-autoregressive pre-trained language model for efficient and effective text generation
Junyi Li, Tianyi Tang, Wayne Xin Zhao, Jian-Yun Nie, and Ji-Rong Wen. 2022 · 2022
Later among the works it cites.
Mega: moving average equipped gated attention
Xuezhe Ma, Chunting Zhou, Xiang Kong, Junxian He, Liangke Gui, Graham Neubig, Jonathan May, and Luke Zettlemoyer. 2022 · 2022
Later among the works it cites.
Sahand Sabour, Wen Zhang, Xiyao Xiao, Yuwei Zhang, Yinhe Zheng, Jiaxin Wen, Jialu Zhao, and Minlie Huang. 2022 · 2022
Later among the works it cites.
Do all languages cost the same? tokenization in the era of commercial language models
Orevaoghene Ahia, Sachin Kumar, Hila Gonen, Jungo Kasai, David R. Mortensen, Noah A. Smith, and Yulia Tsvetkov. 2023 · 2023
Closest in time.
Efficient and effective text encoding for chinese llama and alpaca
Yiming Cui, Ziqing Yang, and Xin Yao. 2023 · 2023
Closest in time.
Eva2. 0: Investigating open-domain chinese dialogue systems with large-scale pre-training
Yuxian Gu, Jiaxin Wen, Hao Sun, Yi Song, Pei Ke, Chujie Zheng, Zheng Zhang, Jianzhu Yao, Lei Liu, Xiaoyan Zhu, et al. 2023 · 2023
Closest in time.
Helping the helper: Supporting peer counselors via ai-empowered practice and feedback
Shang-Ling Hsu, Raj Sanjay Shah, Prathik Senthil, Zahra Ashktorab, Casey Dugan, Werner Geyer, and Diyi Yang. 2023 · 2023
Closest in time.
Study and analysis of chat gpt and its impact on different fields of study
Dinesh Kalla and Nathan Smith. 2023 · 2023
Closest in time.
A chatbot for mental health support: exploring the impact of emohaa on reducing mental distress in china
Sahand Sabour, Wen Zhang, Xiyao Xiao, Yuwei Zhang, Yinhe Zheng, Jiaxin Wen, Jialu Zhao, and Minlie Huang. 2023 · 2023
Closest in time.
Llama: Open and efficient foundation language models
Hugo Touvron, Thibaut Lavril, Gautier Izacard, Xavier Martinet, Marie-Anne Lachaux, Timothée Lacroix, Baptiste Rozière, Naman Goyal, Eric Hambro, Faisal Azhar, et al. 2023 · 2023
Closest in time.
Now it sounds like you: Learning personalized vocabulary on device
Sid Wang, Ashish Shenoy, Pierce Chuang, and John Nguyen. 2023 · 2023
Closest in time.
Qiang Zhang, Jason Naradowsky, and Yusuke Miyao. 2023 · 2023
Closest in time.
Tokenization and the noiseless channel
Vilém Zouhar, Clara Meister, Juan Gastaldi, Li Du, Mrinmaya Sachan, and Ryan Cotterell. 2023 · 2023
Closest in time.