Fetching the paper…
Reading the bibliography…
Since the advent of personal computing devices, intelligent personal assistants (IPAs) have been one of the key technologies that researchers and engineers have focused on, aiming to help users efficiently obtain information and execute tasks, and provide users with more intelligent, convenient, and rich interaction experiences.
Parameter-efficient transfer learning for NLP
Neil Houlsby, Andrei Giurgiu, Stanislaw Jastrzebski, Bruna Morrone, Quentin de Laroussilhe, Andrea Gesmundo, Mona Attariyan, and Sylvain Gelly · 1902
Earlier work this paper cites.
The harpy speech recognition system: performance with large vocabularies
Bruce Lowerre and R Reddy · 1976
Earlier work this paper cites.
On data banks and privacy homomorphisms
Ronald L Rivest, Len Adleman, Michael L Dertouzos, et al · 1978
Earlier work this paper cites.
An introduction to hidden markov models
L. Rabiner and B. Juang · 1986
Earlier work this paper cites.
The dragon continuous speech recognition system: A real-time implementation
Paul G. Bamberg, Yen lu Chow, Larry Gillick, Robert Roth, and Dean G. Sturtevant · 1990
Earlier work this paper cites.
1. 0 TANGORA - a large vocabulary speech recognition system for five languages
Helene Cerf-Danon, Steven DeGennaro, Marco Ferretti, Jorge Gonzalez, and Eric Keppel · 1991
Earlier work this paper cites.
The linux operating system
Sayed Naem Bokhari · 1995
Earlier work this paper cites.
Medspeak: Report creation with continuous speech recognition
Jennifer Lai and John Vergo · 1997
Earlier work this paper cites.
Multi party computations: past and present
Shafi Goldwasser · 1997
Earlier work this paper cites.
Locality-sensitive hashing scheme based on p-stable distributions
Mayur Datar, Nicole Immorlica, Piotr Indyk, and Vahab S. Mirrokni · 2004
Earlier work this paper cites.
Linformer: Self-attention with linear complexity
Sinong Wang, Belinda Z. Li, Madian Khabsa, Han Fang, and Hao Ma · 2006
Earlier work this paper cites.
‘outlines of a world coming into existence’: pervasive computing and the ethics of forgetting
Martin Dodge and Rob Kitchin · 2007
Earlier work this paper cites.
Random projection trees and low dimensional manifolds
Sanjoy Dasgupta and Yoav Freund · 2008
Earlier work this paper cites.
Environmental sound recognition with time–frequency audio features
Selina Chu, Shrikanth Narayanan, and C-C Jay Kuo · 2009
Earlier work this paper cites.
Fully homomorphic encryption using ideal lattices
Craig Gentry · 2009
Earlier work this paper cites.
Guide to protecting the confidentiality of personally identifiable information , volume 800
Erika McCallister · 2010
Earlier work this paper cites.
Pomdp-based statistical spoken dialog systems: A review
Steve Young, Milica Gašić, Blaise Thomson, and Jason D. Williams · 2012
Earlier work this paper cites.
User-driven access control: Rethinking permission granting in modern operating systems
Franziska Roesner, Tadayoshi Kohno, Alexander Moshchuk, Bryan Parno, Helen J Wang, and Crispin Cowan · 2012
Earlier work this paper cites.
Guoguo: Enabling fine-grained indoor localization via smartphone
Kaikai Liu, Xinxin Liu, and Xiaolin Li · 2013
Earlier work this paper cites.
Efficient estimation of word representations in vector space, 2013
Tomas Mikolov, Kai Chen, Greg Corrado, and Jeffrey Dean · 2013
Earlier work this paper cites.
Randomized partition trees for exact nearest neighbor search, 2013
Sanjoy Dasgupta and Kaushik Sinha · 2013
Earlier work this paper cites.
Activity recognition with smartphone sensors
Xing Su, Hanghang Tong, and Ping Ji · 2014
Earlier work this paper cites.
Predictors of life satisfaction based on daily activities from mobile sensor data
Onur Yürüten, Jiyong Zhang, and Pearl HZ Pu · 2014
Earlier work this paper cites.
Lifelogging: Personal big data
Cathal Gurrin, Alan F Smeaton, Aiden R Doherty, et al · 2014
Earlier work this paper cites.
Distributed representations of sentences and documents
Quoc Le and Tomas Mikolov · 2014
Earlier work this paper cites.
Approximate nearest neighbor algorithm based on navigable small world graphs
Yury Malkov, Alexander Ponomarenko, Andrey Logvinov, and Vladimir Krylov · 2014
Earlier work this paper cites.
Taintdroid: an information-flow tracking system for realtime privacy monitoring on smartphones
William Enck, Peter Gilbert, Seungyeop Han, Vasant Tendulkar, Byung-Gon Chun, Landon P Cox, Jaeyeon Jung, Patrick McDaniel, and Anmol N Sheth · 2014
Earlier work this paper cites.
Intriguing properties of neural networks, 2014
Christian Szegedy, Wojciech Zaremba, Ilya Sutskever, Joan Bruna, Dumitru Erhan, Ian Goodfellow, and Rob Fergus · 2014
Earlier work this paper cites.
Smartgpa: how smartphones can assess and predict academic performance of college students
Rui Wang, Gabriella Harari, Peilin Hao, Xia Zhou, and Andrew T Campbell · 2015
Earlier work this paper cites.
Ulink: Enabling user-defined deep linking to app content
Tanzirul Azim, Oriana Riva, and Suman Nath · 2016
Earlier work this paper cites.
"like having a really bad pa": The gulf between user expectation and experience of conversational agents
Ewa Luger and Abigail Sellen · 2016
Earlier work this paper cites.
Identifying user habits through data mining on call data records
Filippo Maria Bianchi, Antonello Rizzi, Alireza Sadeghian, and Corrado Moiso · 2016
Earlier work this paper cites.
An algorithm to transform natural language into sql queries for relational databases
Garima Singh and Arun Solanki · 2016
Earlier work this paper cites.
Cryptonets: Applying neural networks to encrypted data with high throughput and accuracy
Ran Gilad-Bachrach, Nathan Dowlin, Kim Laine, Kristin Lauter, Michael Naehrig, and John Wernsing · 2016
Earlier work this paper cites.
Programming iot devices by demonstration using mobile apps
Toby Jia-Jun Li, Yuanchun Li, Fanglin Chen, and Brad A Myers · 2017
Earlier work this paper cites.
"what can i help you with?": Infrequent users’ experiences of intelligent personal assistants
Benjamin R. Cowan, Nadia Pantidi, David Coyle, Kellie Morrissey, Peter Clarke, Sara Al-Shehri, David Earley, and Natasha Bandeira · 2017
Earlier work this paper cites.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N. Gomez, Łukasz Kaiser, and Illia Polosukhin · 2017
Earlier work this paper cites.
World of bits: An open-domain platform for web-based agents
Tianlin Tim Shi, Andrej Karpathy, Linxi Jim Fan, Jonathan Hernandez, and Percy Liang · 2017
Earlier work this paper cites.
Sugilite: Creating multimodal smartphone automation by demonstration
Toby Jia-Jun Li, Amos Azaria, and Brad A. Myers · 2017
Earlier work this paper cites.
Deep learning-based document modeling for personality detection from text
Navonil Majumder, Soujanya Poria, Alexander Gelbukh, and Erik Cambria · 2017
Earlier work this paper cites.
Reinforcement learning on web interfaces using workflow-guided exploration
Evan Zheran Liu, Kelvin Guu, Panupong Pasupat, Tianlin Shi, and Percy Liang · 2018
Earlier work this paper cites.
Alexa, siri, cortana, and more: An introduction to voice assistants
Matthew B. Hoy · 2018
Earlier work this paper cites.
Izzeddin Gur, Ulrich Rückert, Aleksandra Faust, and Dilek Z. Hakkani-Tür · 2018
Earlier work this paper cites.
Multi-task learning for joint language understanding and dialogue state tracking
Abhinav Rastogi, Raghav Gupta, and Dilek Hakkani-Tur · 2018
Earlier work this paper cites.
Kite: Building conversational bots from mobile apps
Toby Jia-Jun Li and Oriana Riva · 2018
Earlier work this paper cites.
Mapping natural language commands to web elements
Panupong Pasupat, Tian-Shun Jiang, Evan Liu, Kelvin Guu, and Percy Liang · 2018
Earlier work this paper cites.
Moodexplorer: Towards compound emotion detection via smartphone sensing
Xiao Zhang, Wenzhong Li, Xu Chen, and Sanglu Lu · 2018
Earlier work this paper cites.
Automated extraction of personal knowledge from smartphone push notifications
Yuanchun Li, Ziyue Yang, Yao Guo, Xiangqun Chen, Yuvraj Agarwal, and Jason I Hong · 2018
Earlier work this paper cites.
Efficient and robust approximate nearest neighbor search using hierarchical navigable small world graphs
Yu A. Malkov and D. A. Yashunin · 2018
Earlier work this paper cites.
A survey on homomorphic encryption schemes: Theory and implementation
Abbas Acar, Hidayet Aksu, A Selcuk Uluagac, and Mauro Conti · 2018
Earlier work this paper cites.
Slalom: Fast, verifiable and private execution of neural networks in trusted hardware
Florian Tramer and Dan Boneh · 2018
Earlier work this paper cites.
Privacy-preserving neural representations of text
Maximin Coavoux, Shashi Narayan, and Shay B Cohen · 2018
Earlier work this paper cites.
Humanoid: A deep learning-based approach to automated black-box android app testing
Yuanchun Li, Ziyue Yang, Yao Guo, and Xiangqun Chen · 2019
Earlier work this paper cites.
Dom-q-net: Grounded rl on structured language
Sheng Jia, Jamie Ryan Kiros, and Jimmy Ba · 2019
Earlier work this paper cites.
Environmental audio scene and sound event recognition for autonomous surveillance: A survey and comparative studies
S Chandrakala and SL Jayalakshmi · 2019
Earlier work this paper cites.
Predicting personality traits from physical activity intensity
Nan Gao, Wei Shao, and Flora D Salim · 2019
Earlier work this paper cites.
Akupm: Attention-enhanced knowledge-aware user preference model for recommendation
Xiaoli Tang, Tengyun Wang, Haizhi Yang, and Hengjie Song · 2019
Earlier work this paper cites.
Grammar-based neural text-to-sql generation
Kevin Lin, Ben Bogin, Mark Neumann, Jonathan Berant, and Matt Gardner · 2019
Earlier work this paper cites.
Diskann: Fast accurate billion-point nearest neighbor search on a single node
Suhas Jayaram Subramanya, Fnu Devvrit, Harsha Vardhan Simhadri, Ravishankar Krishnawamy, and Rohan Kadekodi · 2019
Earlier work this paper cites.
Billion-scale similarity search with GPUs
Jeff Johnson, Matthijs Douze, and Hervé Jégou · 2019
Earlier work this paper cites.
Quicker adc: Unlocking the hidden potential of product quantization with simd
Fabien Andre, Anne-Marie Kermarrec, and Nicolas Le Scouarnec · 2019
Earlier work this paper cites.
Ggnn: Graph-based gpu nearest neighbor search
Fabian Groh, Lukas Ruppert, Patrick Wieschollek, and Hendrik P. A. Lensch · 2019
Earlier work this paper cites.
Badnets: Evaluating backdooring attacks on deep neural networks
Tianyu Gu, Kang Liu, Brendan Dolan-Gavitt, and Siddharth Garg · 2019
Earlier work this paper cites.
Bert: Pre-training of deep bidirectional transformers for language understanding
Jacob Devlin Ming-Wei Chang Kenton and Lee Kristina Toutanova · 2019
Earlier work this paper cites.
Cgmh: Constrained sentence generation by metropolis-hastings sampling
Ning Miao, Hao Zhou, Lili Mou, Rui Yan, and Lei Li · 2019
Earlier work this paper cites.
Mapping natural language instructions to mobile ui action sequences, 2020
Yang Li, Jiacong He, Xin Zhou, Yuan Zhang, and Jason Baldridge · 2020
Earlier work this paper cites.
Actionbert: Leveraging user actions for semantic understanding of user interfaces
Zecheng He, Srinivas Sunkara, Xiaoxue Zang, Ying Xu, Lijuan Liu, Nevan Wichers, Gabriel Schubiner, Ruby B. Lee, and Jindong Chen · 2020
Earlier work this paper cites.
Scaling laws for neural language models, 2020
Jared Kaplan, Sam McCandlish, Tom Henighan, Tom B. Brown, Benjamin Chess, Rewon Child, Scott Gray, Alec Radford, Jeffrey Wu, and Dario Amodei · 2020
Earlier work this paper cites.
Language models are few-shot learners, 2020
Tom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, Sandhini Agarwal, Ariel Herbert-Voss, Gretchen Krueger, Tom Henighan, Rewon Child, Aditya Ramesh, Daniel M. Ziegler, Jeffrey Wu, Clemens Winter, Christopher Hesse, Mark Chen, Eric Sigler, Mateusz Litwin, Scott Gray, Benjamin Chess, Jack Clark, Christopher Berner, Sam McCandlish, Alec Radford, Ilya Sutskever, and Dario Amodei · 2020
Earlier work this paper cites.
A multi-sensor approach to automatically recognize breaks and work activities of knowledge workers in academia
Elena Di Lascio, Shkurta Gashi, Juan Sebastian Hidalgo, Beatrice Nale, Maike E Debus, and Silvia Santini · 2020
Earlier work this paper cites.
Predicting subjective measures of social anxiety from sparsely collected mobile sensor data
Haroon Rashid, Sanjana Mendu, Katharine E Daniel, Miranda L Beltzer, Bethany A Teachman, Mehdi Boukhechba, and Laura E Barnes · 2020
Earlier work this paper cites.
Healthwalks: Sensing fine-grained individual health condition via mobility data
Zongyu Lin, Shiqing Lyu, Hancheng Cao, Fengli Xu, Yuqiong Wei, Hanan Samet, and Yong Li · 2020
Earlier work this paper cites.
Detecting job promotion in information workers using mobile sensing
Subigya Nepal, Shayan Mirjafari, Gonzalo J Martinez, Pino Audia, Aaron Striegel, and Andrew T Campbell · 2020
Earlier work this paper cites.
Social sensing: assessing social functioning of patients living with schizophrenia using mobile phone sensing
Weichen Wang, Shayan Mirjafari, Gabriella Harari, Dror Ben-Zeev, Rachel Brian, Tanzeem Choudhury, Marta Hauser, John Kane, Kizito Masaba, Subigya Nepal, et al · 2020
Earlier work this paper cites.
Smokingopp: Detecting the smoking’opportunity’context using mobile sensors
Soujanya Chatterjee, Alexander Moreno, Steven Lloyd Lizotte, Sayma Akther, Emre Ertin, Christopher P Fagundes, Cho Lam, James M Rehg, Neng Wan, David W Wetter, et al · 2020
Earlier work this paper cites.
Vision-based human activity recognition: a survey
Djamila Romaissa Beddiar, Brahim Nini, Mohammad Sabokrou, and Abdenour Hadid · 2020
Earlier work this paper cites.
Predicting personality from patterns of behavior collected with smartphones
Clemens Stachl, Quay Au, Ramona Schoedel, Samuel D Gosling, Gabriella M Harari, Daniel Buschek, Sarah Theres Völkel, Tobias Schuwerk, Michelle Oldemeier, Theresa Ullmann, et al · 2020
Earlier work this paper cites.
A survey of automatic personality detection from texts
Sanja Štajner and Seren Yenikent · 2020
Earlier work this paper cites.
Facial emotion detection using deep learning
Akriti Jaiswal, A Krishnama Raju, and Suman Deb · 2020
Earlier work this paper cites.
Keep calm and explore: Language models for action generation in text-based games
Shunyu Yao, Rohan Rao, Matthew Hausknecht, and Karthik Narasimhan · 2020
Earlier work this paper cites.
Retrieval-augmented generation for knowledge-intensive nlp tasks
Patrick Lewis, Ethan Perez, Aleksandra Piktus, Fabio Petroni, Vladimir Karpukhin, Naman Goyal, Heinrich Küttler, Mike Lewis, Wen-tau Yih, Tim Rocktäschel, et al · 2020
Earlier work this paper cites.
Analyticdb-v: A hybrid analytical engine towards query fusion for structured and unstructured data
Chuangxian Wei, Bin Wu, Sheng Wang, Renjie Lou, Chaoqun Zhan, Feifei Li, and Yuanzhe Cai · 2020
Earlier work this paper cites.
Song: Approximate nearest neighbor search on gpu
Weijie Zhao, Shulong Tan, and Ping Li · 2020
Earlier work this paper cites.
Adversarial training for large neural language models
Xiaodong Liu, Hao Cheng, Pengcheng He, Weizhu Chen, Yu Wang, Hoifung Poon, and Jianfeng Gao · 2020
Earlier work this paper cites.
Adversarial attacks and defenses in images, graphs and text: A review
Han Xu, Yao Ma, Hao-Chen Liu, Debayan Deb, Hui Liu, Ji-Liang Tang, and Anil K. Jain · 2020
Earlier work this paper cites.
Don’t stop pretraining: Adapt language models to domains and tasks
Suchin Gururangan, Ana Marasović, Swabha Swayamdipta, Kyle Lo, Iz Beltagy, Doug Downey, and Noah A. Smith · 2020
Earlier work this paper cites.
Retrieval augmented language model pre-training
Kelvin Guu, Kenton Lee, Zora Tung, Panupong Pasupat, and Mingwei Chang · 2020
Earlier work this paper cites.
A survey of the state of explainable ai for natural language processing
Marina Danilevsky, Kun Qian, Ranit Aharonov, Yannis Katsis, Ban Kawas, and Prithviraj Sen · 2020
Earlier work this paper cites.
Glider: A reinforcement learning approach to extract ui scripts from websites
Yuanchun Li and Oriana Riva · 2021
Earlier work this paper cites.
Understanding mobile gui: from pixel-words to screen-sentences
Jingwen Fu, Xiaoyi Zhang, Yuwang Wang, Wenjun Zeng, Sam Yang, and Grayson Hilliard · 2021
Earlier work this paper cites.
Vut: Versatile ui transformer for multi-modal multi-task user interface modeling
Yang Li, Gang Li, Xin Zhou, Mostafa Dehghani, and Alexey A. Gritsenko · 2021
Earlier work this paper cites.
Uibert: Learning generic multimodal representations for ui understanding
Chongyang Bai, Xiaoxue Zang, Ying Xu, Srinivas Sunkara, Abhinav Rastogi, Jindong Chen, and Blaise Agüera y Arcas · 2021
Earlier work this paper cites.
Actionbert: Leveraging user actions for semantic understanding of user interfaces
Zecheng He, Srinivas Sunkara, Xiaoxue Zang, Ying Xu, Lijuan Liu, Nevan Wichers, Gabriel Schubiner, Ruby Lee, and Jindong Chen · 2021
Earlier work this paper cites.
Androidenv: A reinforcement learning platform for android
Daniel Toyama, Philippe Hamel, Anita Gergely, Gheorghe Comanici, Amelia Glaese, Zafarali Ahmed, Tyler Jackson, Shibl Mourad, and Doina Precup · 2021
Earlier work this paper cites.
Robust inertial motion tracking through deep sensor fusion across smart earbuds and smartphone
Jian Gong, Xinyu Zhang, Yuanjun Huang, Ju Ren, and Yaoxue Zhang · 2021
Earlier work this paper cites.
Attend and discriminate: Beyond the state-of-the-art for human activity recognition using wearable sensors
Alireza Abedin, Mahsa Ehsanpour, Qinfeng Shi, Hamid Rezatofighi, and Damith C Ranasinghe · 2021
Earlier work this paper cites.
mteeth: Identifying brushing teeth surfaces using wrist-worn inertial sensors
Sayma Akther, Nazir Saleheen, Mithun Saha, Vivek Shetty, and Santosh Kumar · 2021
Earlier work this paper cites.
Identifying mobile sensing indicators of stress-resilience
Daniel A Adler, Vincent W-S Tseng, Gengmo Qi, Joseph Scarpa, Srijan Sen, and Tanzeem Choudhury · 2021
Earlier work this paper cites.
Emotion detection of textual data: An interdisciplinary survey
Samira Zad, Maryam Heidari, H James Jr, and Ozlem Uzuner · 2021
Earlier work this paper cites.
Show your work: Scratchpads for intermediate computation with language models
Maxwell Nye, Anders Johan Andreassen, Guy Gur-Ari, Henryk Michalewski, Jacob Austin, David Bieber, David Dohan, Aitor Lewkowycz, Maarten Bosma, David Luan, et al · 2021
Earlier work this paper cites.
Grounding language to entities and dynamics for generalization in reinforcement learning
Austin W Hanjie, Victor Y Zhong, and Karthik Narasimhan · 2021
Earlier work this paper cites.
Lifelong pretraining: Continually adapting language models to emerging corpora
Xisen Jin, Dejiao Zhang, Henghui Zhu, Wei Xiao, Shang-Wen Li, Xiaokai Wei, Andrew Arnold, and Xiang Ren · 2021
Earlier work this paper cites.
Continual learning for named entity recognition
Natawut Monaikul, Giuseppe Castellucci, Simone Filice, and Oleg Rokhlenko · 2021
Earlier work this paper cites.
SPANN: Highly-efficient billion-scale approximate nearest neighborhood search
Qi Chen, Bing Zhao, Haidong Wang, Mingqin Li, Chuanjie Liu, Zengzhong Li, Mao Yang, and Jingdong Wang · 2021
Earlier work this paper cites.
Milvus: A purpose-built vector data management system
Jianguo Wang, Xiaomeng Yi, Rentong Guo, Hai Jin, Peng Xu, Shengjun Li, Xiangyu Wang, Xiangzhou Guo, Chengming Li, Xiaohai Xu, Kun Yu, Yuxing Yuan, Yinghao Zou, Jiquan Long, Yudong Cai, Zhenxiang Li, Zhifeng Zhang, Yihua Mo, Jun Gu, Ruiyi Jiang, Yi Wei, and Charles Xie · 2021
Earlier work this paper cites.
Understanding and overcoming the challenges of efficient transformer quantization
Yelysei Bondarenko, Markus Nagel, and Tijmen Blankevoort · 2021
Earlier work this paper cites.
Irene: Interpretable energy prediction for transformers
Qingqing Cao, Yash Kumar Lal, H. Trivedi, Aruna Balasubramanian, and Niranjan Balasubramanian · 2021
Earlier work this paper cites.
Prefix-tuning: Optimizing continuous prompts for generation, 2021
Xiang Lisa Li and Percy Liang · 2021
Earlier work this paper cites.
Cheetah: Optimizing and accelerating homomorphic encryption for private inference
Brandon Reagen, Woo-Seok Choi, Yeongil Ko, Vincent T Lee, Hsien-Hsin S Lee, Gu-Yeon Wei, and David Brooks · 2021
Cited alongside, same era.
Crypten: Secure multi-party computation meets machine learning
Brian Knott, Shobha Venkataraman, Awni Hannun, Shubho Sengupta, Mark Ibrahim, and Laurens van der Maaten · 2021
Cited alongside, same era.
Security vulnerabilities of sgx and countermeasures: A survey
Shufan Fei, Zheng Yan, Wenxiu Ding, and Haomeng Xie · 2021
Cited alongside, same era.
Extracting training data from large language models
Nicholas Carlini, Florian Tramèr, Eric Wallace, Matthew Jagielski, Ariel Herbert-Voss, Katherine Lee, Adam Roberts, Tom Brown, Dawn Song, Úlfar Erlingsson, Alina Oprea, and Colin Raffel · 2021
Cited alongside, same era.
Finetuned language models are zero-shot learners
Jason Wei, Maarten Bosma, Vincent Zhao, Kelvin Guu, Adams Wei Yu, Brian Lester, Nan Du, Andrew M Dai, and Quoc V Le · 2021
Cited alongside, same era.
Palr: Personalization aware llms for recommendation
Zheng Chen · 2023
Later among the works it cites.
Conversational health agents: A personalized llm-powered agent framework
Mahyar Abbasian, Iman Azimi, Amir M Rahmani, and Ramesh Jain · 2023
Later among the works it cites.
Generative agents: Interactive simulacra of human behavior
Joon Sung Park, Joseph O’Brien, Carrie Jun Cai, Meredith Ringel Morris, Percy Liang, and Michael S Bernstein · 2023
Later among the works it cites.
Mot: Memory-of-thought enables chatgpt to self-improve
Xiaonan Li and Xipeng Qiu · 2023
Later among the works it cites.
Prompt-guided retrieval augmentation for non-knowledge-intensive tasks
Zhicheng Guo, Sijie Cheng, Yile Wang, Peng Li, and Yang Liu · 2023
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Sarah Wiegreffe and Ana Marasović · 2021
Cited alongside, same era.
Spotlight: Mobile ui understanding using vision-language models with a focus
Gang Li and Yang Li · 2022
Cited alongside, same era.
A data-driven approach for learning to control computers
Peter C Humphreys, David Raposo, Tobias Pohlen, Gregory Thornton, Rachita Chhaparia, Alistair Muldal, Josh Abramson, Petko Georgiev, Adam Santoro, and Timothy Lillicrap · 2022
Cited alongside, same era.
Training language models to follow instructions with human feedback, 2022
Long Ouyang, Jeff Wu, Xu Jiang, Diogo Almeida, Carroll L. Wainwright, Pamela Mishkin, Chong Zhang, Sandhini Agarwal, Katarina Slama, Alex Ray, John Schulman, Jacob Hilton, Fraser Kelton, Luke Miller, Maddie Simens, Amanda Askell, Peter Welinder, Paul Christiano, Jan Leike, and Ryan Lowe · 2022
Cited alongside, same era.
Webgpt: Browser-assisted question-answering with human feedback, 2022
Reiichiro Nakano, Jacob Hilton, Suchir Balaji, Jeff Wu, Long Ouyang, Christina Kim, Christopher Hesse, Shantanu Jain, Vineet Kosaraju, William Saunders, Xu Jiang, Karl Cobbe, Tyna Eloundou, Gretchen Krueger, Kevin Button, Matthew Knight, Benjamin Chess, and John Schulman · 2022
Cited alongside, same era.
Recent advances in end-to-end automatic speech recognition, 2022
Jinyu Li · 2022
Cited alongside, same era.
Mrkl systems: A modular, neuro-symbolic architecture that combines large language models, external knowledge sources and discrete reasoning, 2022
Ehud Karpas, Omri Abend, Yonatan Belinkov, Barak Lenz, Opher Lieber, Nir Ratner, Yoav Shoham, Hofit Bata, Yoav Levine, Kevin Leyton-Brown, Dor Muhlgay, Noam Rozen, Erez Schwartz, Gal Shachaf, Shai Shalev-Shwartz, Amnon Shashua, and Moshe Tenenholtz · 2022
Cited alongside, same era.
Cognitive architectures for language agents
Theodore Sumers, Shunyu Yao, Karthik Narasimhan, and Thomas L Griffiths · 2023
Later among the works it cites.
Baolin Peng, Michel Galley, Pengcheng He, Hao Cheng, Yujia Xie, Yu Hu, Qiuyuan Huang, Lars Liden, Zhou Yu, Weizhu Chen, et al · 2023
Later among the works it cites.
Human-assisted continual robot learning with foundation models
Meenal Parakh, Alisha Fong, Anthony Simeonov, Abhishek Gupta, Tao Chen, and Pulkit Agrawal · 2023
Later among the works it cites.
Dreamcoder: growing generalizable, interpretable knowledge with wake–sleep bayesian program learning
Kevin Ellis, Lionel Wong, Maxwell Nye, Mathias Sable-Meyer, Luc Cary, Lore Anaya Pozo, Luke Hewitt, Armando Solar-Lezama, and Joshua B Tenenbaum · 2023
Later among the works it cites.
Awq: Activation-aware weight quantization for llm compression and acceleration
Ji Lin, Jiaming Tang, Haotian Tang, Shang Yang, Xingyu Dang, and Song Han · 2023
Later among the works it cites.
Smoothquant: Accurate and efficient post-training quantization for large language models
Guangxuan Xiao, Ji Lin, Mickael Seznec, Hao Wu, Julien Demouth, and Song Han · 2023
Later among the works it cites.
Llm-pruner: On the structural pruning of large language models
Xinyin Ma, Gongfan Fang, and Xinchao Wang · 2023
Later among the works it cites.
Sparsegpt: Massive language models can be accurately pruned in one-shot
Elias Frantar and Dan Alistarh · 2023
Later among the works it cites.
Inar Timiryasov and Jean-Loup Tastet · 2023
Later among the works it cites.
Knowledge distillation of large language models
Yuxian Gu, Li Dong, Furu Wei, and Minlie Huang · 2023
Later among the works it cites.
Cheng-Yu Hsieh, Chun-Liang Li, Chih-Kuan Yeh, Hootan Nakhost, Yasuhisa Fujii, Alexander Ratner, Ranjay Krishna, Chen-Yu Lee, and Tomas Pfister · 2023
Later among the works it cites.
Llmlingua: Compressing prompts for accelerated inference of large language models
Huiqiang Jiang, Qianhui Wu, Chin-Yew Lin, Yuqing Yang, and Lili Qiu · 2023
Later among the works it cites.
Adapting language models to compress contexts
Alexis Chevalier, Alexander Wettig, Anirudh Ajith, and Danqi Chen · 2023
Later among the works it cites.
Dynamic context pruning for efficient and interpretable autoregressive transformers
Sotiris Anagnostidis, Dario Pavllo, Luca Biggio, Lorenzo Noci, Aurelien Lucchi, and Thomas Hoffmann · 2023
Later among the works it cites.
Flashattention-2: Faster attention with better parallelism and work partitioning
Tri Dao · 2023
Later among the works it cites.
Fast inference from transformers via speculative decoding
Yaniv Leviathan, Matan Kalman, and Yossi Matias · 2023
Later among the works it cites.
Flexgen: High-throughput generative inference of large language models with a single gpu, 2023
Ying Sheng, Lianmin Zheng, Binhang Yuan, Zhuohan Li, Max Ryabinin, Daniel Y. Fu, Zhiqiang Xie, Beidi Chen, Clark Barrett, Joseph E. Gonzalez, Percy Liang, Christopher Ré, Ion Stoica, and Ce Zhang · 2023
Later among the works it cites.
Powerinfer: Fast large language model serving with a consumer-grade gpu
Yixin Song, Zeyu Mi, Haotong Xie, and Haibo Chen · 2023
Later among the works it cites.
Llm in a flash: Efficient large language model inference with limited memory
Keivan Alizadeh, Iman Mirzadeh, Dmitry Belenko, Karen Khatamifard, Minsik Cho, Carlo C Del Mundo, Mohammad Rastegari, and Mehrdad Farajtabar · 2023
Later among the works it cites.
Snapdragon 8 gen 3 mobile platform
Qualcomm · 2023
Later among the works it cites.
Efficient deployment of transformer models on edge tpu accelerators: A real system evaluation
Brendan C Reidy, Mohammadreza Mohammadi, Mohammed E Elbtity, and Ramtin Zand · 2023
Later among the works it cites.
Full parameter fine-tuning for large language models with limited resources, 2023
Kai Lv, Yuqing Yang, Tengxiao Liu, Qinghui Gao, Qipeng Guo, and Xipeng Qiu · 2023
Later among the works it cites.
Textbooks are all you need, 2023
Suriya Gunasekar, Yi Zhang, Jyoti Aneja, Caio César Teodoro Mendes, Allie Del Giorno, Sivakanth Gopi, Mojan Javaheripi, Piero Kauffmann, Gustavo de Rosa, Olli Saarikivi, Adil Salim, Shital Shah, Harkirat Singh Behl, Xin Wang, Sébastien Bubeck, Ronen Eldan, Adam Tauman Kalai, Yin Tat Lee, and Yuanzhi Li · 2023
Later among the works it cites.
Phi-2: The surprising power of small language models
Mojan Javaheripi and Sébastien Bubeck · 2023
Later among the works it cites.
Cxl-anns: Software-hardware collaborative memory disaggregation and computation for billion-scale approximate nearest neighbor search
Junhyeok Jang, Hanjin Choi, Hanyeoreum Bae, Seungjun Lee, Miryeong Kwon, and Myoungsoo Jung · 2023
Later among the works it cites.
ggerganov/llama.cpp: Port of facebook’s llama model in c/c++
llama.cpp developers · 2023
Later among the works it cites.
MLC-LLM, 2023
MLC team · 2023
Later among the works it cites.
Efficient memory management for large language model serving with pagedattention
Woosuk Kwon, Zhuohan Li, Siyuan Zhuang, Ying Sheng, Lianmin Zheng, Cody Hao Yu, Joseph Gonzalez, Hao Zhang, and Ion Stoica · 2023
Later among the works it cites.
Colt5: Faster long-range transformers with conditional computation
Joshua Ainslie, Tao Lei, Michiel de Jong, Santiago Ontan’on, Siddhartha Brahma, Yury Zemlyanskiy, David C. Uthus, Mandy Guo, James Lee-Thorp, Yi Tay, Yun-Hsuan Sung, and Sumit K. Sanghai · 2023
Later among the works it cites.
Skipdecode: Autoregressive skip decoding with batching and caching for efficient llm inference
Luciano Del Corro, Allie Del Giorno, Sahaj Agarwal, Bin Yu, Ahmed Awadallah, and Subhabrata Mukherjee · 2023
Later among the works it cites.
Xupeng Miao, Gabriele Oliaro, Zhihao Zhang, Xinhao Cheng, Zeyu Wang, Rae Ying Yee Wong, Zhuoming Chen, Daiyaan Arfeen, Reyna Abhyankar, and Zhihao Jia · 2023
Later among the works it cites.
Accelerating llm inference with staged speculative decoding
Benjamin Spector and Chris Re · 2023
Later among the works it cites.
Accelerating attention mechanism on fpgas based on efficient reconfigurable systolic array
Wenhua Ye, Xu Zhou, Joey Zhou, Cen Chen, and Kenli Li · 2023
Later among the works it cites.
From words to watts: Benchmarking the energy costs of large language model inference
Siddharth Samsi, Dan Zhao, Joseph McDonald, Baolin Li, Adam Michaleas, Michael Jones, William Bergeron, Jeremy Kepner, Devesh Tiwari, and Vijay Gadepally · 2023
Later among the works it cites.
Llmcarbon: Modeling the end-to-end carbon footprint of large language models
Ahmad Faiz, Sotaro Kaneda, Ruhan Wang, Rita Osi, Parteek Sharma, Fan Chen, and Lei Jiang · 2023
Later among the works it cites.
Prompt cache: Modular attention reuse for low-latency inference
In Gim, Guojun Chen, Seung seob Lee, Nikhil Sarda, Anurag Khandelwal, and Lin Zhong · 2023
Later among the works it cites.
Enhancing llm intelligence with arm-rag: Auxiliary rationale memory for retrieval augmented generation, 2023
Eric Melz · 2023
Later among the works it cites.
A comprehensive survey on vector database: Storage and retrieval technique, challenge
Yikun Han, Chunjiang Liu, and Pengfei Wang · 2023
Later among the works it cites.
Survey of vector database management systems, 2023
James Jie Pan, Jianguo Wang, and Guoliang Li · 2023
Later among the works it cites.
Vector database management systems: Fundamental concepts, use-cases, and current challenges
Toni Taipalus · 2023
Later among the works it cites.
Ret-llm: Towards a general read-write memory for large language models
Ali Modarressi, Ayyoob Imani, Mohsen Fayyaz, and Hinrich Schütze · 2023
Later among the works it cites.
Filtered-diskann: Graph algorithms for approximate nearest neighbor search with filters
Siddharth Gollapudi, Neel Karia, Varun Sivashankar, Ravishankar Krishnaswamy, Nikit Begwani, Swapnil Raz, Yiyong Lin, Yin Zhang, Neelam Mahapatro, Premkumar Srinivasan, Amit Singh, and Harsha Vardhan Simhadri · 2023
Later among the works it cites.
Approximate nearest neighbor search in high dimensional vector databases: Current research and future directions, 2023
Yao Tian, Ziyang Yue, Ruiyuan Zhang, Xi Zhao, Bolong Zheng, and Xiaofang Zhou · 2023
Later among the works it cites.
Jiongkang Ni, Xiaoliang Xu, Yuxiang Wang, Can Li, Jiajie Yao, Shihai Xiao, and Xuecang Zhang · 2023
Later among the works it cites.
Cagra: Highly parallel graph construction and approximate nearest neighbor search for gpus
Hiroyuki Ootomo, Akira Naruse, Corey J. Nolet, Ray Wang, Tamas B. Fehér, and Y. Wang · 2023
Later among the works it cites.
Llama: Open and efficient foundation language models, 2023
Hugo Touvron, Thibaut Lavril, Gautier Izacard, Xavier Martinet, Marie-Anne Lachaux, Timothée Lacroix, Baptiste Rozière, Naman Goyal, Eric Hambro, Faisal Azhar, Aurelien Rodriguez, Armand Joulin, Edouard Grave, and Guillaume Lample · 2023
Later among the works it cites.
Bluelm: An open multilingual 7b language model
BlueLM Team · 2023
Later among the works it cites.
Spqr: A sparse-quantized representation for near-lossless llm weight compression
Tim Dettmers, Ruslan Svirschevski, Vage Egiazarian, Denis Kuznedelev, Elias Frantar, Saleh Ashkboos, Alexander Borzunov, Torsten Hoefler, and Dan Alistarh · 2023
Later among the works it cites.
Textobfuscator: Making pre-trained language model a privacy protector via obfuscating word representations
Xin Zhou, Yi Lu, Ruotian Ma, Tao Gui, Yuran Wang, Yong Ding, Yibo Zhang, Qi Zhang, and Xuan-Jing Huang · 2023
Later among the works it cites.
Certifying llm safety against adversarial prompting, 2023
Aounon Kumar, Chirag Agarwal, Suraj Srinivas, Aaron Jiaxun Li, Soheil Feizi, and Himabindu Lakkaraju · 2023
Later among the works it cites.
On the adversarial robustness of multi-modal foundation models
Christian Schlarmann and Matthias Hein · 2023
Later among the works it cites.
Misusing tools in large language models with visual adversarial examples, 2023
Xiaohan Fu, Zihan Wang, Shuheng Li, Rajesh K. Gupta, Niloofar Mireshghallah, Taylor Berg-Kirkpatrick, and Earlence Fernandes · 2023
Later among the works it cites.
Backdoor attacks for in-context learning with language models
Nikhil Kandpal, Matthew Jagielski, Florian Tramèr, and Nicholas Carlini · 2023
Later among the works it cites.
Not what you’ve signed up for: Compromising real-world llm-integrated applications with indirect prompt injection
Sahar Abdelnabi, Kai Greshake, Shailesh Mishra, Christoph Endres, Thorsten Holz, and Mario Fritz · 2023
Later among the works it cites.
Jailbreak in pieces: Compositional adversarial attacks on multi-modal language models, 2023
Erfan Shayegani, Yue Dong, and Nael Abu-Ghazaleh · 2023
Later among the works it cites.
Jailbreaking black box large language models in twenty queries, 2023
Patrick Chao, Alexander Robey, Edgar Dobriban, Hamed Hassani, George J. Pappas, and Eric Wong · 2023
Later among the works it cites.
Smoothllm: Defending large language models against jailbreaking attacks, 2023
Alexander Robey, Eric Wong, Hamed Hassani, and George J. Pappas · 2023
Later among the works it cites.
Survey of hallucination in natural language generation
Ziwei Ji, Nayeon Lee, Rita Frieske, Tiezheng Yu, Dan Su, Yan Xu, Etsuko Ishii, Ye Jin Bang, Andrea Madotto, and Pascale Fung · 2023
Later among the works it cites.
A survey of hallucination in large foundation models
Vipula Rawte, Amit Sheth, and Amitava Das · 2023
Later among the works it cites.
Dera: enhancing large language model completions with dialog-enabled resolving agents
Varun Nair, Elliot Schumacher, Geoffrey Tso, and Anitha Kannan · 2023
Later among the works it cites.
Learning from mistakes makes llm better reasoner
Shengnan An, Zexiong Ma, Zeqi Lin, Nanning Zheng, Jian-Guang Lou, and Weizhu Chen · 2023
Later among the works it cites.
Self-refine: Iterative refinement with self-feedback
Aman Madaan, Niket Tandon, Prakhar Gupta, Skyler Hallinan, Luyu Gao, Sarah Wiegreffe, Uri Alon, Nouha Dziri, Shrimai Prabhumoye, Yiming Yang, et al · 2023
Later among the works it cites.
Reflexion: Language agents with verbal reinforcement learning, 2023
Noah Shinn, Federico Cassano, Edward Berman, Ashwin Gopinath, Karthik Narasimhan, and Shunyu Yao · 2023
Later among the works it cites.
Selfcheckgpt: Zero-resource black-box hallucination detection for generative large language models
Potsawee Manakul, Adian Liusie, and Mark JF Gales · 2023
Later among the works it cites.
Improving factuality and reasoning in language models through multiagent debate
Yilun Du, Shuang Li, Antonio Torralba, Joshua B Tenenbaum, and Igor Mordatch · 2023
Later among the works it cites.
Large language models can be easily distracted by irrelevant context
Freda Shi, Xinyun Chen, Kanishka Misra, Nathan Scales, David Dohan, Ed H Chi, Nathanael Schärli, and Denny Zhou · 2023
Later among the works it cites.
Chain-of-note: Enhancing robustness in retrieval-augmented language models
Wenhao Yu, Hongming Zhang, Xiaoman Pan, Kaixin Ma, Hongwei Wang, and Dong Yu · 2023
Later among the works it cites.
Self-rag: Learning to retrieve, generate, and critique through self-reflection
Akari Asai, Zeqiu Wu, Yizhong Wang, Avirup Sil, and Hannaneh Hajishirzi · 2023
Later among the works it cites.
Critic: Large language models can self-correct with tool-interactive critiquing
Zhibin Gou, Zhihong Shao, Yeyun Gong, Yelong Shen, Yujiu Yang, Nan Duan, and Weizhu Chen · 2023
Later among the works it cites.
Large language models are better reasoners with self-verification
Yixuan Weng, Minjun Zhu, Fei Xia, Bin Li, Shizhu He, Shengping Liu, Bin Sun, Kang Liu, and Jun Zhao · 2023
Later among the works it cites.
Rationalization for explainable nlp: A survey
Sai Gurrapu, Ajay Kulkarni, Lifu Huang, Ismini Lourentzou, and Feras A Batarseh · 2023
Later among the works it cites.
Overthinking the truth: Understanding how language models process false demonstrations
Danny Halawi, Jean-Stanislas Denain, and Jacob Steinhardt · 2023
Later among the works it cites.
Google is taking questions (spoken, via iphone)
John Markoff · 2024
Closest in time.
Octopus v2: On-device language model for super agent
Wei Chen and Zhiyuan Li · 2024
Closest in time.
Osworld: Benchmarking multimodal agents for open-ended tasks in real computer environments
Tianbao Xie, Danyang Zhang, Jixuan Chen, Xiaochuan Li, Siheng Zhao, Ruisheng Cao, Toh Jing Hua, Zhoujun Cheng, Dongchan Shin, Fangyu Lei, et al · 2024
Closest in time.
Seeclick: Harnessing gui grounding for advanced visual gui agents
Kanzhi Cheng, Qiushi Sun, Yougang Chu, Fangzhi Xu, Yantao Li, Jianbing Zhang, and Zhiyong Wu · 2024
Closest in time.
Ferret-ui: Grounded mobile ui understanding with multimodal llms
Keen You, Haotian Zhang, Eldon Schoop, Floris Weers, Amanda Swearngin, Jeffrey Nichols, Yinfei Yang, and Zhe Gan · 2024
Closest in time.
Raghav Kapoor, Yash Parag Butala, Melisa Russak, Jing Yu Koh, Kiran Kamble, Waseem Alshikh, and Ruslan Salakhutdinov · 2024
Closest in time.
Autowebglm: Bootstrap and reinforce a large language model-based web navigating agent
Hanyu Lai, Xiao Liu, Iat Long Iong, Shuntian Yao, Yuxuan Chen, Pengbo Shen, Hao Yu, Hanchen Zhang, Xiaohan Zhang, Yuxiao Dong, et al · 2024
Closest in time.
Screenagent: A vision language model-driven computer control agent
Runliang Niu, Jindong Li, Shiqi Wang, Yali Fu, Xiyu Hu, Xueyuan Leng, He Kong, Yi Chang, and Qi Wang · 2024
Closest in time.
Hargpt: Are llms zero-shot human activity recognizers?
Sijie Ji, Xinzhe Zheng, and Chenshu Wu · 2024
Closest in time.
Are you being tracked? discover the power of zero-shot trajectory tracing with llms!
Huanqi Yang, Sijie Ji, Rucheng Wu, and Weitao Xu · 2024
Closest in time.
Prompting multi-modal tokens to enhance end-to-end autonomous driving imitation learning with llms
Yiqun Duan, Qiang Zhang, and Renjing Xu · 2024
Closest in time.
A survey on large language model-based game agents
Sihao Hu, Tiansheng Huang, Fatih Ilhan, Selim Tekin, Gaowen Liu, Ramana Kompella, and Ling Liu · 2024
Closest in time.
Chattracer: Large language model powered real-time bluetooth device tracking system
Qijun Wang, Shichen Zhang, Kunzhe Song, and Huacheng Zeng · 2024
Closest in time.
Organa: A robotic assistant for automated chemistry experimentation and characterization
Kourosh Darvish, Marta Skreta, Yuchi Zhao, Naruki Yoshikawa, Sagnik Som, Miroslav Bogdanovic, Yang Cao, Han Hao, Haoping Xu, Alán Aspuru-Guzik, et al · 2024
Closest in time.
Health-llm: Large language models for health prediction via wearable sensor data
Yubin Kim, Xuhai Xu, Daniel McDuff, Cynthia Breazeal, and Hae Won Park · 2024
Closest in time.
Depression detection on social media with large language models
Xiaochong Lan, Yiming Cheng, Li Sheng, Chen Gao, and Yong Li · 2024
Closest in time.
Zita Lifelo, Huansheng Ning, and Sahraoui Dhelim · 2024
Closest in time.
Large language model for mental health: A systematic review
Zhijun Guo, Alvina Lai, Johan Hilge Thygesen, Joseph Farrington, Thomas Keen, and Kezhi Li · 2024
Closest in time.
Llmsense: Harnessing llms for high-level reasoning over spatiotemporal sensor traces
Xiaomin Ouyang and Mani Srivastava · 2024
Closest in time.
Model tells you what to discard: Adaptive kv cache compression for llms
Suyu Ge, Yunan Zhang, Liyuan Liu, Minjia Zhang, Jiawei Han, and Jianfeng Gao · 2024
Closest in time.
Piperag: Fast retrieval-augmented generation via algorithm-system co-design, 2024
Wenqi Jiang, Shuai Zhang, Boran Han, Jie Wang, Bernie Wang, and Tim Kraska · 2024
Closest in time.
Ragcache: Efficient knowledge caching for retrieval-augmented generation, 2024
Chao Jin, Zili Zhang, Xuanlin Jiang, Fangyue Liu, Xin Liu, Xuanzhe Liu, and Xin Jin · 2024
Closest in time.
Generative representational instruction tuning, 2024
Niklas Muennighoff, Hongjin Su, Liang Wang, Nan Yang, Furu Wei, Tao Yu, Amanpreet Singh, and Douwe Kiela · 2024
Closest in time.
Transformer-lite: High-efficiency deployment of large language models on mobile phone gpus
Luchang Li, Sheng Qian, Jie Lu, Lunxi Yuan, Rui Wang, and Qin Xie · 2024
Closest in time.
Towards greener llms: Bringing energy-efficiency to the forefront of llm inference
Jovan Stojkovic, Esha Choukse, Chaojie Zhang, Íñigo Goiri, and Josep Torrellas · 2024
Closest in time.
Melting point: Mobile evaluation of language transformers, 2024
Stefanos Laskaridis, Kleomenis Katevas, Lorenzo Minto, and Hamed Haddadi · 2024
Closest in time.
Second-order fine-tuning without pain for llms: A hessian informed zeroth-order optimizer
Yanjun Zhao, Sizhe Dang, Haishan Ye, Guang Dai, Yi Qian, and Ivor Wai-Hung Tsang · 2024
Closest in time.
Promptcrypt: Prompt encryption for secure communication with large language models
Guo Lin, Wenyue Hua, and Yongfeng Zhang · 2024
Closest in time.
Whispers in the machine: Confidentiality in llm-integrated systems
Jonathan Evertz, Merlin Chlosta, Lea Schönherr, and Thorsten Eisenhofer · 2024
Closest in time.
On the effectiveness of distillation in mitigating backdoors in pre-trained encoder, 2024
Tingxu Han, Shenghan Huang, Ziqi Ding, Weisong Sun, Yebo Feng, Chunrong Fang, Jun Li, Hanwei Qian, Cong Wu, Quanjun Zhang, Yang Liu, and Zhenyu Chen · 2024
Closest in time.