Fetching the paper…
Reading the bibliography…
In this work, we conceptualize the learning process as information compression.
A mathematical theory of communication
Claude Elwood Shannon · 1948
Earlier work this paper cites.
Arithmetic coding for data compression
Ian H. Witten, Radford M. Neal, and John G. Cleary · 1987
Earlier work this paper cites.
Okapi at trec-3
Stephen Robertson, S. Walker, S. Jones, M. M. Hancock-Beaulieu, and M. Gatford · 1995
Earlier work this paper cites.
GZIP file format specification version 4.3
L. Peter Deutsch · 1996
Earlier work this paper cites.
Bidirectional recurrent neural networks
Mike Schuster and Kuldip K Paliwal · 1997
Earlier work this paper cites.
Information distance
Charles H Bennett, Péter Gács, Ming Li, Paul MB Vitányi, and Wojciech H Zurek · 1998
Earlier work this paper cites.
Relaxing the triangle inequality in pattern matching
R. Fagin and L. Stockmeyer · 1998
Earlier work this paper cites.
Elements of information theory
Thomas M Cover · 1999
Earlier work this paper cites.
An information-based sequence distance and its application to whole mitochondrial genome phylogeny
Ming Li, Jonathan H. Badger, Xin Chen, Sam Kwong, Paul Kearney, and Haoyong Zhang · 2001
Earlier work this paper cites.
Shape matching: similarity measures and algorithms
R.C. Veltkamp · 2001
Earlier work this paper cites.
Clustering by compression
Rudi Cilibrasi and Paul M. B. Vitányi · 2003
Earlier work this paper cites.
Shared information and program plagiarism detection
Xin Chen, Brent Francia, Ming Li, Brian Mckinnon, and Amit Seker · 2004
Earlier work this paper cites.
Towards parameter-free data mining
Eamonn Keogh, Stefano Lonardi, and Chotirat Ann Ratanamahatana · 2004
Earlier work this paper cites.
The similarity metric
Ming Li, Xin Chen, Xin Li, Bin Ma, and P.M.B. Vitanyi · 2004
Earlier work this paper cites.
Information distance and its applications
Ming Li · 2006
Earlier work this paper cites.
The complexity of finite objects and the development of the concepts of information and randomness by means of the theory of algorithms
A Zvonkin and L Levin · 2007
Earlier work this paper cites.
An Introduction to Kolmogorov Complexity and its Applications
Ming Li and Paul M. B. Vitányi · 2008
Earlier work this paper cites.
Recursive deep models for semantic compositionality over a sentiment treebank
Richard Socher, Alex Perelygin, Jean Wu, Jason Chuang, Christopher D. Manning, Andrew Ng, and Christopher Potts · 2013
Earlier work this paper cites.
Domain-adaptive discriminative one-shot learning of gestures
Tomas Pfister, James Charles, and Andrew Zisserman · 2014
Earlier work this paper cites.
Siamese neural networks for one-shot image recognition
Gregory Koch, Richard Zemel, Ruslan Salakhutdinov, et al · 2015
Earlier work this paper cites.
Variational inference with normalizing flows
Danilo Rezende and Shakir Mohamed · 2015
Earlier work this paper cites.
An overview of the bioasq large-scale biomedical semantic indexing and question answering competition
George Tsatsaronis, Georgios Balikas, Prodromos Malakasiotis, Ioannis Partalas, Matthias Zschunke, Michael R Alvers, Dirk Weissenborn, Anastasia Krithara, Sergios Petridis, Dimitris Polychronopoulos, et al · 2015
Earlier work this paper cites.
Character-level convolutional networks for text classification
Xiang Zhang, Junbo Jake Zhao, and Yann LeCun · 2015
Earlier work this paper cites.
Aligning books and movies: Towards story-like visual explanations by watching movies and reading books
Yukun Zhu, Ryan Kiros, Richard S. Zemel, Ruslan Salakhutdinov, Raquel Urtasun, Antonio Torralba, and Sanja Fidler · 2015
Earlier work this paper cites.
Ms marco: A human generated machine reading comprehension dataset
Daniel Fernando Campos, Tri Nguyen, Mir Rosenberg, Xia Song, Jianfeng Gao, Saurabh Tiwary, Rangan Majumder, Li Deng, and Bhaskar Mitra · 2016
Earlier work this paper cites.
Towards a neural statistician
Harrison Edwards and Amos Storkey · 2016
Earlier work this paper cites.
One-shot learning of scene locations via feature trajectory transfer
Roland Kwitt, Sebastian Hegenbart, and Marc Niethammer · 2016
Cited alongside, same era.
Matching networks for one shot learning
Oriol Vinyals, Charles Blundell, Timothy Lillicrap, Daan Wierstra, et al · 2016
Cited alongside, same era.
SemEval-2017 task 1: Semantic textual similarity multilingual and crosslingual focused evaluation
Daniel Cer, Mona Diab, Eneko Agirre, Iñigo Lopez-Gazpio, and Lucia Specia · 2017
Cited alongside, same era.
Model-agnostic meta-learning for fast adaptation of deep networks
Chelsea Finn, Pieter Abbeel, and Sergey Levine · 2017
Cited alongside, same era.
Prototypical networks for few-shot learning
Jake Snell, Kevin Swersky, and Richard Zemel · 2017
Cited alongside, same era.
Attention is all you need
Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N Gomez, Łukasz Kaiser, and Illia Polosukhin · 2017
Document ranking with a pretrained sequence-to-sequence model
Rodrigo Nogueira, Zhiying Jiang, Ronak Pradeep, and Jimmy Lin · 2020
Later among the works it cites.
Wikipedia citations: A comprehensive data set of citations with identifiers extracted from english wikipedia
Harshdeep Singh, Robert West, and Giovanni Colavizza · 2020
Later among the works it cites.
A survey on semi-supervised learning
Jesper E Van Engelen and Holger H Hoos · 2020
Later among the works it cites.
Fact or fiction: Verifying scientific claims
David Wadden, Shanchuan Lin, Kyle Lo, Lucy Lu Wang, Madeleine van Zuylen, Arman Cohan, and Hannaneh Hajishirzi · 2020
Later among the works it cites.
Luke: Deep contextualized entity representations with entity-aware self-attention
Ikuya Yamada, Akari Asai, Hiroyuki Shindo, Hideaki Takeda, and Yuji Matsumoto · 2020
Later among the works it cites.
Lossless data compression with transformer, 2021
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Low-shot learning with large-scale diffusion
Matthijs Douze, Arthur Szlam, Bharath Hariharan, and Hervé Jégou · 2018
Cited alongside, same era.
Low-shot learning via covariance-preserving adversarial augmentation networks
Hang Gao, Zheng Shou, Alireza Zareian, Hanwang Zhang, and Shih-Fu Chang · 2018
Cited alongside, same era.
Www’18 open challenge: Financial opinion mining and question answering
Macedo Maia, Siegfried Handschuh, André Freitas, Brian Davis, Ross McDermott, Manel Zarrouk, and Alexandra Balahur · 2018
Cited alongside, same era.
Improving language understanding by generative pre-training, 2018
Alec Radford, Karthik Narasimhan, Tim Salimans, and Ilya Sutskever · 2018
Cited alongside, same era.
Learning to compare: Relation network for few-shot learning
Flood Sung, Yongxin Yang, Li Zhang, Tao Xiang, Philip HS Torr, and Timothy M Hospedales · 2018
Cited alongside, same era.
Retrieval of the best counterargument without prior topic knowledge
Henning Wachsmuth, Shahbaz Syed, and Benno Stein · 2018
Cited alongside, same era.
Fabrice Bellard · 2021
Later among the works it cites.
Lossless neural text compression, 2021
Michael Herrera and Kasey Luo · 2021
Later among the works it cites.
Prefix-tuning: Optimizing continuous prompts for generation
Xiang Lisa Li and Percy Liang · 2021
Later among the works it cites.
True few-shot learning with language models
Ethan Perez, Douwe Kiela, and Kyunghyun Cho · 2021
Later among the works it cites.
Trec-covid: Constructing a pandemic information retrieval test collection
Ellen Voorhees, Tasmeer Alam, Steven Bedrick, Dina Demner-Fushman, William R. Hersh, Kyle Lo, Kirk Roberts, Ian Soboroff, and Lucy Lu Wang · 2021
Later among the works it cites.
An empirical study of pre-trained vision models on out-of-distribution generalization
Yaodong Yu, Heinrich Jiang, Dara Bahri, Hossein Mobahi, Seungyeon Kim, Ankit Singh Rawat, Andreas Veit, and Yi Ma · 2021
Later among the works it cites.
Calibrate before use: Improving few-shot performance of language models
Zihao Zhao, Eric Wallace, Shi Feng, Dan Klein, and Sameer Singh · 2021
Later among the works it cites.
When does pretraining help?: assessing self-supervised learning for law and the casehold dataset of 53,000+ legal holdings
Lucia Zheng, Neel Guha, Brandon R. Anderson, Peter Henderson, and Daniel E. Ho · 2021
Later among the works it cites.
Pada: Example-based prompt learning for on-the-fly adaptation to unseen domains
Eyal Ben-David, Nadav Oved, and Roi Reichart · 2022
Later among the works it cites.
Few-shot non-parametric learning with deep latent variable model
Zhiying Jiang, Yiqin Dai, Ji Xin, Ming Li, and Jimmy Lin · 2022
Later among the works it cites.
Few-shot parameter-efficient fine-tuning is better and cheaper than in-context learning
Haokun Liu, Derek Tam, Mohammed Muqeeth, Jay Mohta, Tenghao Huang, Mohit Bansal, and Colin A Raffel · 2022
Later among the works it cites.
Trace: A fast transformer-based general-purpose lossless compressor
Yu Mao, Yufei Cui, Tei-Wei Kuo, and Chun Jason Xue · 2022
Later among the works it cites.
An unsupervised sentence embedding method by maximizing the mutual information of augmented text representations
Tianye Sheng, Lisong Wang, Zongfeng He, Mingjie Sun, and Guohua Jiang · 2022
Later among the works it cites.
Fast lossless neural compression with integer-only discrete flows
Siyu Wang, Jianfei Chen, Chongxuan Li, Jun Zhu, and Bo Zhang · 2022
Later among the works it cites.
A theory of human-like few-shot learning, 01 2023
Zhiying Jiang, Rui Wang, Dongbo Bu, and Ming Li · 2023
Closest in time.
“low-resource” text classification: A parameter-free classification method with compressors
Zhiying Jiang, Matthew Yang, Mikhail Tsirlin, Raphael Tang, Yiqin Dai, and Jimmy Lin · 2023
Closest in time.
Pre-train, prompt, and predict: A systematic survey of prompting methods in natural language processing
Pengfei Liu, Weizhe Yuan, Jinlan Fu, Zhengbao Jiang, Hiroaki Hayashi, and Graham Neubig · 2023
Closest in time.
Gpt-4 technical report, 2023
OpenAI · 2023
Closest in time.
Evaluating unsupervised text classification: Zero-shot and similarity-based approaches
Tim Schopf, Daniel Braun, and Florian Matthes · 2023
Closest in time.
Self-adaptive in-context learning: An information compression perspective for in-context example selection and ordering, 2023
Zhiyong Wu, Yaoxiang Wang, Jiacheng Ye, and Lingpeng Kong · 2023
Closest in time.