Fetching the paper…
Reading the bibliography…
The quest for human imitative AI has been an enduring topic in AI research since its inception.
Rigidity of behavior: A variational approach to the effect of Einstellung
A. S. Luchins and E. H. Luchins · 1959
Earlier work this paper cites.
The associative basis of the creative process
S. Mednick · 1962
Earlier work this paper cites.
The remote associates test
S. A. Mednick · 1968
Earlier work this paper cites.
Encoding and retrieval from long-term storage
G. Wood and J. Pennington · 1973
Earlier work this paper cites.
Algorithm as 136: A k-means clustering algorithm
J. A. Hartigan and M. A. Wong · 1979
Earlier work this paper cites.
A method for comparing two hierarchical clusterings
E. B. Fowlkes and C. L. Mallows · 1983
Earlier work this paper cites.
Comparing partitions
L. Hubert and P. Arabie · 1985
Earlier work this paper cites.
Incubation effects
S. M. Smith and S. E. Blankenship · 1989
Earlier work this paper cites.
Incubation and the persistence of fixation in problem solving
S. M. Smith and S. E. Blankenship · 1991
Earlier work this paper cites.
WordNet: A lexical database for English
G. A. Miller · 1992
Earlier work this paper cites.
Memory blocks in word fragment completion caused by involuntary retrieval of orthographically related primes
S. M. Smith and D. R. Tindell · 1997
Earlier work this paper cites.
Expertise as mental set: The effects of domain knowledge in creative problem solving
J. Wiley · 1998
Earlier work this paper cites.
Constrained k-means clustering
P. S. Bradley, K. P. Bennett, and A. Demiriz · 2000
Earlier work this paper cites.
Constrained k-means clustering with background knowledge
K. Wagstaff, C. Cardie, S. Rogers, and S. Schrödl · 2001
Earlier work this paper cites.
The constraining effects of initial ideas
S. M. Smith · 2003
Earlier work this paper cites.
ROUGE: A package for automatic evaluation of summaries
C.-Y. Lin · 2004
Earlier work this paper cites.
Constrained clustering: Advances in algorithms, theory, and applications
S. Basu, I. Davidson, and K. Wagstaff · 2008
Earlier work this paper cites.
Visualizing data using t-SNE
L. Van der Maaten and G. Hinton · 2008
Earlier work this paper cites.
Partly versus completely out of your mind: Effects of incubation and distraction on resolving fixation
N. Kohn and S. M. Smith · 2009
Earlier work this paper cites.
Word sense disambiguation: A survey
R. Navigli · 2009
Earlier work this paper cites.
Information theoretic measures for clusterings comparison , volume 09
N. X. Vinh, J. Epps, and B. J · 2009
Earlier work this paper cites.
The winograd schema challenge
H. Levesque, E. Davis, and L. Morgenstern · 2012
Earlier work this paper cites.
Revisiting Mednick’s model on creativity-related differences in associative hierarchies. Evidence for a common path to uncommon thought
M. Benedek and A. C. Neubauer · 2013
Earlier work this paper cites.
On handling negative transfer and imbalanced distributions in multiple source transfer learning
J. Gao, L. Ge, K. Li, H. Q. Ngo, and A. Zhang · 2013
Earlier work this paper cites.
GloVe: Global vectors for word representation
J. Pennington, R. Socher, and C. Manning · 2014
Earlier work this paper cites.
A tutorial on principal component analysis
J. Shlens · 2014
Earlier work this paper cites.
On wasserstein two-sample testing and related families of nonparametric tests
A. Ramdas, N. García Trillos, and M. Cuturi · 2017
Earlier work this paper cites.
Attention is all you need
A. Vaswani, N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, L. Kaiser, and I. Polosukhin · 2017
Earlier work this paper cites.
Chasing red herrings: Memory of distractors causes fixation in creative problem solving
Z. Beda and S. M. Smith · 2018
Earlier work this paper cites.
Learning word vectors for 157 languages
E. Grave, P. Bojanowski, P. Gupta, A. Joulin, and T. Mikolov · 2018
Cited alongside, same era.
BPEmb: Tokenization-free pre-trained subword embeddings in 275 languages
B. Heinzerling and M. Strube · 2018
Cited alongside, same era.
Deep contextualized word representations
M. E. Peters, M. Neumann, M. Iyyer, M. Gardner, C. Clark, K. Lee, and L. Zettlemoyer · 2018
Cited alongside, same era.
FLAIR: An easy-to-use framework for state-of-the-art NLP
A. Akbik, T. Bergmann, D. Blythe, K. Rasul, S. Schweter, and R. Vollgraf · 2019
Cited alongside, same era.
BERT: Pre-training of deep bidirectional transformers for language understanding
J. Devlin, M.-W. Chang, K. Lee, and K. Toutanova · 2019
Cited alongside, same era.
How contextual are contextualized word representations? Comparing the geometry of BERT, ELMo, and GPT-2 embeddings
K. Ethayarajh · 2019
Improving language models by retrieving from trillions of tokens
S. Borgeaud, A. Mensch, J. Hoffmann, T. Cai, E. Rutherford, K. Millican, G. van den Driessche, J. Lespiau, B. Damoc, A. Clark, D. de Las Casas, A. Guy, J. Menick, R. Ring, T. Hennigan, S. Huang, L. Maggiore, C. Jones, A. Cassirer, A. Brock, M. Paganini, G. Irving, O. Vinyals, S. Osindero, K. Simonyan, J. W. Rae, E. Elsen, and L. Sifre · 2022
Later among the works it cites.
PaLM: Scaling language modeling with pathways. 2022
A. Chowdhery, S. Narang, J. Devlin, M. Bosma, G. Mishra, A. Roberts, P. Barham, H. W. Chung, C. Sutton, S. Gehrmann, et al · 2022
Later among the works it cites.
Glam: Efficient scaling of language models with mixture-of-experts
N. Du, Y. Huang, A. M. Dai, S. Tong, D. Lepikhin, Y. Xu, M. Krikun, Y. Zhou, A. W. Yu, O. Firat, B. Zoph, L. Fedus, M. P. Bosma, Z. Zhou, T. Wang, Y. E. Wang, K. Webster, M. Pellat, K. Robinson, K. S. Meier-Hellstern, T. Duke, L. Dixon, K. Zhang, Q. V. Le, Y. Wu, Z. Chen, and C. Cui · 2022
Later among the works it cites.
Few-shot learning with retrieval augmented language models
G. Izacard, P. Lewis, M. Lomeli, L. Hosseini, F. Petroni, T. Schick, J. A. Yu, A. Joulin, S. Riedel, and E. Grave · 2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Cited alongside, same era.
Artificial intelligence—the revolution hasn’t happened yet
M. I. Jordan · 2019
Cited alongside, same era.
Roberta: A robustly optimized bert pretraining approach
Y. Liu, M. Ott, N. Goyal, J. Du, M. Joshi, D. Chen, O. Levy, M. Lewis, L. Zettlemoyer, and V. Stoyanov · 2019
Cited alongside, same era.
Remote associates test: An empirical proof of concept
M. Marko, D. Michalko, and I. Riečanskỳ · 2019
Cited alongside, same era.
DistilBERT, a distilled version of BERT: smaller, faster, cheaper and lighter
V. Sanh, L. Debut, J. Chaumond, and T. Wolf · 2019
Cited alongside, same era.
Megatron-lm: Training multi-billion parameter language models using model parallelism
M. Shoeybi, M. Patwary, R. Puri, P. LeGresley, J. Casper, and B. Catanzaro · 2019
Cited alongside, same era.
Superglue: A stickier benchmark for general-purpose language understanding systems
A. Wang, Y. Pruksachatkun, N. Nangia, A. Singh, J. Michael, F. Hill, O. Levy, and S. R. Bowman · 2019
Cited alongside, same era.
Holistic evaluation of language models
P. Liang, R. Bommasani, T. Lee, D. Tsipras, D. Soylu, M. Yasunaga, Y. Zhang, D. Narayanan, Y. Wu, A. Kumar, et al · 2022
Later among the works it cites.
Training language models to follow instructions with human feedback
L. Ouyang, J. Wu, X. Jiang, D. Almeida, C. Wainwright, P. Mishkin, C. Zhang, S. Agarwal, K. Slama, A. Ray, et al · 2022
Later among the works it cites.
Bloom: A 176b-parameter open-access multilingual language model
T. L. Scao, A. Fan, C. Akiki, E. Pavlick, S. Ilić, D. Hesslow, R. Castagné, A. S. Luccioni, F. Yvon, M. Gallé, et al · 2022
Later among the works it cites.
Beyond the imitation game: Quantifying and extrapolating the capabilities of language models
A. Srivastava, A. Rastogi, A. Rao, A. A. M. Shoeb, A. Abid, A. Fisch, A. R. Brown, A. Santoro, A. Gupta, A. Garriga-Alonso, et al · 2022
Later among the works it cites.
Text embeddings by weakly-supervised contrastive pre-training
L. Wang, N. Yang, X. Huang, B. Jiao, L. Yang, D. Jiang, R. Majumder, and F. Wei · 2022
Later among the works it cites.
Finetuned language models are zero-shot learners
J. Wei, M. Bosma, V. Y. Zhao, K. Guu, A. W. Yu, B. Lester, N. Du, A. M. Dai, and Q. V. Le · 2022
Later among the works it cites.
American== white in multimodal language-and-image ai
R. Wolfe and A. Caliskan · 2022
Later among the works it cites.
OPT: Open pre-trained transformer language models
S. Zhang, S. Roller, N. Goyal, M. Artetxe, M. Chen, S. Chen, C. Dewan, M. Diab, X. Li, X. V. Lin, et al · 2022
Later among the works it cites.
https://github.com/microsoft/guidance
Microsoft guidance: A guidance language for controlling large language models · 2023
Closest in time.
https://platform.openai.com/docs/api-reference/completions
OpenAI API completions reference · 2023
Closest in time.
Planning for AGI and beyond
S. Altman · 2023
Closest in time.
Sparks of artificial general intelligence: Early experiments with GPT-4
S. Bubeck, V. Chandrasekaran, R. Eldan, J. Gehrke, E. Horvitz, E. Kamar, P. Lee, Y. T. Lee, Y. Li, S. Lundberg, et al · 2023
Closest in time.
A. Jacovi, A. Caciularu, O. Goldman, and Y. Goldberg · 2023
Closest in time.
Challenges and applications of large language models
J. Kaddour, J. Harris, M. Mozes, H. Bradley, R. Raileanu, and R. McHardy · 2023
Closest in time.
Large language models can be guided to evade AI-generated text detection
N. Lu, S. Liu, R. He, and K. Tang · 2023
Closest in time.
OpenAI · 2023
Closest in time.
What in-context learning "learns" in-context: Disentangling task recognition and task learning
J. Pan, T. Gao, H. Chen, and D. Chen · 2023
Closest in time.
O. Sainz, O. L. de Lacalle, E. Agirre, and G. Rigau · 2023
Closest in time.
GlobalBench: A benchmark for global progress in natural language processing
Y. Song, C. Cui, S. Khanuja, P. Liu, F. Faisal, A. Ostapenko, G. I. Winata, A. F. Aji, S. Cahyawijaya, Y. Tsvetkov, et al · 2023
Closest in time.
LLaMA: Open and efficient foundation language models
H. Touvron, T. Lavril, G. Izacard, X. Martinet, M.-A. Lachaux, T. Lacroix, B. Rozière, N. Goyal, E. Hambro, F. Azhar, et al · 2023
Closest in time.
Finding word sense embeddings of known meaning
L. White, R. Togneri, W. Liu, and M. Bennamoun · 2023
Closest in time.
Only connect — Wikipedia, the free encyclopedia
Wikipedia contributors · 2023
Closest in time.
Tree of thoughts: Deliberate problem solving with large language models
S. Yao, D. Yu, J. Zhao, I. Shafran, T. L. Griffiths, Y. Cao, and K. Narasimhan · 2023
Closest in time.
Sentence-BERT and k-means based clustering technology for scientific and technical literature
B. Yin, M. Zhao, L. Guo, and L. Qiao · 2023
Closest in time.
A survey of large language models
W. X. Zhao, K. Zhou, J. Li, T. Tang, X. Wang, Y. Hou, Y. Min, B. Zhang, J. Zhang, Z. Dong, et al · 2023
Closest in time.