Fetching the paper…
Reading the bibliography…
Human similarity judgments are a powerful supervision signal for machine learning applications based on techniques such as contrastive learning, information retrieval, and model alignment, but classical methods for collecting human similarity judgments are too expensive to be used at scale.
Some experimental results in the correlation of mental abilities 1
William Brown · 1910
Earlier work this paper cites.
Features of similarity
Amos Tversky · 1977
Earlier work this paper cites.
Multidimensional scaling, tree-fitting, and clustering
Roger N Shepard · 1980
Earlier work this paper cites.
Toward a universal law of generalization for psychological science
Roger N Shepard · 1987
Earlier work this paper cites.
Generalization, similarity, and bayesian inference
Joshua B Tenenbaum and Thomas L Griffiths · 2001
Earlier work this paper cites.
The big book of concepts
Gregory Murphy · 2004
Earlier work this paper cites.
A package for automatic evaluation of summaries
Lin CY Rouge · 2004
Earlier work this paper cites.
Labeling images with a computer game
Luis Von Ahn and Laura Dabbish · 2004
Earlier work this paper cites.
A bayesian view of language evolution by iterated learning
Thomas L Griffiths and Michael L Kalish · 2005
Earlier work this paper cites.
Speakers optimize information density through syntactic reduction
T Jaeger and Roger Levy · 2006
Earlier work this paper cites.
TagATune: A game for music and sound annotation
Edith LM Law, Luis Von Ahn, Roger B Dannenberg, and Mike Crawford · 2007
Earlier work this paper cites.
Cumulative cultural evolution in the laboratory: An experimental approach to the origins of structure in human language
Simon Kirby, Hannah Cornish, and Kenny Smith · 2008
Earlier work this paper cites.
Designing games with a purpose
Luis Von Ahn and Laura Dabbish · 2008
Earlier work this paper cites.
ImageNet: A large-scale hierarchical image database
Jia Deng, Wei Dong, Richard Socher, Li-Jia Li, Kai Li, and Li Fei-Fei · 2009
Earlier work this paper cites.
Low-dimensional embedding using adaptively selected ordinal data
Kevin G Jamieson and Robert D Nowak · 2011
Earlier work this paper cites.
Scikit-learn: Machine learning in Python
F. Pedregosa, G. Varoquaux, A. Gramfort, V. Michel, B. Thirion, O. Grisel, M. Blondel, P. Prettenhofer, R. Weiss, V. Dubourg, J. Vanderplas, A. Passos, D. Cournapeau, M. Brucher, M. Perrot, and E. Duchesnay · 2011
Earlier work this paper cites.
Word lengths are optimized for efficient communication
Steven T Piantadosi, Harry Tily, and Edward Gibson · 2011
Earlier work this paper cites.
Introducing lextale: A quick and valid lexical test for advanced learners of english
Kristin Lemhöfer and Mirjam Broersma · 2012
Earlier work this paper cites.
Performance-optimized hierarchical models predict neural responses in higher visual cortex
Daniel LK Yamins, Ha Hong, Charles F Cadieu, Ethan A Solomon, Darren Seibert, and James J DiCarlo · 2014
Earlier work this paper cites.
Variations of the similarity function of textrank for automated summarization
Federico Barrios, Federico López, Luis Argerich, and Rosa Wachenchauzer · 2016
Earlier work this paper cites.
Paper recommender systems: a literature survey
Joeran Beel, Bela Gipp, Stefan Langer, and Corinna Breitinger · 2016
Earlier work this paper cites.
Self-report captures 27 distinct categories of emotion bridged by continuous gradients
Alan S Cowen and Dacher Keltner · 2017
Earlier work this paper cites.
The Kinetics human action video dataset
Will Kay, Joao Carreira, Karen Simonyan, Brian Zhang, Chloe Hillier, Sudheendra Vijayanarasimhan, Fabio Viola, Tim Green, Trevor Back, Paul Natsev, et al · 2017
Cited alongside, same era.
Visual genome: Connecting language and vision using crowdsourced dense image annotations
Ranjay Krishna, Yuke Zhu, Oliver Groth, Justin Johnson, Kenji Hata, Joshua Kravitz, Stephanie Chen, Yannis Kalantidis, Li-Jia Li, David A Shamma, et al · 2017
Cited alongside, same era.
What are the visual features underlying human versus machine vision?
Drew Linsley, Sven Eberhardt, Tarun Sharma, Pankaj Gupta, and Thomas Serre · 2017
Cited alongside, same era.
Conceptnet 5.5: An open multilingual graph of general knowledge
Robyn Speer, Joshua Chin, and Catherine Havasi · 2017
Cited alongside, same era.
Headphone screening to facilitate web-based auditory experiments
Kevin JP Woods, Max H Siegel, James Traer, and Josh H McDermott · 2017
Cited alongside, same era.
Crisscrossed captions: Extended intramodal and intermodal semantic similarity judgments for MS-COCO
Zarana Parekh, Jason Baldridge, Daniel Cer, Austin Waters, and Yinfei Yang · 2020
Later among the works it cites.
Brain-Score: Which artificial neural network for object recognition is most brain-like?
Martin Schrimpf, Jonas Kubilius, Ha Hong, Najib J Majaj, Rishi Rajalingham, Elias B Issa, Kohitij Kar, Pouya Bashivan, Jonathan Prescott-Roy, Franziska Geiger, et al · 2020
Later among the works it cites.
Cultural influences on word meanings revealed through large-scale semantic alignment
Bill Thompson, Seán G Roberts, and Gary Lupyan · 2020
Later among the works it cites.
Transformers: State-of-the-art natural language processing
Thomas Wolf, Lysandre Debut, Victor Sanh, Julien Chaumond, Clement Delangue, Anthony Moi, Pierric Cistac, Tim Rault, Rémi Louf, Morgan Funtowicz, Joe Davison, Sam Shleifer, Patrick von Platen, Clara Ma, Yacine Jernite, Julien Plu, Canwen Xu, Teven Le Scao, Sylvain Gugger, Mariama Drame, Quentin Lhoest, and Alexander M. Rush · 2020
Later among the works it cites.
An optimization-based approach to understanding sensory systems
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Generative timbre spaces: regularizing variational auto-encoders with perceptual metrics
Philippe Esling, Adrien Bitton, et al · 2018
Cited alongside, same era.
A task-optimized neural network replicates human auditory behavior, predicts brain responses, and reveals a cortical processing hierarchy
Alexander JE Kell, Daniel LK Yamins, Erica N Shook, Sam V Norman-Haignere, and Josh H McDermott · 2018
Cited alongside, same era.
The Ryerson audio-visual database of emotional speech and song (RAVDESS): A dynamic, multimodal set of facial and vocal expressions in north american english
Steven R Livingstone and Frank A Russo · 2018
Cited alongside, same era.
Evaluating (and improving) the correspondence between deep neural networks and human representations
Joshua C Peterson, Joshua T Abbott, and Thomas L Griffiths · 2018
Cited alongside, same era.
Rethinking spatiotemporal feature learning: Speed-accuracy trade-offs in video classification
Saining Xie, Chen Sun, Jonathan Huang, Zhuowen Tu, and Kevin Murphy · 2018
Cited alongside, same era.
Efficient compression in color naming and its evolution
Noga Zaslavsky, Charles Kemp, Terry Regier, and Naftali Tishby · 2018
Cited alongside, same era.
SlowFast networks for video recognition
Christoph Feichtenhofer, Haoqi Fan, Jitendra Malik, and Kaiming He · 2019
Cited alongside, same era.
Daniel Yamins · 2020
Later among the works it cites.
Bertscore: Evaluating text generation with bert
Tianyi Zhang, Varsha Kishore, Felix Wu, Kilian Q. Weinberger, and Yoav Artzi · 2020
Later among the works it cites.
BEiT: BERT pre-training of image transformers
Hangbo Bao, Li Dong, and Furu Wei · 2021
Later among the works it cites.
WavLM: Large-scale self-supervised pre-training for full stack speech processing
Sanyuan Chen, Chengyi Wang, Zhengyang Chen, Yu Wu, Shujie Liu, Zhuo Chen, Jinyu Li, Naoyuki Kanda, Takuya Yoshioka, Xiong Xiao, et al · 2021
Later among the works it cites.
PyTorchVideo: A deep learning library for video understanding
Haoqi Fan, Tullie Murrell, Heng Wang, Kalyan Vasudev Alwala, Yanghao Li, Yilei Li, Bo Xiong, Nikhila Ravi, Meng Li, Haichuan Yang, Jitendra Malik, Ross Girshick, Matt Feiszli, Aaron Adcock, Wan-Yen Lo, and Christoph Feichtenhofer · 2021
Later among the works it cites.
SimCSE: Simple contrastive learning of sentence embeddings
Tianyu Gao, Xingcheng Yao, and Danqi Chen · 2021
Later among the works it cites.
HuBERT: Self-supervised speech representation learning by masked prediction of hidden units
Wei-Ning Hsu, Benjamin Bolte, Yao-Hung Hubert Tsai, Kushal Lakhotia, Ruslan Salakhutdinov, and Abdelrahman Mohamed · 2021
Later among the works it cites.
Passive attention in artificial neural networks predicts human visual selectivity
Thomas Langlois, Haicheng Zhao, Erin Grant, Ishita Dasgupta, Tom Griffiths, and Nori Jacoby · 2021
Later among the works it cites.
Swin transformer: Hierarchical vision transformer using shifted windows
Ze Liu, Yutong Lin, Yue Cao, Han Hu, Yixuan Wei, Zheng Zhang, Stephen Lin, and Baining Guo · 2021
Later among the works it cites.
An online headphone screening test based on dichotic pitch
Alice E Milne, Roberta Bianco, Katarina C Poole, Sijia Zhao, Andrew J Oxenham, Alexander J Billig, and Maria Chait · 2021
Later among the works it cites.
SpeechBrain: A general-purpose speech toolkit, 2021
Mirco Ravanelli, Titouan Parcollet, Peter Plantinga, Aku Rouhe, Samuele Cornell, Loren Lugosch, Cem Subakan, Nauman Dawalatabad, Abdelwahab Heba, Jianyuan Zhong, Ju-Chieh Chou, Sung-Lin Yeh, Szu-Wei Fu, Chien-Feng Liao, Elena Rastorgueva, François Grondin, William Aris, Hwidong Na, Yan Gao, Renato De Mori, and Yoshua Bengio · 2021
Later among the works it cites.
Enriching ImageNet with human similarity judgments and psychological embeddings
Brett D Roads and Bradley C Love · 2021
Later among the works it cites.
SUPERB: Speech Processing Universal PERformance Benchmark
Shu wen Yang, Po-Han Chi, Yung-Sung Chuang, Cheng-I Jeff Lai, Kushal Lakhotia, Yist Y. Lin, Andy T. Liu, Jiatong Shi, Xuankai Chang, Guan-Ting Lin, Tzu-Hsien Huang, Wei-Cheng Tseng, Ko tik Lee, Da-Rong Liu, Zili Huang, Shuyan Dong, Shang-Wen Li, Shinji Watanabe, Abdelrahman Mohamed, and Hung yi Lee · 2021
Later among the works it cites.
Torchaudio: Building blocks for audio and speech processing
Yao-Yuan Yang, Moto Hira, Zhaoheng Ni, Anjali Chourdia, Artyom Astafurov, Caroline Chen, Ching-Feng Yeh, Christian Puhrsch, David Pollack, Dmitriy Genzel, Donny Greenberg, Edward Z. Yang, Jason Lian, Jay Mahadeokar, Jeff Hwang, Ji Chen, Peter Goldsborough, Prabhat Roy, Sean Narenthiran, Shinji Watanabe, Soumith Chintala, Vincent Quenneville-Bélair, and Yangyang Shi · 2021
Later among the works it cites.
data2vec: A general framework for self-supervised learning in speech, vision and language, 2022
Alexei Baevski, Wei-Ning Hsu, Qiantong Xu, Arun Babu, Jiatao Gu, and Michael Auli · 2022
Closest in time.
Zhuang Liu, Hanzi Mao, Chao-Yuan Wu, Christoph Feichtenhofer, Trevor Darrell, and Saining Xie · 2022
Closest in time.
Predicting human similarity judgments using large language models
Raja Marjieh, Ilia Sucholutsky, Theodore R Sumers, Nori Jacoby, and Thomas L Griffiths · 2022
Closest in time.
Dawn of the transformer era in speech emotion recognition: closing the valence gap, 2022
Johannes Wagner, Andreas Triantafyllopoulos, Hagen Wierstorf, Maximilian Schmitt, Felix Burkhardt, Florian Eyben, and Björn W. Schuller · 2022
Closest in time.