Fetching the paper…
Reading the bibliography…
Relations are basic building blocks of human cognition.
Topological structure in visual perception
Chen L · 1982
Earlier work this paper cites.
Perception of partly occluded objects in infancy
Kellman PJ, Spelke ES · 1983
Earlier work this paper cites.
Lexicalization patterns: Semantic structure in lexical forms
Talmy L · 1985
Earlier work this paper cites.
Initial knowledge: Six suggestions
Spelke E · 1994
Earlier work this paper cites.
The principle of semantic compositionality
Pelletier FJ · 1994
Earlier work this paper cites.
Conceptual precursors to language
Hespos SJ, Spelke ES · 2004
Earlier work this paper cites.
Social evaluation by preverbal infants
Hamlin JK, Wynn K, Bloom P · 2007
Earlier work this paper cites.
Describing scenes hardly seen
Dobel C, Gumnior H, Bölte J, Zwitserlood P · 2007
Earlier work this paper cites.
The psychophysics of chasing: A case study in the perception of animacy
Gao T, Newman GE, Scholl BJ · 2009
Earlier work this paper cites.
Help or hinder: Bayesian models of social goal inference
Ullman T, Baker C, Macindoe O, Evans O, Goodman N, Tenenbaum J · 2009
Earlier work this paper cites.
Neuronal arithmetic
Silver RA · 2010
Earlier work this paper cites.
Event completion: Event based inferences distort memory in a matter of seconds
Strickland B, Keil F · 2011
Earlier work this paper cites.
Improved semantic representations from tree-structured long short-term memory networks
Tai KS, Socher R, Manning CD · 2015
Earlier work this paper cites.
Perceiving fully occluded objects via physical simulation
Yildirim I, Siegel MH, Tenenbaum JB · 2016
Cited alongside, same era.
The automaticity of perceiving animacy: Goal-directed motion in simple shapes influences visuomotor behavior even when task-irrelevant
van Buren B, Uddenberg S, Scholl BJ · 2016
Cited alongside, same era.
Rapid apprehension of the coherence of action scenes
Glanemann R, Zwitserlood P, Bölte J, Dobel C · 2016
Cited alongside, same era.
Semantic compositionality
Pelletier FJ · 2016
Cited alongside, same era.
Why neurons mix: high dimensionality for higher cognition
Fusi S, Miller EK, Rigotti M · 2016
Cited alongside, same era.
Topological relations between objects are categorically coded
Lovett A, Franconeri SL · 2017
Cited alongside, same era.
The perception of relations
Hafri A, Firestone C · 2021
Later among the works it cites.
Glide: Towards photorealistic image generation and editing with text-guided diffusion models
Nichol A, Dhariwal P, Ramesh A, Shyam P, Mishkin P, McGrew B, et al · 2021
Later among the works it cites.
Learning transferable visual models from natural language supervision
Radford A, Kim JW, Hallacy C, Ramesh A, Goh G, Agarwal S, et al · 2021
Later among the works it cites.
CLIPort: What and Where Pathways for Robotic Manipulation
Shridhar M, Manuelli L, Fox D · 2021
Later among the works it cites.
Photorealistic Text-to-Image Diffusion Models with Deep Language Understanding
Saharia C, Chan W, Saxena S, Li L, Whang J, Denton E, et al · 2022
Closest in time.
Hierarchical text-conditional image generation with clip latents
Ramesh A, Dhariwal P, Nichol A, Chu C, Chen M · 2022
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Seeing physics in the blink of an eye
Firestone C, Scholl B · 2017
Cited alongside, same era.
Clevr: A diagnostic dataset for compositional language and elementary visual reasoning
Johnson J, Hariharan B, Van Der Maaten L, Fei-Fei L, Lawrence Zitnick C, Girshick R · 2017
Cited alongside, same era.
Beyond the Turk: Alternative platforms for crowdsourcing behavioral research
Peer E, Brandimarte L, Samat S, Acquisti A · 2017
Cited alongside, same era.
Seeing what’s possible: Disconnected visual parts are confused for their potential wholes
Guan C, Firestone C · 2020
Cited alongside, same era.
A phone in a basket looks like a knife in a cup: The perception of abstract relations
Hafri A, Bonner MF, Landau B, Firestone C · 2020
Cited alongside, same era.
RELATE: Physically plausible multi-object scene synthesis using structured latent spaces
Ehrhardt S, Groth O, Monszpart A, Engelcke M, Posner I, Mitra N, et al · 2020
Cited alongside, same era.
Closest in time.
A very preliminary analysis of DALL-E 2
Marcus G, Davis E, Aaronson S · 2022
Closest in time.
Available from: https://www.lesswrong.com/posts/uKp6tBFStnsvrot5t/what-dall-e-2-can-and-cannot-do#DALLE_s_weaknesses
Swimmer963. What DALL-E 2 can and cannot do; 2022 · 2022
Closest in time.
Perspective (In) consistency of Paint by Text
Farid H · 2022
Closest in time.
Compositional Visual Generation with Composable Diffusion Models
Liu N, Li S, Du Y, Torralba A, Tenenbaum JB · 2022
Closest in time.
Vqgan-clip: Open domain image generation and editing with natural language guidance
Crowson K, Biderman S, Kornis D, Stander D, Hallahan E, Castricato L, et al · 2022
Closest in time.
High-resolution image synthesis with latent diffusion models
Rombach R, Blattmann A, Lorenz D, Esser P, Ommer B · 2022
Closest in time.
Associative memory of structured knowledge
Steinberg J, Sompolinsky H · 2022
Closest in time.