Fetching the paper…
Reading the bibliography…
The Abstraction and Reasoning Corpus (ARC) is a visual program synthesis benchmark designed to test challenging out-of-distribution generalization in humans and machines.
On the measure of intelligence
Chollet, F. (2019) · 1911
Earlier work this paper cites.
Scaling laws for neural language models
Kaplan, J., McCandlish, S., Henighan, T., Brown, T. B., Chess, B., Child, R., Gray, S., Radford, A., Wu, J., and Amodei, D. (2020) · 2020
Earlier work this paper cites.
Finding categories through words: More nameable features improve category learning
Zettersten, M. and Lupyan, G. (2020) · 2020
Earlier work this paper cites.
Fast and flexible: Human program induction in abstract reasoning tasks
Johnson, A., Vong, W. K., Lake, B., and Gureckis, T. M. (2021) · 2021
Earlier work this paper cites.
Communicating natural programs to humans and machines
Acquaviva, S., Pu, Y., Kryven, M., Sechopoulos, T., Wong, C., Ecanow, G., Nye, M., Tessler, M., and Tenenbaum, J. (2022) · 2022
Earlier work this paper cites.
Time spent thinking in online chess reflects the value of computation
Russek, E., Acosta-Kane, D., van Opheusden, B., Mattar, M. G., and Griffiths, T. (2022) · 2022
Earlier work this paper cites.
Emergent abilities of large language models
Wei, J., Tay, Y., Bommasani, R., Raffel, C., Zoph, B., Borgeaud, S., Yogatama, D., Bosma, M., Zhou, D., Metzler, D., et al. (2022) · 2022
Cited alongside, same era.
Achiam, J., Adler, S., Agarwal, S., Ahmad, L., Akkaya, I., Aleman, F. L., Almeida, D., Altenschmidt, J., Altman, S., Anadkat, S., et al. (2023) · 2023
Cited alongside, same era.
Evaluating cloudresearch’s approved group as a solution for problematic data quality on mturk
Hauser, D. J., Moss, A. J., Rosenzweig, C., Jaffe, S. N., Robinson, J., and Litman, L. (2023) · 2023
Cited alongside, same era.
Comparing humans, gpt-4, and gpt-4v on abstraction and reasoning tasks
Mitchell, M., Palmarini, A. B., and Moskvichev, A. (2023) · 2023
Cited alongside, same era.
The conceptarc benchmark: Evaluating understanding and generalization in the arc domain
Hypothesis search: Inductive reasoning with language models
Wang, R., Zelikman, E., Poesia, G., Pu, Y., Haber, N., and Goodman, N. D. (2023) · 2023
Later among the works it cites.
Claude 3.5 sonnet model card addendum
Anthropic (2024) · 2024
Closest in time.
Response to Difficulty Drives Variation in IQ Test Performance
Cheyette, S. J. and Piantadosi, S. T. (2024) · 2024
Closest in time.
Getting 50% (SoTA) on ARC-AGI with GPT-4o
Greenblatt, R. (2024) · 2024
Closest in time.
Using frontier models on arc agi via langchain
Kamradt, G. (2024) · 2024
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Moskvichev, A., Odouard, V. V., and Mitchell, M. (2023) · 2023
Cited alongside, same era.