Fetching the paper…
Reading the bibliography…
Large-scale pre-trained video generation models excel in content creation but are not reliable as physically accurate world simulators out of the box.
The kolmogorov-smirnov test for goodness of fit
Massey Jr, F. J · 1951
Earlier work this paper cites.
The nature of explanation , volume 445
Craik, K. J. W · 1967
Earlier work this paper cites.
Origins of knowledge
Spelke, E. S., Breinlinger, K., Macomber, J., and Jacobson, K · 1992
Earlier work this paper cites.
Infants’ physical world
Baillargeon, R · 2004
Earlier work this paper cites.
Bullet physics engine
Coumans, E. et al · 2010
Earlier work this paper cites.
Simulation as an engine of physical scene understanding
Battaglia, P. W., Hamrick, J. B., and Tenenbaum, J. B · 2013
Earlier work this paper cites.
Unsupervised learning of video representations using lstms
Srivastava, N., Mansimov, E., and Salakhudinov, R · 2015
Earlier work this paper cites.
Galileo: Perceiving physical object properties by integrating a physics engine with deep learning
Wu, J., Yildirim, I., Lim, J. J., Freeman, B., and Tenenbaum, J · 2015
Earlier work this paper cites.
Inferring mass in complex scenes by mental simulation
Hamrick, J. B., Battaglia, P. W., Griffiths, T. L., and Tenenbaum, J. B · 2016
Earlier work this paper cites.
Improved techniques for training gans
Salimans, T., Goodfellow, I., Zaremba, W., Cheung, V., Radford, A., and Chen, X · 2016
Earlier work this paper cites.
Visual dynamics: Probabilistic future frame synthesis via cross convolutional networks
Xue, T., Wu, J., Bouman, K., and Freeman, B · 2016
Earlier work this paper cites.
Gans trained by a two time-scale update rule converge to a local nash equilibrium
Heusel, M., Ramsauer, H., Unterthiner, T., Nessler, B., and Hochreiter, S · 2017
Earlier work this paper cites.
Building machines that learn and think like people
Lake, B. M., Ullman, T. D., Tenenbaum, J. B., and Gershman, S. J · 2017
Earlier work this paper cites.
Mind games: Game engines as an architecture for intuitive physics
Ullman, T. D., Spelke, E., Battaglia, P., and Tenenbaum, J. B · 2017
Earlier work this paper cites.
Blender - a 3d modelling and rendering package, 2018
Community, B. O · 2018
Earlier work this paper cites.
Recurrent world models facilitate policy evolution
Ha, D. and Schmidhuber, J · 2018
Cited alongside, same era.
Towards accurate generative models of video: A new metric & challenges
Unterthiner, T., Van Steenkiste, S., Kurach, K., Marinier, R., Michalski, M., and Gelly, S · 2018
Cited alongside, same era.
Dream to control: Learning behaviors by latent imagination
Hafner, D., Lillicrap, T., Ba, J., and Norouzi, M · 2019
Cited alongside, same era.
Raft: Recurrent all-pairs field transforms for optical flow
Teed, Z. and Deng, J · 2020
Cited alongside, same era.
Physion: Evaluating physical prediction from vision in humans and machines
Bear, D. M., Wang, E., Mrowca, D., Binder, F. J., Tung, H.-Y. F., Pramod, R., Holdaway, C., Tao, S., Smith, K., Sun, F.-Y., et al · 2021
Cited alongside, same era.
Pyramidal flow matching for efficient video generative modeling
Jin, Y., Sun, Z., Li, N., Xu, K., Jiang, H., Zhuang, N., Huang, Q., Song, Y., Mu, Y., and Lin, Z · 2024
Later among the works it cites.
How far is video generation from world model: A physical law perspective
Kang, B., Yue, Y., Lu, R., Lin, Z., Zhao, Y., Wang, K., Huang, G., and Feng, J · 2024
Later among the works it cites.
Kling, 2024
Kuaishou · 2024
Later among the works it cites.
Dream machine, 2024
Luma · 2024
Later among the works it cites.
Towards world simulator: Crafting physical commonsense-based benchmark for video generation
Meng, F., Liao, J., Tan, X., Shao, W., Lu, Q., Zhang, K., Cheng, Y., Li, D., Qiao, Y., and Luo, P · 2024
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Google scanned objects: A high-quality dataset of 3d scanned household items
Downs, L., Francis, A., Koenig, N., Kinman, B., Hickman, R., Reymann, K., McHugh, T. B., and Vanhoucke, V · 2022
Cited alongside, same era.
Kubric: A scalable dataset generator
Greff, K., Belletti, F., Beyer, L., Doersch, C., Du, Y., Duckworth, D., Fleet, D. J., Gnanapragasam, D., Golemo, F., Herrmann, C., et al · 2022
Cited alongside, same era.
A path towards autonomous machine intelligence version 0.9. 2, 2022-06-27
LeCun, Y · 2022
Cited alongside, same era.
Flow straight and fast: Learning to generate and transfer data with rectified flow
Liu, X., Gong, C., and Liu, Q · 2022
Cited alongside, same era.
Mastering diverse domains through world models
Hafner, D., Pasukonis, J., Ba, J., and Lillicrap, T · 2023
Cited alongside, same era.
Dynamicrafter: Animating open-domain images with video diffusion priors
Xing, J., Xia, M., Zhang, Y., Chen, H., Yu, W., Liu, H., Wang, X., Wong, T.-T., and Shan, Y · 2023
Cited alongside, same era.
Learning interactive real-world simulators
Yang, M., Du, Y., Ghasemipour, K., Tompson, J., Schuurmans, D., and Abbeel, P · 2023
Cited alongside, same era.
Sora, 2024
OpenAI · 2024
Later among the works it cites.
Video diffusion alignment via reward gradients
Prabhudesai, M., Mendonca, R., Qin, Z., Fragkiadaki, K., and Pathak, D · 2024
Later among the works it cites.
Sam 2: Segment anything in images and videos
Ravi, N., Gabeur, V., Hu, Y.-T., Hu, R., Ryali, C., Ma, T., Khedr, H., Rädle, R., Rolland, C., Gustafson, L., Mintun, E., Pan, J., Alwala, K. V., Carion, N., Wu, C.-Y., Girshick, R., Dollár, P., and Feichtenhofer, C · 2024
Later among the works it cites.
Gen-3 alpha, 2024
Runway · 2024
Later among the works it cites.
Open-sora: Democratizing efficient video production for all, March 2024
Zheng, Z., Peng, X., Yang, T., Shen, C., Li, S., Liu, H., Zhou, Y., Li, T., and You, Y · 2024
Later among the works it cites.
Cosmos world foundation model platform for physical AI
Agarwal, N., Ali, A., Bala, M., Balaji, Y., Barker, E., Cai, T., Chattopadhyay, P., Chen, Y., Cui, Y., Ding, Y., et al · 2025
Closest in time.
Do generative video models learn physical principles from watching videos?
Motamed, S., Culp, L., Swersky, K., Jaini, P., and Geirhos, R · 2025
Closest in time.
Coca-cola causes controversy with ai-made ad, 2025
NBC · 2025
Closest in time.
AIFF 2025: AI Film Festival, 2025
Runway · 2025
Closest in time.