Fetching the paper…
Reading the bibliography…
Evaluating the quality of videos generated from text-to-video (T2V) models is important if they are to produce plausible outputs that convince a viewer of their authenticity.
Lindeberg, T.: Feature detection with automatic scale selection. International Journal of Computer Vision 30
1998
Earlier work this paper cites.
Cover, T.M., Thomas, J.A.: Elements of information theory second edition solutions to problems. Internet Access pp. 19–20 (2006)
2006
Earlier work this paper cites.
Sheikh, H.R., Bovik, A.C.: Image information and visual quality. IEEE Transactions on Image Processing 15
2006
Earlier work this paper cites.
Abdi, H., Williams, L.J.: Tukey’s honestly significant difference (HSD) test. Encyclopedia of research design 3
2010
Earlier work this paper cites.
Hoßfeld, T., Schatz, R., Egger, S.: SOS: The MOS is not enough! In: 2011 Third International Workshop on Quality of Multimedia Experience. pp. 131–136 (2011)
2011
Earlier work this paper cites.
Rublee, E., Rabaud, V., Konolige, K., Bradski, G.: Orb: An efficient alternative to sift or surf. In: 2011 International conference on computer vision. pp. 2564–2571. Ieee (2011)
2011
Earlier work this paper cites.
Mittal, A., Moorthy, A.K., Bovik, A.C.: No-reference image quality assessment in the spatial domain. IEEE Transactions on image processing 21
2012
Earlier work this paper cites.
Mittal, A., Soundararajan, R., Bovik, A.C.: Making a “completely blind” image quality analyzer. IEEE Signal Processing Letters 20
2012
Earlier work this paper cites.
Podpora, M., Korbas, G.P., Kawala-Janik, A.: Yuv vs rgb-choosing a color space for human-machine interaction. In: FedCSIS (Position Papers). pp. 29–34. Citeseer (2014)
2014
Earlier work this paper cites.
Salimans, T., Goodfellow, I., Zaremba, W., Cheung, V., Radford, A., Chen, X.: Improved techniques for training gans. Advances in Neural Information Processing Systems 29
2016
Cited alongside, same era.
Streijl, R.C., Winkler, S., Hands, D.S.: Mean opinion score (MOS) revisited: methods and applications, limitations and alternatives. Multimedia Systems 22
2016
Cited alongside, same era.
Szegedy, C., Vanhoucke, V., Ioffe, S., Shlens, J., Wojna, Z.: Rethinking the inception architecture for computer vision. In: IEEE Conference on Computer Vision and Pattern Recognition. pp. 2818–2826 (2016)
2016
Cited alongside, same era.
Heusel, M., Ramsauer, H., Unterthiner, T., Nessler, B., Hochreiter, S.: GANs trained by a two time-scale update rule converge to a local Nash equilibrium. AAdvances in Neural Information Processing Systems 30
2017
Cited alongside, same era.
Gao, Y., Min, X., Zhu, Y., Li, J., Zhang, X.P., Zhai, G.: Image quality assessment: From mean opinion score to opinion score distribution. In: Proceedings of the 30th ACM International Conference on Multimedia. p. 997–1005 (2022)
2022
Later among the works it cites.
Rombach, R., Blattmann, A., Lorenz, D., Esser, P., Ommer, B.: High-resolution image synthesis with latent diffusion models. In: IEEE/CVF Conference on Computer Vision and Pattern Recognition. pp. 10684–10695 (2022)
2022
Later among the works it cites.
2022
Later among the works it cites.
Epstein, V.: Aphantasia text to video model. https://github.com/eps696/aphantasia , last Acccessed 29 August 2023
2023
Closest in time.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
2018
Cited alongside, same era.
Unterthiner, T., van Steenkiste, S., Kurach, K., Marinier, R., Michalski, M., Gelly, S.: FVD: A new Metric for Video Generation (May 2019), iCLR Workshop on Deep Generative Models for Highly Structured Data
2019
Cited alongside, same era.
Radford, A., Kim, J.W., Hallacy, C., Ramesh, A., Goh, G., Agarwal, S., Sastry, G., Askell, A., Mishkin, P., Clark, J., et al.: Learning transferable visual models from natural language supervision. In: International Conference on Machine Learning. pp. 8748–8763. PMLR (2021)
2021
Cited alongside, same era.
2021
Cited alongside, same era.
2023
Closest in time.
2023
Closest in time.
2023
Closest in time.
Luo, Z., Chen, D., Zhang, Y., Huang, Y., Wang, L., Shen, Y., Zhao, D., Zhou, J., Tan, T.: VideoFusion: Decomposed Diffusion Models for High-Quality Video Generation. In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition. pp. 10209–10218 (2023)
2023
Closest in time.