Fetching the paper…
Reading the bibliography…
Video streaming usage has seen a significant rise as entertainment, education, and business increasingly rely on online video.
Self-improving reactive agents based on reinforcement learning, planning and teaching
L.-J. Lin · 1992
Earlier work this paper cites.
A new rate control scheme using quadratic rate distortion model
Tihao Chiang and Ya-Qin Zhang · 1996
Earlier work this paper cites.
Constrained Markov decision processes
E. Altman · 1999
Earlier work this paper cites.
Trellis-based r-d optimal quantization in h.263+
Jiangtao Wen, M. Luttrell, and J. Villasenor · 2000
Earlier work this paper cites.
Calculation of average PSNR differences between RD-curves
G. Bjontegaard · 2001
Earlier work this paper cites.
Optimum bit allocation and accurate rate control for video coding via p-domain source modeling
Z. He and S. K. Mitra · 2002
Earlier work this paper cites.
Image quality assessment: from error visibility to structural similarity
Z. Wang, A. Bovik, H. Sheikh, and E. Simoncelli · 2003
Earlier work this paper cites.
Rate-distortion analysis for h.264/avc video coding and its application to rate control
S. Ma, Wen Gao, and Yan Lu · 2005
Earlier work this paper cites.
On enhancing h.264/avc video rate control by psnr-based frame complexity estimation
Minqiang Jiang and Nam Ling · 2005
Earlier work this paper cites.
Efficient selectivity and backup operators in monte-carlo tree search
R. Coulom · 2006
Earlier work this paper cites.
Rate control for h.264 video with enhanced rate and distortion models
D. Kwon, M. Shen, and C. . J. Kuo · 2007
Earlier work this paper cites.
Rbf-based qp estimation model for vbr control in h.264/svc
S. Sanz-Rodriguez and F. Diaz-de-Maria · 2011
Earlier work this paper cites.
Playing atari with deep reinforcement learning
V. Mnih, K. Kavukcuoglu, D. Silver, A. Graves, I. Antonoglou, D. Wierstra, and M. Riedmiller · 2013
Earlier work this paper cites.
The latest open-source video codec VP9 - an overview and preliminary results
D. Mukherjee, J. Bankoski, A. Grange, J. Han, J. Koleszar, P. Wilkins, Y. Xu, and R. Bultje · 2013
Earlier work this paper cites.
λ \lambda domain rate control algorithm for high efficiency video coding
B. Li, H. Li, L. Li, and J. Zhang · 2014
Earlier work this paper cites.
J. L. Ba, J. R. Kiros, and G. E. Hinton · 2016
Earlier work this paper cites.
Ssim-based game theory approach for rate-distortion optimized intra frame ctu-level bit allocation
W. Gao, S. Kwong, Y. Zhou, and H. Yuan · 2016
Earlier work this paper cites.
Identity mappings in deep residual networks
K. He, X. Zhang, S. Ren, and J. Sun · 2016
Earlier work this paper cites.
End-to-end training of deep visuomotor policies
S. Levine, C. Finn, T. Darrell, and P. Abbeel · 2016
Cited alongside, same era.
Toward a practical perceptual video quality metric, 2016
Z. Li, A. Aaron, I. Katsavounidis, A. Moorthy, and M. Manohara · 2016
Cited alongside, same era.
Mastering the game of Go with deep neural networks and tree search
D. Silver, A. Huang, C. J. Maddison, A. Guez, L. Sifre, G. Van Den Driessche, J. Schrittwieser, I. Antonoglou, V. Panneershelvam, M. Lanctot, et al · 2016
Cited alongside, same era.
Constrained policy optimization
J. Achiam, D. Held, A. Tamar, and P. Abbeel · 2017
Cited alongside, same era.
Evolution strategies as a scalable alternative to reinforcement learning
T. Salimans, J. Ho, X. Chen, S. Sidor, and I. Sutskever · 2017
Cited alongside, same era.
Constrained reinforcement learning has zero duality gap
S. Paternain, L. Chamon, M. Calvo-Fullana, and A. Ribeiro · 2019
Later among the works it cites.
Self-play learning without a reward metric
D. Schmidt, N. Moran, J. S. Rosenfeld, J. Rosenthal, and J. Yedidia · 2019
Later among the works it cites.
Grandmaster level in starcraft ii using multi-agent reinforcement learning
O. Vinyals, I. Babuschkin, W. M. Czarnecki, M. Mathieu, A. Dudzik, J. Chung, D. H. Choi, R. Powell, T. Ewalds, P. Georgiev, et al · 2019
Later among the works it cites.
Youtube ugc dataset for video compression research
Y. Wang, S. Inguva, and B. Adsumilli · 2019
Later among the works it cites.
URL https://chromium.googlesource.com/webm/libvpx/+/master/vp9/simple_encode.h
WebM, 2019 · 2019
Later among the works it cites.
Cisco annual internet report (2018–2023) white paper, 2020
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
A deep hierarchical approach to lifelong learning in minecraft
C. Tessler, S. Givony, T. Zahavy, D. Mankowitz, and S. Mannor · 2017
Cited alongside, same era.
URL https://developers.google.com/media/vp9/bitrate-modes/
WebM, 2017 · 2017
Cited alongside, same era.
JAX: composable transformations of Python+NumPy programs, 2018
J. Bradbury, R. Frostig, P. Hawkins, M. J. Johnson, C. Leary, D. Maclaurin, G. Necula, A. Paszke, J. VanderPlas, S. Wanderman-Milne, and Q. Zhang · 2018
Cited alongside, same era.
Reinforcement learning for hevc/h.265 frame-level bit allocation
L.-C. Chen, J.-H. Hu, and W.-H. Peng · 2018
Cited alongside, same era.
A lyapunov-based approach to safe reinforcement learning, 2018
Y. Chow, O. Nachum, E. Duenez-Guzman, and M. Ghavamzadeh · 2018
Cited alongside, same era.
Distributed prioritized experience replay
D. Horgan, J. Quan, D. Budden, G. Barth-Maron, M. Hessel, H. Van Hasselt, and D. Silver · 2018
Cited alongside, same era.
Reinforcement learning for hevc/h.265 intra-frame rate control
J. Hu, W. Peng, and C. Chung · 2018
Cited alongside, same era.
Cisco · 2020
Later among the works it cites.
Exploration-exploitation in constrained mdps, 2020
Y. Efroni, S. Mannor, and M. Pirotta · 2020
Later among the works it cites.
Haiku: Sonnet for JAX, 2020
T. Hennigan, T. Cai, T. Norman, and I. Babuschkin · 2020
Later among the works it cites.
Optax: Composable gradient transformation and optimisation, in JAX!, 2020
M. Hessel, D. Budden, F. Viola, M. Rosca, E. Sezener, and T. Hennigan · 2020
Later among the works it cites.
Rate control method based on deep reinforcement learning for dynamic video sequences in hevc
S. Kwong, M. Zhou, W. Xuekai, W. Jia, and B. Fang · 2020
Later among the works it cites.
Neural rate control for video encoding using imitation learning, 2020
H. Mao, C. Gu, M. Wang, A. Chen, N. Lazic, N. Levine, D. Pang, R. Claus, M. Hechtman, C.-H. Chiang, C. Chen, and J. Han · 2020
Later among the works it cites.
Chip placement with deep reinforcement learning
A. Mirhoseini, A. Goldie, M. Yazgan, J. Jiang, E. Songhori, S. Wang, Y.-J. Lee, E. Johnson, O. Pathak, S. Bae, et al · 2020
Later among the works it cites.
Mastering Atari, Go, Chess and Shogi by Planning with a Learned Model
J. Schrittwieser, I. Antonoglou, T. Hubert, K. Simonyan, L. Sifre, S. Schmitt, A. Guez, E. Lockhart, D. Hassabis, T. Graepel, T. P. Lillicrap, and D. Silver · 2020
Later among the works it cites.
Reward constrained interactive recommendation with natural language feedback, 2020
R. Zhang, T. Yu, Y. Shen, H. Jin, C. Chen, and L. Carin · 2020
Later among the works it cites.
Rate control method based on deep reinforcement learning for dynamic video sequences in hevc
M. Zhou, X. Wei, S. Kwong, W. Jia, and B. Fang · 2020
Later among the works it cites.
Balancing constraints and rewards with meta-gradient D4PG
D. A. Calian, D. J. Mankowitz, T. Zahavy, Z. Xu, J. Oh, N. Levine, and T. Mann · 2021
Later among the works it cites.
A dual-critic reinforcement learning framework for frame-level bit allocation in hevc/h. 265
Y.-H. Ho, G.-L. Jin, Y. Liang, W.-H. Peng, and X. Li · 2021
Later among the works it cites.