Fetching the paper…
Reading the bibliography…
Bird's-Eye View (BEV) Perception has received increasing attention in recent years as it provides a concise and unified spatial representation across views and benefits a diverse set of downstream driving applications.
P. Isola, J.-Y. Zhu, T. Zhou, and A. A. Efros, “Image-To-Image Translation With Conditional Adversarial Networks,” in
2017
Earlier work this paper cites.
A. van den Oord, O. Vinyals, and k. kavukcuoglu, “Neural Discrete Representation Learning,” in
2017
Earlier work this paper cites.
K. Regmi and A. Borji, “Cross-View Image Synthesis Using Conditional GANs,” in
2018
Earlier work this paper cites.
M.-F. Chang, J. Lambert, P. Sangkloy, J. Singh, S. Bak, A. Hartnett, D. Wang, P. Carr, S. Lucey, D. Ramanan, and J. Hays, “Argoverse: 3D Tracking and Forecasting With Rich Maps,” in
2019
Earlier work this paper cites.
I. Loshchilov and F. Hutter, “Decoupled Weight Decay Regularization,” Jan. 2019
2019
Earlier work this paper cites.
A. Amini, I. Gilitschenski, J. Phillips, J. Moseyko, R. Banerjee, S. Karaman, and D. Rus, “Learning Robust Control Policies for End-to-End Autonomous Driving From Data-Driven Simulation,”
2020
Earlier work this paper cites.
J. Li, X. Zhang, C. Jia, J. Xu, L. Zhang, Y. Wang, S. Ma, and W. Gao, “Direct Speech-to-image Translation,”
2020
Earlier work this paper cites.
O. Wiles, G. Gkioxari, R. Szeliski, and J. Johnson, “SynSin: End-to-End View Synthesis From a Single Image,” in
2020
Earlier work this paper cites.
H. Caesar, V. Bankiti, A. H. Lang, S. Vora, V. E. Liong, Q. Xu, A. Krishnan, Y. Pan, G. Baldan, and O. Beijbom, “nuScenes: A Multimodal Dataset for Autonomous Driving,” in
2020
Earlier work this paper cites.
J. Philion and S. Fidler, “Lift, Splat, Shoot: Encoding Images from Arbitrary Camera Rigs by Implicitly Unprojecting to 3D,” in
2020
Earlier work this paper cites.
M. Zaheer, G. Guruganesh, K. A. Dubey, J. Ainslie, C. Alberti, S. Ontanon, P. Pham, A. Ravula, Q. Wang, L. Yang, and A. Ahmed, “Big Bird: Transformers for Longer Sequences,” in
2020
Earlier work this paper cites.
Y. Yuan, X. Chen, and J. Wang, “Object-Contextual Representations for Semantic Segmentation,” in
2020
Cited alongside, same era.
O. Makansi, O. Cicek, Y. Marrakchi, and T. Brox, “On Exposing the Challenging Long Tail in Future Prediction of Traffic Actors,”
2021
Cited alongside, same era.
J. Breitenstein, J.-A. Termöhlen, D. Lipinski, and T. Fingscheidt, “Corner Cases for Visual Perception in Automated Driving: Some Guidance on Detection Approaches,”
2021
Cited alongside, same era.
S. W. Kim, J. Philion, A. Torralba, and S. Fidler, “DriveGAN: Towards a Controllable High-Quality Neural Simulation,” in
2021
Cited alongside, same era.
A. Ramesh, M. Pavlov, G. Goh, S. Gray, C. Voss, A. Radford, M. Chen, and I. Sutskever, “Zero-Shot Text-to-Image Generation,” in
2021
Cited alongside, same era.
2022
Later among the works it cites.
N. Gosala and A. Valada, “Bird’s-Eye-View Panoptic Segmentation Using Monocular Frontal View Images,”
2022
Later among the works it cites.
B. Zhou and P. Krähenbühl, “Cross-view Transformers for real-time Map-view Semantic Segmentation,” in
2022
Later among the works it cites.
R. Xu, Z. Tu, H. Xiang, W. Shao, B. Zhou, and J. Ma, “CoBEVT: Cooperative bird’s eye view semantic segmentation with sparse transformers,” in
2022
Later among the works it cites.
X. Ren and X. Wang, “Look Outside the Room: Synthesizing A Consistent Long-Term 3D Scene Video from A Single Image,” in
2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Z. Li, J. Wu, I. Koh, Y. Tang, and L. Sun, “Image Synthesis from Layout with Locality-Aware Mask Adaption,” in
2021
Cited alongside, same era.
A. Toker, Q. Zhou, M. Maximov, and L. Leal-Taixe, “Coming Down to Earth: Satellite-to-Street View Synthesis for Geo-Localization,” in
2021
Cited alongside, same era.
R. Rombach, P. Esser, and B. Ommer, “Geometry-Free View Synthesis: Transformers and no 3D Priors,” in
2021
Cited alongside, same era.
P. Esser, R. Rombach, and B. Ommer, “Taming Transformers for High-Resolution Image Synthesis,” in
2021
Cited alongside, same era.
B. Wilson, W. Qi, T. Agarwal, J. Lambert, J. Singh, S. Khandelwal, B. Pan, R. Kumar, A. Hartnett, J. K. Pontes, D. Ramanan, P. Carr, and J. Hays, “Argoverse 2: Next Generation Datasets for Self-Driving Perception and Forecasting,” in
2021
Cited alongside, same era.
J. Sun, Z. Shen, Y. Wang, H. Bao, and X. Zhou, “LoFTR: Detector-Free Local Feature Matching With Transformers,” in
2021
Cited alongside, same era.
A. Radford, J. Wu, R. Child, D. Luan, D. Amodei, and I. Sutskever, “Language Models are Unsupervised Multitask Learners,” p. 24
Cited in the paper.
G. Parmar, R. Zhang, and J.-Y. Zhu, “On aliased resizing and surprising subtleties in GAN evaluation,” in
2022
Later among the works it cites.
Q. Li, Z. Peng, L. Feng, Q. Zhang, Z. Xue, and B. Zhou, “Metadrive: Composing diverse driving scenarios for generalizable reinforcement learning,”
2022
Later among the works it cites.
Z. Li, W. Wang, H. Li, E. Xie, C. Sima, T. Lu, Y. Qiao, and J. Dai, “BEVFormer: Learning Bird’s-Eye-View Representation from Multi-camera Images via Spatiotemporal Transformers,” in
2022
Later among the works it cites.
2023
Closest in time.
L. Feng, Q. Li, Z. Peng, S. Tan, and B. Zhou, “TrafficGen: Learning to Generate Diverse and Realistic Traffic Scenarios,” in
2023
Closest in time.
Q. Li, Z. Peng, L. Feng, Z. Liu, C. Duan, W. Mo, and B. Zhou, “ScenarioNet: Open-Source Platform for Large-Scale Traffic Scenario Simulation and Modeling,” Oct. 2023
2023
Closest in time.