Fetching the paper…
Reading the bibliography…
In multitask learning, conflicts between task gradients are a frequent issue degrading a model's training performance.
Caruana, R.: Multitask learning. Machine Learning 28
1997
Earlier work this paper cites.
Deng, J., Dong, W., Socher, R., Li, L.J., Li, K., Fei-Fei, L.: Imagenet: A large-scale hierarchical image database. In: 2009 IEEE Conference on Computer Vision and Pattern Recognition. pp. 248–255 (2009). https://doi.org/10.1109/CVPR.2009.5206848
2009
Earlier work this paper cites.
Krizhevsky, A., Hinton, G.: Learning multiple layers of features from tiny images. Master’s thesis, Department of Computer Science, University of Toronto (2009)
2009
Earlier work this paper cites.
Liu, Z., Luo, P., Wang, X., Tang, X.: Deep learning face attributes in the wild. 2015 IEEE International Conference on Computer Vision (ICCV) pp. 3730–3738 (2014), https://api.semanticscholar.org/CorpusID:459456
2014
Earlier work this paper cites.
He, K., Zhang, X., Ren, S., Sun, J.: Deep residual learning for image recognition. 2016 IEEE Conference on Computer Vision and Pattern Recognition (CVPR) pp. 770–778 (2015), https://api.semanticscholar.org/CorpusID:206594692
2015
Earlier work this paper cites.
Kendall, A., Gal, Y., Cipolla, R.: Multi-task learning using uncertainty to weigh losses for scene geometry and semantics. 2018 IEEE/CVF Conference on Computer Vision and Pattern Recognition pp. 7482–7491 (2017), https://api.semanticscholar.org/CorpusID:4800342
2017
Earlier work this paper cites.
2017
Earlier work this paper cites.
2018
Earlier work this paper cites.
Guo, M., Haque, A., Huang, D.A., Yeung, S., Fei-Fei, L.: Dynamic Task Prioritization for Multitask Learning. In: Ferrari, V., Hebert, M., Sminchisescu, C., Weiss, Y. (eds.) Computer Vision – ECCV 2018, vol. 11220, pp. 282–299. Springer International Publishing, Cham (2018). https://doi.org/10.1007/978-3-030-01270-0_17, https://link.springer.com/10.1007/978-3-030-01270-0_17 , series Title: Lecture Notes in Computer Science
2018
Earlier work this paper cites.
Lin, T.Y., Goyal, P., Girshick, R., He, K., Dollár, P.: Focal loss for dense object detection. IEEE Transactions on Pattern Analysis and Machine Intelligence 42
2018
Cited alongside, same era.
Sener, O., Koltun, V.: Multi-task learning as multi-objective optimization. In: Bengio, S., Wallach, H., Larochelle, H., Grauman, K., Cesa-Bianchi, N., Garnett, R. (eds.) Advances in Neural Information Processing Systems. vol. 31. Curran Associates, Inc. (2018), https://proceedings.neurips.cc/paper_files/paper/2018/file/432aca3a1e345e339f35a30c8f65edce-Paper.pdf
2018
Cited alongside, same era.
Caesar, H., Bankiti, V., Lang, A.H., Vora, S., Liong, V., Xu, Q., Krishnan, A., Pan, Y., Baldan, G., Beijbom, O.: nuscenes: A multimodal dataset for autonomous driving. In: 2020 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR). pp. 11618–11628. IEEE Computer Society, Los Alamitos, CA, USA (jun 2020). https://doi.org/10.1109/CVPR42600.2020.01164, https://doi.ieeecomputersociety.org/10.1109/CVPR42600.2020.01164
2020
Cited alongside, same era.
Yu, T., Kumar, S., Gupta, A., Levine, S., Hausman, K., Finn, C.: Gradient surgery for multi-task learning. In: Proceedings of the 34th International Conference on Neural Information Processing Systems. NIPS’20, Curran Associates Inc., Red Hook, NY, USA (2020)
2020
Later among the works it cites.
Liu, B., Liu, X., Jin, X., Stone, P., Liu, Q.: Conflict-Averse Gradient Descent for Multi-task learning. In: Advances in Neural Information Processing Systems. vol. 34, pp. 18878–18890. Curran Associates, Inc. (2021), https://proceedings.neurips.cc/paper/2021/hash/9d27fdf2477ffbff837d73ef7ae23db9-Abstract.html
2021
Later among the works it cites.
Li, Z., Wang, W., Li, H., Xie, E., Sima, C., Lu, T., Qiao, Y., Dai, J.: Bevformer: Learning bird’s-eye-view representation from multi-camera images via spatiotemporal transformers. In: Avidan, S., Brostow, G., Cissé, M., Farinella, G.M., Hassner, T. (eds.) Computer Vision – ECCV 2022. pp. 1–18. Springer Nature Switzerland, Cham (2022)
2022
Later among the works it cites.
alphaXiv searches the wider corpus for related work and actual follow-ups.
alphaXiv is searching for related work…
Chen, Z., Ngiam, J., Huang, Y., Luong, T., Kretzschmar, H., Chai, Y., Anguelov, D.: Just pick a sign: Optimizing deep multitask models with gradient sign dropout (2020)
2020
Cited alongside, same era.
Leang, I., Sistu, G., Bürger, F., Bursuc, A., Yogamani, S.K.: Dynamic task weighting methods for multi-task networks in autonomous driving systems. 2020 IEEE 23rd International Conference on Intelligent Transportation Systems (ITSC) pp. 1–8 (2020), https://api.semanticscholar.org/CorpusID:210023792
2020
Cited alongside, same era.
Philion, J., Fidler, S.: Lift, splat, shoot: Encoding images from arbitrary camera rigs by implicitly unprojecting to 3d. In: Proceedings of the European Conference on Computer Vision (2020)
2020
Cited alongside, same era.
Tseng, W.C.: github.com/weichengtseng/pytorch-pcgrad.git (2020), https://github.com/WeiChengTseng/Pytorch-PCGrad.git
2020
Cited alongside, same era.
Wang, Y., Chao, W.L., Garg, D., Hariharan, B., Campbell, M., Weinberger, K.Q.: Pseudo-lidar from visual depth estimation: Bridging the gap in 3d object detection for autonomous driving (2020)
2020
Cited alongside, same era.
2020
Cited alongside, same era.
2022
Later among the works it cites.
Xie, E., Yu, Z., Zhou, D., Philion, J., Anandkumar, A., Fidler, S., Luo, P., Alvarez, J.M.: M 2 BEV: Multi-camera joint 3D detection and segmentation with unified birds-eye view representation. In: ArXiv Preprint (April 2022)
2022
Later among the works it cites.
Yang, C., Chen, Y., Tian, H., Tao, C., Zhu, X., Zhang, Z., Huang, G., Li, H., Qiao, Y., Lu, L., Zhou, J., Dai, J.: Bevformer v2: Adapting modern image backbones to bird’s-eye-view recognition via perspective supervision. 2023 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) pp. 17830–17839 (2022), https://api.semanticscholar.org/CorpusID:253708369
2022
Later among the works it cites.
Bin-ze: Bevformer_segmentation_detection (2023), https://github.com/Bin-ze/BEVFormer_segmentation_detection.git
2023
Later among the works it cites.
Li, Z., Yu, Z., Wang, W., Anandkumar, A., Lu, T., Alvarez, J.M.: Fb-bev: Bev representation from forward-backward view transformations. In: 2023 IEEE/CVF International Conference on Computer Vision (ICCV). pp. 6896–6905. IEEE Computer Society, Los Alamitos, CA, USA (oct 2023). https://doi.org/10.1109/ICCV51070.2023.00637, https://doi.ieeecomputersociety.org/10.1109/ICCV51070.2023.00637
2023
Later among the works it cites.