1. Aerts HJ, Velazquez ER, Leijenaar RT, Parmar C, Grossmann P, Carvalho S, Bussink J, Monshouwer R, Haibe-Kains B, Rietveld D et al (2014) Decoding tumour phenotype by noninvasive imaging using a quantitative radiomics approach. Nat Commun 5:1–9
2. Agrawal H, Desai K, Wang Y, Chen X, Jain R, Johnson M, Batra D, Parikh D, Lee S, Anderson P (2019) nocaps: novel object captioning at scale. In: Proceedings of the IEEE international conference on computer vision, Seoul, Korea, pp 8948–8957
3. Anderson P, Fernando B, Johnson M, Gould S (2016) SPICE: semantic propositional image caption evaluation. In: Proceedings of the European conference on computer vision, Amsterdam, Netherlands, pp 382–398
4. Bai S, An S (2018) A survey on automatic image caption generation. Neurocomputing 311:291–304
5. Banerjee S, Lavie A (2005) METEOR: an automatic metric for MT evaluation with improved correlation with human judgments. In: Proceedings of the workshop on intrinsic and extrinsic evaluation measures for machine translation and/or summarization of the annual conference of the association for computational linguistics, Ann Arbor, MI, USA, pp 65–72