Every reference with a DOI in the deposited reference list resolved to a known
work in Crossref or DataCite at the dated check, and none carried a retraction,
withdrawal, or removal notice.
The 29 checked references that resolve
resolves10.1109/TPAMI.2017.2699184DeepLab: Semantic Image Segmentation with Deep Convolutional Nets, Atrous Convolution, and Fully Connected CRFs
resolves10.1007/s11263-014-0777-6Indoor Scene Understanding with RGB-D Images: Bottom-up Segmentation, Object Detection and Semantic Segmentation
resolves10.1109/CVPR.2013.79Perceptual Organization and Recognition of Indoor Scenes from RGB-D Images
resolves10.1109/CVPR.2017.345Knowing When to Look: Adaptive Attention via a Visual Sentinel for Image Captioning
resolves10.1109/CVPR.2017.556Learning Object Interactions and Descriptions for Semantic Image Segmentation
resolves10.1109/CVPR.2017.122Discriminative Bimodal Networks for Visual Localization and Detection with Natural Language Queries
The 9 references without a DOI — listed, not checked
no DOI — not checkedChen, L.C., Papandreou, G., Schroff, F., Adam, H.: Rethinking atrous convolution for semantic image segmentation. CoRR (2017)
no DOI — not checkedLiu, W., Rabinovich, A., Berg, A.C.: ParseNet: looking wider to see better. CoRR abs/1506.04579 (2015)
no DOI — not checkedLu, J., Yang, J., Batra, D., Parikh, D.: Hierarchical question-image co-attention for visual question answering. In: NIPS (2016)
no DOI — not checkedLuo, B., Li, H., Meng, F., Wu, Q., Huang, C.: Video object segmentation via global consistency aware query strategy. IEEE Trans. Multimed. PP(99), 1 (2017)
no DOI — not checkedMeng, F., Li, H., Wu, Q., Luo, B., Huang, C., Ngan, K.: Globally measuring the similarity of superpixels by binary edge maps for superpixel clustering. IEEE Trans. Circuits Syst. Video Technol. PP(99), 1 (2016)
no DOI — not checkedPeng, Z., Zhang, R., Liang, X., Liu, X., Lin, L.: Geometric scene parsing with hierarchical LSTM. In: Proceedings of the Twenty-Fifth International Joint Conference on Artificial Intelligence, pp. 3439–3445 (2016)
no DOI — not checkedSimonyan, K., Zisserman, A.: Very deep convolutional networks for large-scale image recognition. In: International Conference on Learning Representations (2015)
no DOI — not checkedXu, K., et al.: Show, attend and tell: neural image caption generation with visual attention. In: International Conference on Machine Learning, pp. 2048–2057 (2015)
no DOI — not checkedYu, F., Koltun, V.: Multi-scale context aggregation by dilated convolutions. In: International Conference on Learning Representations (2016)
checked 2026-08-09 — re-checked daily as this page is visited;
titles and statuses come from Crossref and DataCite and are not part of the signed record
Both snippets point at the live badge image and link back to this page. The
badge re-renders from the daily check, so an embed never goes stale by more than a day of visits.