Rethinking Object Detection in Retail Stores-Reference-Cited by-同舟云学术

Rethinking Object Detection in Retail Stores

Published:2021-05-18 Issue:2 Volume:35 Page:947-954
ISSN:2374-3468
Container-title:Proceedings of the AAAI Conference on Artificial Intelligence
language:
Short-container-title:AAAI

Author:

Cai Yuanqiang,Wen Longyin,Zhang Libo,Du Dawei,Wang Weiqiang

Abstract

The conventional standard for object detection uses a bounding box to represent each individual object instance. However, it is not practical in the industry-relevant applications in the context of warehouses due to severe occlusions among groups of instances of the same categories. In this paper, we propose a new task, i.e., simultaneously object localization and counting, abbreviated as Locount, which requires algorithms to localize groups of objects of interest with the number of instances. However, there does not exist a dataset or benchmark designed for such a task. To this end, we collect a large-scale object localization and counting dataset with rich annotations in retail stores, which consists of 50,394 images with more than 1.9 million object instances in 140 categories. Together with this dataset, we provide a new evaluation protocol and divide the training and testing subsets to fairly evaluate the performance of algorithms for Locount, developing a new benchmark for the Locount task. Moreover, we present a cascaded localization and counting network as a strong baseline, which gradually classifies and regresses the bounding boxes of objects with the predicted numbers of instances enclosed in the bounding boxes, trained in an end-to-end manner. Extensive experiments are conducted on the proposed dataset to demonstrate its significance and the analysis is provided to indicate future directions. Dataset is available at https://isrc.iscas.ac.cn/gitlab/research/locount-dataset.

Publisher

Association for the Advancement of Artificial Intelligence (AAAI)

Subject

General Medicine

Cited by 9 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. X-Detect: explainable adversarial patch detection for object detectors in retail;Machine Learning;2024-06-19

2. Multimodal fine-grained grocery product recognition using image and OCR text;Machine Vision and Applications;2024-06-07

3. Object Detection Datasets and Product Detection Datasets Based on Deep Learning Methodology: Review;2023 7th International Conference On Computing, Communication, Control And Automation (ICCUBEA);2023-08-18

4. Augmenting Vibration-Based Customer-Product Interaction Recognition with Sparse Load Sensing;Proceedings of Cyber-Physical Systems and Internet of Things Week 2023;2023-05-09

5. Multimodal Fine-Grained Grocery Product Recognition Using Image and Ocr Text;2023