Using the forest to see the trees

Author:

Torralba A.1,Murphy K. P.2,Freeman W. T.1

Affiliation:

1. Massachusetts Institute of Technology, Cambridge, MA

2. University of British Columbia, Vancouver, Canada

Abstract

Recognizing objects in images is an active area of research in computer vision. In the last two decades, there has been much progress and there are already object recognition systems operating in commercial products. However, most of the algorithms for detecting objects perform an exhaustive search across all locations and scales in the image comparing local image regions with an object model. That approach ignores the semantic structure of scenes and tries to solve the recognition problem by brute force. In the real world, objects tend to covary with other objects, providing a rich collection of contextual associations. These contextual associations can be used to reduce the search space by looking only in places in which the object is expected to be; this also increases performance, by rejecting patterns that look like the target but appear in unlikely places. Most modeling attempts so far have defined the context of an object in terms of other previously recognized objects. The drawback of this approach is that inferring the context becomes as difficult as detecting each object. An alternative view of context relies on using the entire scene information holistically. This approach is algorithmically attractive since it dispenses with the need for a prior step of individual object recognition. In this paper, we use a probabilistic framework for encoding the relationships between context and object properties and we show how an integrated system provides improved performance. We view this as a significant step toward general purpose machine vision systems.

Funder

NGA

Office of Naval Research

Division of Information and Intelligent Systems

Publisher

Association for Computing Machinery (ACM)

Subject

General Computer Science

Cited by 65 articles. 订阅此论文施引文献 订阅此论文施引文献,注册后可以免费订阅5篇论文的施引文献,订阅后可以查看论文全部施引文献

1. A face retrieval technique combining large models and artificial neural networks;Concurrency and Computation: Practice and Experience;2024-03-25

2. Driving Environment Inference from POI of Navigation Map: Fuzzy Logic and Machine Learning Approaches;Sensors;2023-11-13

3. Context understanding in computer vision: A survey;Computer Vision and Image Understanding;2023-03

4. Research Review of Dispensing Based on Machine Vision;2022 4th International Conference on Applied Machine Learning (ICAML);2022-07

5. From Node to Graph: Joint Reasoning on Visual-Semantic Relational Graph for Zero-Shot Detection;2022 IEEE/CVF Winter Conference on Applications of Computer Vision (WACV);2022-01

同舟云学术

1.学者识别学者识别

2.学术分析学术分析

3.人才评估人才评估

"同舟云学术"是以全球学者为主线,采集、加工和组织学术论文而形成的新型学术文献查询和分析系统,可以对全球学者进行文献检索和人才价值评估。用户可以通过关注某些学科领域的顶尖人物而持续追踪该领域的学科进展和研究前沿。经过近期的数据扩容,当前同舟云学术共收录了国内外主流学术期刊6万余种,收集的期刊论文及会议论文总量共计约1.5亿篇,并以每天添加12000余篇中外论文的速度递增。我们也可以为用户提供个性化、定制化的学者数据。欢迎来电咨询!咨询电话:010-8811{复制后删除}0370

www.globalauthorid.com

TOP

Copyright © 2019-2024 北京同舟云网络信息技术有限公司
京公网安备11010802033243号  京ICP备18003416号-3