Learning the Dynamics of Visual Relational Reasoning via Reinforced Path Routing-Reference-Cited by-同舟云学术

Learning the Dynamics of Visual Relational Reasoning via Reinforced Path Routing

Published:2022-06-28 Issue:1 Volume:36 Page:1122-1130
ISSN:2374-3468
Container-title:Proceedings of the AAAI Conference on Artificial Intelligence
language:
Short-container-title:AAAI

Author:

Jing Chenchen,Jia Yunde,Wu Yuwei,Li Chuanhao,Wu Qi

Abstract

Reasoning is a dynamic process. In cognitive theories, the dynamics of reasoning refers to reasoning states over time after successive state transitions. Modeling the cognitive dynamics is of utmost importance to simulate human reasoning capability. In this paper, we propose to learn the reasoning dynamics of visual relational reasoning by casting it as a path routing task. We present a reinforced path routing method that represents an input image via a structured visual graph and introduces a reinforcement learning based model to explore paths (sequences of nodes) over the graph based on an input sentence to infer reasoning results. By exploring such paths, the proposed method represents reasoning states clearly and characterizes state transitions explicitly to fully model the reasoning dynamics for accurate and transparent visual relational reasoning. Extensive experiments on referring expression comprehension and visual question answering demonstrate the effectiveness of our method.

Publisher

Association for the Advancement of Artificial Intelligence (AAAI)

Subject

General Medicine

Cited by 3 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Dual-Branch Collaborative Learning for Visual Question Answering;Lecture Notes in Computer Science;2024

2. InterREC: An Interpretable Method for Referring Expression Comprehension;IEEE Transactions on Multimedia;2023

3. Maintaining Reasoning Consistency in Compositional Visual Question Answering;2022 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR);2022-06