Learning the Dynamics of Visual Relational Reasoning via Reinforced Path Routing

Chenchen Jing,Yuwei Wu,Chuanhao Li,Yunde Jia,Qi Wu

doi:10.1609/aaai.v36i1.19997

Abstract

Reasoning is a dynamic process. In cognitive theories, the dynamics of reasoning refers to reasoning states over time after successive state transitions. Modeling the cognitive dynamics is of utmost importance to simulate human reasoning capability. In this paper, we propose to learn the reasoning dynamics of visual relational reasoning by casting it as a path routing task. We present a reinforced path routing method that represents an input image via a structured visual graph and introduces a reinforcement learning based model to explore paths (sequences of nodes) over the graph based on an input sentence to infer reasoning results. By exploring such paths, the proposed method represents reasoning states clearly and characterizes state transitions explicitly to fully model the reasoning dynamics for accurate and transparent visual relational reasoning. Extensive experiments on referring expression comprehension and visual question answering demonstrate the effectiveness of our method.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Learning the Dynamics of Visual Relational Reasoning via Reinforced Path Routing

Abstract

Talk to us

Similar Papers

More From: Proceedings of the AAAI Conference on Artificial Intelligence

Lead the way for us

Journal: Proceedings of the AAAI Conference on Artificial Intelligence	Publication Date: Jun 28, 2022
Citations: 4

Similar Papers

Dynamic Graph Attention for Referring Expression Comprehension
Sibei Yang ... Yizhou Yu
-
Sibei Yang, et. al.Sibei Yang ... Yizhou Yu
01 Oct 2019
01 Oct 2019

Cops-Ref: A New Dataset and Task on Compositional Referring Expression Comprehension
Zhenfang Chen ... Lin Ma
-
Zhenfang Chen, et. al.Zhenfang Chen ... Lin Ma
01 Jun 2020
01 Jun 2020

A Proposal-Free One-Stage Framework for Referring Expression Comprehension and Generation via Dense Cross-Attention
Mengyang Sun ... Yanning Zhang
IEEE Transactions on Multimedia | VOL. 25
Mengyang Sun, et. al.Mengyang Sun ... Yanning Zhang
01 Jan 2023
IEEE Transactions on Multimedia | VOL. 25

Proposal-free One-stage Referring Expression via Grid-Word Cross-Attention
Wei Suo ... Peng Wang
-
Wei Suo, et. al.Wei Suo ... Peng Wang
01 Aug 2021
01 Aug 2021

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Learning the Dynamics of Visual Relational Reasoning via Reinforced Path Routing

Abstract

Talk to us

Similar Papers

More From: Proceedings of the AAAI Conference on Artificial Intelligence