Devil's on the Edges: Selective Quad Attention for Scene Graph Generation

Jung, Deunsol; Kim, Sanghyun; Kim, Won Hwa; Cho, Minsu

Computer Science > Computer Vision and Pattern Recognition

arXiv:2304.03495 (cs)

[Submitted on 7 Apr 2023]

Title:Devil's on the Edges: Selective Quad Attention for Scene Graph Generation

Authors:Deunsol Jung, Sanghyun Kim, Won Hwa Kim, Minsu Cho

View PDF

Abstract:Scene graph generation aims to construct a semantic graph structure from an image such that its nodes and edges respectively represent objects and their relationships. One of the major challenges for the task lies in the presence of distracting objects and relationships in images; contextual reasoning is strongly distracted by irrelevant objects or backgrounds and, more importantly, a vast number of irrelevant candidate relations. To tackle the issue, we propose the Selective Quad Attention Network (SQUAT) that learns to select relevant object pairs and disambiguate them via diverse contextual interactions. SQUAT consists of two main components: edge selection and quad attention. The edge selection module selects relevant object pairs, i.e., edges in the scene graph, which helps contextual reasoning, and the quad attention module then updates the edge features using both edge-to-node and edge-to-edge cross-attentions to capture contextual information between objects and object pairs. Experiments demonstrate the strong performance and robustness of SQUAT, achieving the state of the art on the Visual Genome and Open Images v6 benchmarks.

Comments:	Accepted at CVPR 2023; Project page at this https URL
Subjects:	Computer Vision and Pattern Recognition (cs.CV)
Cite as:	arXiv:2304.03495 [cs.CV]
	(or arXiv:2304.03495v1 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.2304.03495

Submission history

From: Deunsol Jung [view email]
[v1] Fri, 7 Apr 2023 06:33:46 UTC (940 KB)

Computer Science > Computer Vision and Pattern Recognition

Title:Devil's on the Edges: Selective Quad Attention for Scene Graph Generation

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:Devil's on the Edges: Selective Quad Attention for Scene Graph Generation

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators