Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

Attentive Relational Networks for Mapping Images to Scene Graphs

Type de record:

paperdataset
Créateur:
Qi,Li,YanWan
Hôte:avatar
Scene graph generation refers to the task of automatically mapping an image into a semantic structural graph, which requires correctly labeling each extracted object and their interaction relationships. Despite the recent success in object detection using deep learning techniques, inferring complex contextual relationships and structured graph representations from visual data remains a challenging topic. In this study, we propose a novel Attentive Relational Network that consists of two key modules with an object detection backbone to approach this problem. The first module is a semantic transformation module utilized to capture semantic embedded relation features, by translating visual features and linguistic features into a common semantic space. The other module is a graph self-attention module introduced to embed a joint graph representation through assigning various importance weights to neighboring nodes. Finally, accurate scene graphs are produced by the relation inference module to recognize all entities and the corresponding relations. We evaluate our proposed method on the widely-adopted Visual Genome Dataset, and the results demonstrate the effectiveness and superiority of our model. Accepted by CVPR 2019

Visit

arxiv.org

Tasks

computer visionimage-text retrieval

Tags

Computer Vision and Pattern Recognition

Similaires

Attentive Sequence-to-Sequence Learning for Diacritic Restoration of Yorùbá Language TextAn End-to-End Scene Text Recognition for Bilingual TextDeep Active Learning Approach for Traffic Sign and Panel Guide Arabic-Latin Text Content Annotation in Natural Scene ImagesSource data for graphs and chartsRepresentation Learning for Texts and GraphsDeep Neural Networks and Transfer Learning for Food Crop Identification in UAV Images

Attentive Sequence-to-Sequence Learning for Diacritic Restoration of Yorùbá Language Text

Yorùbá is a widely spoken West African language with a writing system rich in tonal and orthographic diacritics. With very few exceptions, diacritics are omitted from electronic texts, due to limited device and application support. Diacritics provide morphological

An End-to-End Scene Text Recognition for Bilingual Text

Text localization and recognition from natural scene images has gained a lot of attention recently d

Deep Active Learning Approach for Traffic Sign and Panel Guide Arabic-Latin Text Content Annotation in Natural Scene Images

The detection and recognition of road traffic signs and panel guides content has become cha

Source data for graphs and charts

Topsoil arsenic (As) contamination threatens the ecological environment and human healt

Representation Learning for Texts and Graphs

[...] This thesis is situated between natural language processing and graph representation learning

Deep Neural Networks and Transfer Learning for Food Crop Identification in UAV Images

Accurate projections of seasonal agricultural output are essential for improving food security. Howe