کد مقاله کد نشریه سال انتشار مقاله انگلیسی نسخه تمام متن
6865302 1439555 2018 9 صفحه PDF دانلود رایگان
عنوان انگلیسی مقاله ISI
A region-based image caption generator with refined descriptions
ترجمه فارسی عنوان
یک ژنراتور عنصر تصویر مبتنی بر منطقه با توضیحات تصفیه شده
کلمات کلیدی
نسل توصیف تصویر، شبکه های عصبی مصنوعی و مجدد، تولید شرح،
موضوعات مرتبط
مهندسی و علوم پایه مهندسی کامپیوتر هوش مصنوعی
چکیده انگلیسی
Describing the content of an image is a challenging task. To enable detailed description, it requires the detection and recognition of objects, people, relationships and associated attributes. Currently, the majority of the existing research relies on holistic techniques, which may lose details relating to important aspects in a scene. In order to deal with such a challenge, we propose a novel region-based deep learning architecture for image description generation. It employs a regional object detector, recurrent neural network (RNN)-based attribute prediction, and an encoder-decoder language generator embedded with two RNNs to produce refined and detailed descriptions of a given image. Most importantly, the proposed system focuses on a local based approach to further improve upon existing holistic methods, which relates specifically to image regions of people and objects in an image. Evaluated with the IAPR TC-12 dataset, the proposed system shows impressive performance and outperforms state-of-the-art methods using various evaluation metrics. In particular, the proposed system shows superiority over existing methods when dealing with cross-domain indoor scene images.
ناشر
Database: Elsevier - ScienceDirect (ساینس دایرکت)
Journal: Neurocomputing - Volume 272, 10 January 2018, Pages 416-424
نویسندگان
, , ,