دانلود رایگان مقاله: تعویض ها و سطوح سطح شخصیت برای ترازبندی چندجملهای

کد مقاله	کد نشریه	سال انتشار	مقاله انگلیسی	نسخه تمام متن
6940682	1450017	2018	8 صفحه PDF	دانلود رایگان

عنوان انگلیسی مقاله ISI

Order embeddings and character-level convolutions for multimodal alignment

ترجمه فارسی عنوان

تعویض ها و سطوح سطح شخصیت برای ترازبندی چندجملهای

دانلود مقاله + سفارش ترجمه

دانلود مقاله ISI انگلیسی

رایگان برای ایرانیان

کلمات کلیدی

41A05 65D05 41A10 65D17

موضوعات مرتبط

مهندسی و علوم پایه مهندسی کامپیوتر چشم انداز کامپیوتر و تشخیص الگو

پیش نمایش مقاله

تعویض ها و سطوح سطح شخصیت برای ترازبندی چندجملهای

چکیده انگلیسی

With the novel and fast advances in the area of deep neural networks, several challenging image-based tasks have been recently approached by researchers in pattern recognition and computer vision. In this paper, we address one of these tasks, which is to match image content with natural language descriptions, sometimes referred as multimodal content retrieval. Such a task is particularly challenging considering that we must find a semantic correspondence between captions and the respective image, a challenge for both computer vision and natural language processing areas. For such, we propose a novel multimodal approach based solely on convolutional neural networks for aligning images with their captions by directly convolving raw characters. Our proposed character-based textual embeddings allow the replacement of both word-embeddings and recurrent neural networks for text understanding, saving processing time and requiring fewer learnable parameters. Our method is based on the idea of projecting both visual and textual information into a common embedding space. For training such embeddings we optimize a contrastive loss function that is computed to minimize order-violations between images and their respective descriptions. We achieve state-of-the-art performance in the most well-known image-text alignment datasets, namely MicrosoftÂ COCO, Flickr8k, and Flickr30k, with a method that is conceptually much simpler and that possesses considerably fewer parameters than current approaches.

ناشر

Database: Elsevier - ScienceDirect (ساینس دایرکت)
Journal: Pattern Recognition Letters - Volume 102, 15 January 2018, Pages 15-22

نویسندگان

JÃ´natas Wehrmann, Anderson Mattjie, Rodrigo C. Barros,

علوم انسانی و هنر

فنی، مهندسی و علوم پایه

پزشکی و سلامت

بیو تکنولوژی

پذیرش سفارش ترجمه

دانلود رایگان مقاله ISI : تعویض ها و سطوح سطح شخصیت برای ترازبندی چندجملهای

دسترسی سریع

ارتباط

English Website