کد مقاله کد نشریه سال انتشار مقاله انگلیسی نسخه تمام متن
5102705 1480089 2017 30 صفحه PDF دانلود رایگان
عنوان انگلیسی مقاله ISI
From language identification to language distance
ترجمه فارسی عنوان
از شناسایی زبان به فاصله زبان
موضوعات مرتبط
مهندسی و علوم پایه ریاضیات فیزیک ریاضی
چکیده انگلیسی
In this paper, we define two quantitative distances to measure how far apart two languages are. The distance measure that we have identified as more accurate is based on the perplexity of n-gram models extracted from text corpora. An experiment to compare forty-four European languages has been performed. For this purpose, we computed the distances for all the possible language pairs and built a network whose nodes are languages and edges are distances. The network we have built on the basis of linguistic distances represents the current map of similarities and divergences among the main languages of Europe.
ناشر
Database: Elsevier - ScienceDirect (ساینس دایرکت)
Journal: Physica A: Statistical Mechanics and its Applications - Volume 484, 15 October 2017, Pages 152-162
نویسندگان
, , ,