Improving the polarity of text through word2vec embedding for primary classical arabic sentiment...- 学术资源搜索

Improving the polarity of text through word2vec embedding for primary classical arabic sentiment analysis

NE Aoumeur, Z Li, EM Alshari - Neural processing letters, 2023 - Springer

Neural processing letters, 2023•Springer

Over the past decade, Sentiment analysis has attracted significant researcher attention.
Despite a huge number of studies in this field, Sentiment analysis of authors' books
(classical Arabic) with extracting the embedding features has not yet been done. The recent
feature extraction of Arabic text depends on the frequency of the words within the corpus
without extracting the relation between these words. This paper aims to create a new
classical Arabic dataset CASAD from many art books by collecting sentences from several …

Abstract

Over the past decade, Sentiment analysis has attracted significant researcher attention. Despite a huge number of studies in this field, Sentiment analysis of authors’ books (classical Arabic) with extracting the embedding features has not yet been done. The recent feature extraction of Arabic text depends on the frequency of the words within the corpus without extracting the relation between these words. This paper aims to create a new classical Arabic dataset CASAD from many art books by collecting sentences from several stories with human-expert labeling. Additionally, the feature extraction of those datasets is created by word embedding techniques equivalent to Word2vec that are able to extract the deep relation which means features of the formal Arabic language. These features are evaluated by several types of machine learning for classical Arabic, for example, support vector machines (SVM), Logistic Regression (LR), Naive Bayes (NB) K-Nearest Neighbors (KNN), Latent Dirichlet Allocation (LDA) and Classification And Regression Trees (CART). Moreover, statistical methods such as validation and reliability are applied to evaluate this dataset’s label. Finally, our experiments evaluated the classification rate of the feature-extraction matrices in two and three classes using six machine-learning algorithms for tenfold cross-validation that showed that the Logistic Regression with Word2Vec approach is the most accurate in predicting topic-polarity occurrence.

Springer

展开收起

被引用次数：19 相关文章所有 7 个版本

以上显示的是最相近的搜索结果。查看全部搜索结果

高级搜索

QQ 群

Improving the polarity of text through word2vec embedding for primary classical arabic sentiment analysis

引用