Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Sergio Ortiz Rojas

Statistical sentiment analysis performance in Opinum

Mar 03, 2013

Boyan Bonev, Gema Ramírez-Sánchez, Sergio Ortiz Rojas

Figure 1 for Statistical sentiment analysis performance in Opinum

Figure 2 for Statistical sentiment analysis performance in Opinum

Figure 3 for Statistical sentiment analysis performance in Opinum

Figure 4 for Statistical sentiment analysis performance in Opinum

Abstract:The classification of opinion texts in positive and negative is becoming a subject of great interest in sentiment analysis. The existence of many labeled opinions motivates the use of statistical and machine-learning methods. First-order statistics have proven to be very limited in this field. The Opinum approach is based on the order of the words without using any syntactic and semantic information. It consists of building one probabilistic model for the positive and another one for the negative opinions. Then the test opinions are compared to both models and a decision and confidence measure are calculated. In order to reduce the complexity of the training corpus we first lemmatize the texts and we replace most named-entities with wildcards. Opinum presents an accuracy above 81% for Spanish opinions in the financial products domain. In this work we discuss which are the most important factors that have impact on the classification performance.

Via

Access Paper or Ask Questions