In this project, we compared Spanish BERT and Multilingual BERT in the Sentiment Analysis task.

Last update: Jan 03, 2022

Overview

Applying BERT Fine Tuning to Sentiment Classification on Amazon Reviews

Abstract

Sentiment analysis has made great progress in recent years, due to the fact that companies want to have a better understanding of how their products are classified by their consumers. However, despite the great advances that emerge in the field of artificial intelligence to solve this task, the most robust models are found in the English language. In the present work, we compare two Artificial Intelligence models that have monolingual and Multilingual approaches, which are Spanish BERT and Multilingual BERT, models based on BERT's transformer Architecture, to which the fine tuned technique was applied for the task of Sentiment analysis on the Amazon reviews dataset in Spanish using the accuracy and F1 score metrics. Finally, it was found that the Spanish BERT model has the best results for the sentiment analysis task on the Amazon reviews dataset in Spanish.

this paper is available here

Pipeline

Prerequisites

Linux / Window
Python3

Clone this Repository

git clone https://github.com/alexliqu09/Sentiment-Analysis-on-Amazon-Reviews.git

Train model

If you want to train the models use the colab Notebooks

Beto
MBert

Run the work in local

If you want to proof the work , you should run the following commands:

First , Install requeriments file:

pip install -r requeriments.txt

Second , download the Weights of Beto & MBERT and put them in this directory
Third , Start Streamlit server:

streamlit run main.py

Note:

Local host : http://localhost:8501 
Network URL:  http://192.168.0.5:8501

Run with Docker 🐋

#Bulding docker image 

docker build -t bert .

#RUN container
docker run -t -p 5000:5000 --name betocontainer bert

open http://172.17.0.2:8501

If you find useful our work , please cite this paper:

@inproceedings{@lvrBERT,
  title={Applying BERT Fine Tuning to Sentiment Classification on Amazon Reviews},
  author={Lique, Alexander and Vásquez, Diego and Rios, Manuel },
  year={2021}
}

In this project, we compared Spanish BERT and Multilingual BERT in the Sentiment Analysis task.

Related tags

Overview

Applying BERT Fine Tuning to Sentiment Classification on Amazon Reviews

Abstract

Pipeline

Prerequisites

Clone this Repository

Train model

Run the work in local

Run with Docker 🐋

Owner

Alexander Leonardo Lique Lamas

Transformer-based Text Auto-encoder (T-TA) using TensorFlow 2.

SimCSE: Simple Contrastive Learning of Sentence Embeddings

BERT score for text generation

Sequence model architectures from scratch in PyTorch

Lumped-element impedance calculator and frequency-domain plotter.

ThinkTwice: A Two-Stage Method for Long-Text Machine Reading Comprehension

CrossNER: Evaluating Cross-Domain Named Entity Recognition (AAAI-2021)

TunBERT is the first release of a pre-trained BERT model for the Tunisian dialect using a Tunisian Common-Crawl-based dataset.

Reproducing the Linear Multihead Attention introduced in Linformer paper (Linformer: Self-Attention with Linear Complexity)

Constituency Tree Labeling Tool

Codes to pre-train Japanese T5 models

A repository to run gpt-j-6b on low vram machines (4.2 gb minimum vram for 2000 token context, 3.5 gb for 1000 token context). Model loading takes 12gb free ram.

A PyTorch-based model pruning toolkit for pre-trained language models

Natural Language Processing for Adverse Drug Reaction (ADR) Detection

An open collection of annotated voices in Japanese language

Simple multilingual lemmatizer for Python, especially useful for speed and efficiency

customer care chatbot made with Rasa Open Source.

This is the writeup of all the challenges from Advent-of-cyber-2019 of TryHackMe

SpikeX - SpaCy Pipes for Knowledge Extraction

QVHighlights: Detecting Moments and Highlights in Videos via Natural Language Queries