Traditional Machine Learning Models and Bidirectional Encoder Representations From Transformer (BERT)-Based Automatic Classification of Tweets About Eating Disorders: Algorithm Development and Validation Study

Jose Alberto Benítez-Andrades; Jose Manuel Alija-Perez; Maria Esther Vidal; Rafael Pastor-Vargas; María Teresa García-Ordas

doi:10.2196/34492

Details

Originalsprache	Englisch
Aufsatznummer	e34492
Seitenumfang	13
Fachzeitschrift	JMIR Medical Informatics
Jahrgang	10
Ausgabenummer	2
Publikationsstatus	Veröffentlicht - 1 Feb. 2022

Abstract

Background: Eating disorders affect an increasing number of people. Social networks provide information that can help. Objective: We aimed to find machine learning models capable of efficiently categorizing tweets about eating disorders domain. Methods: We collected tweets related to eating disorders, for 3 consecutive months. After preprocessing, a subset of 2000 tweets was labeled: (1) messages written by people suffering from eating disorders or not, (2) messages promoting suffering from eating disorders or not, (3) informative messages or not, and (4) scientific or nonscientific messages. Traditional machine learning and deep learning models were used to classify tweets. We evaluated accuracy, F1 score, and computational time for each model. Results: A total of 1,058,957 tweets related to eating disorders were collected. were obtained in the 4 categorizations, with The bidirectional encoder representations from transformer-based models had the best score among the machine learning and deep learning techniques applied to the 4 categorization tasks (F1 scores 71.1%-86.4%). Conclusions: Bidirectional encoder representations from transformer-based models have better performance, although their computational cost is significantly higher than those of traditional techniques, in classifying eating disorder-related tweets.

ASJC Scopus Sachgebiete

Medizin (insg.)
Gesundheitsinformatik
Gesundheitsberufe (insg.)
Gesundheits-Informationsmanagement

Ziele für nachhaltige Entwicklung

SDG 3 – Gute Gesundheit und Wohlergehen

Zitieren

Traditional Machine Learning Models and Bidirectional Encoder Representations From Transformer (BERT)-Based Automatic Classification of Tweets About Eating Disorders: Algorithm Development and Validation Study. / Benítez-Andrades, Jose Alberto; Alija-Perez, Jose Manuel; Vidal, Maria Esther et al.
in: JMIR Medical Informatics, Jahrgang 10, Nr. 2, e34492, 01.02.2022.

Publikation: Beitrag in Fachzeitschrift › Artikel › Forschung › Peer-Review

Benítez-Andrades, JA, Alija-Perez, JM, Vidal, ME, Pastor-Vargas, R & García-Ordas, MT 2022, 'Traditional Machine Learning Models and Bidirectional Encoder Representations From Transformer (BERT)-Based Automatic Classification of Tweets About Eating Disorders: Algorithm Development and Validation Study', JMIR Medical Informatics, Jg. 10, Nr. 2, e34492. https://doi.org/10.2196/34492

Benítez-Andrades, J. A., Alija-Perez, J. M., Vidal, M. E., Pastor-Vargas, R., & García-Ordas, M. T. (2022). Traditional Machine Learning Models and Bidirectional Encoder Representations From Transformer (BERT)-Based Automatic Classification of Tweets About Eating Disorders: Algorithm Development and Validation Study. JMIR Medical Informatics, 10(2), Artikel e34492. https://doi.org/10.2196/34492

Benítez-Andrades JA, Alija-Perez JM, Vidal ME, Pastor-Vargas R, García-Ordas MT. Traditional Machine Learning Models and Bidirectional Encoder Representations From Transformer (BERT)-Based Automatic Classification of Tweets About Eating Disorders: Algorithm Development and Validation Study. JMIR Medical Informatics. 2022 Feb 1;10(2):e34492. doi: 10.2196/34492

Benítez-Andrades, Jose Alberto ; Alija-Perez, Jose Manuel ; Vidal, Maria Esther et al. / Traditional Machine Learning Models and Bidirectional Encoder Representations From Transformer (BERT)-Based Automatic Classification of Tweets About Eating Disorders : Algorithm Development and Validation Study. in: JMIR Medical Informatics. 2022 ; Jahrgang 10, Nr. 2.

Download

@article{5ef20f346a294a9487992e4f7851721b,

title = "Traditional Machine Learning Models and Bidirectional Encoder Representations From Transformer (BERT)-Based Automatic Classification of Tweets About Eating Disorders: Algorithm Development and Validation Study",

abstract = "Background: Eating disorders affect an increasing number of people. Social networks provide information that can help. Objective: We aimed to find machine learning models capable of efficiently categorizing tweets about eating disorders domain. Methods: We collected tweets related to eating disorders, for 3 consecutive months. After preprocessing, a subset of 2000 tweets was labeled: (1) messages written by people suffering from eating disorders or not, (2) messages promoting suffering from eating disorders or not, (3) informative messages or not, and (4) scientific or nonscientific messages. Traditional machine learning and deep learning models were used to classify tweets. We evaluated accuracy, F1 score, and computational time for each model. Results: A total of 1,058,957 tweets related to eating disorders were collected. were obtained in the 4 categorizations, with The bidirectional encoder representations from transformer-based models had the best score among the machine learning and deep learning techniques applied to the 4 categorization tasks (F1 scores 71.1%-86.4%). Conclusions: Bidirectional encoder representations from transformer-based models have better performance, although their computational cost is significantly higher than those of traditional techniques, in classifying eating disorder-related tweets.",

keywords = "BERT, bidirectional encoder representations from transformer, classification, data, deep learning, diet, disorder, eating disorder, machine learning, mental health, model, natural language processing, NLP, nutrition, performance, social media, Twitter, weight",

author = "Ben{\'i}tez-Andrades, {Jose Alberto} and Alija-Perez, {Jose Manuel} and Vidal, {Maria Esther} and Rafael Pastor-Vargas and Garc{\'i}a-Ordas, {Mar{\'i}a Teresa}",

year = "2022",

month = feb,

day = "1",

doi = "10.2196/34492",

language = "English",

volume = "10",

number = "2",

}

Download

TY - JOUR

T1 - Traditional Machine Learning Models and Bidirectional Encoder Representations From Transformer (BERT)-Based Automatic Classification of Tweets About Eating Disorders

T2 - Algorithm Development and Validation Study

AU - Benítez-Andrades, Jose Alberto

AU - Alija-Perez, Jose Manuel

AU - Vidal, Maria Esther

AU - Pastor-Vargas, Rafael

AU - García-Ordas, María Teresa

PY - 2022/2/1

Y1 - 2022/2/1

N2 - Background: Eating disorders affect an increasing number of people. Social networks provide information that can help. Objective: We aimed to find machine learning models capable of efficiently categorizing tweets about eating disorders domain. Methods: We collected tweets related to eating disorders, for 3 consecutive months. After preprocessing, a subset of 2000 tweets was labeled: (1) messages written by people suffering from eating disorders or not, (2) messages promoting suffering from eating disorders or not, (3) informative messages or not, and (4) scientific or nonscientific messages. Traditional machine learning and deep learning models were used to classify tweets. We evaluated accuracy, F1 score, and computational time for each model. Results: A total of 1,058,957 tweets related to eating disorders were collected. were obtained in the 4 categorizations, with The bidirectional encoder representations from transformer-based models had the best score among the machine learning and deep learning techniques applied to the 4 categorization tasks (F1 scores 71.1%-86.4%). Conclusions: Bidirectional encoder representations from transformer-based models have better performance, although their computational cost is significantly higher than those of traditional techniques, in classifying eating disorder-related tweets.

AB - Background: Eating disorders affect an increasing number of people. Social networks provide information that can help. Objective: We aimed to find machine learning models capable of efficiently categorizing tweets about eating disorders domain. Methods: We collected tweets related to eating disorders, for 3 consecutive months. After preprocessing, a subset of 2000 tweets was labeled: (1) messages written by people suffering from eating disorders or not, (2) messages promoting suffering from eating disorders or not, (3) informative messages or not, and (4) scientific or nonscientific messages. Traditional machine learning and deep learning models were used to classify tweets. We evaluated accuracy, F1 score, and computational time for each model. Results: A total of 1,058,957 tweets related to eating disorders were collected. were obtained in the 4 categorizations, with The bidirectional encoder representations from transformer-based models had the best score among the machine learning and deep learning techniques applied to the 4 categorization tasks (F1 scores 71.1%-86.4%). Conclusions: Bidirectional encoder representations from transformer-based models have better performance, although their computational cost is significantly higher than those of traditional techniques, in classifying eating disorder-related tweets.

KW - BERT

KW - bidirectional encoder representations from transformer

KW - classification

KW - data

KW - deep learning

KW - diet

KW - disorder

KW - eating disorder

KW - machine learning

KW - mental health

KW - model

KW - natural language processing

KW - NLP

KW - nutrition

KW - performance

KW - social media

KW - Twitter

KW - weight

UR - http://www.scopus.com/inward/record.url?scp=85126462382&partnerID=8YFLogxK

U2 - 10.2196/34492

DO - 10.2196/34492

M3 - Article

AN - SCOPUS:85126462382

VL - 10

JO - JMIR Medical Informatics

JF - JMIR Medical Informatics

IS - 2

M1 - e34492

ER -

Research@Leibniz University

Traditional Machine Learning Models and Bidirectional Encoder Representations From Transformer (BERT)-Based Automatic Classification of Tweets About Eating Disorders: Algorithm Development and Validation Study

Autoren

Organisationseinheiten

Externe Organisationen

Details

Abstract

ASJC Scopus Sachgebiete

Ziele für nachhaltige Entwicklung

Zitieren