TY - GEN
T1 - Work hard, play hard
T2 - 11th Workshop on Graph-Based Methods for Natural Language Processing, TextGraphs 2017, in conjunction with the 55th Annual Meeting of the Association for Computational Linguistics, ACL 2017
AU - Alkhereyf, Sakhar
AU - Rambow, Owen
N1 - Publisher Copyright:
© 2017 Association for Computational Linguistics
PY - 2020
Y1 - 2020
N2 - In this paper, we present an empirical study of email classification into two main categories “Business” and “Personal”. We train on the Enron email corpus, and test on the Enron and Avocado email corpora. We show that information from the email exchange networks improves the performance of classification. We represent the email exchange networks as social networks with graph structures. For this classification task, we extract social networks features from the graphs in addition to lexical features from email content and we compare the performance of SVM and Extra-Trees classifiers using these features. Combining graph features with lexical features improves the performance on both classifiers. We also provide manually annotated sets of the Avocado and Enron email corpora as a supplementary contribution.
AB - In this paper, we present an empirical study of email classification into two main categories “Business” and “Personal”. We train on the Enron email corpus, and test on the Enron and Avocado email corpora. We show that information from the email exchange networks improves the performance of classification. We represent the email exchange networks as social networks with graph structures. For this classification task, we extract social networks features from the graphs in addition to lexical features from email content and we compare the performance of SVM and Extra-Trees classifiers using these features. Combining graph features with lexical features improves the performance on both classifiers. We also provide manually annotated sets of the Avocado and Enron email corpora as a supplementary contribution.
UR - https://www.scopus.com/pages/publications/85081541281
M3 - Conference contribution
AN - SCOPUS:85081541281
T3 - Proceedings of TextGraphs@ACL 2017: The 11th Workshop on Graph-Based Methods for Natural Language Processing
SP - 57
EP - 65
BT - Proceedings of TextGraphs@ACL 2017
A2 - Riedl, Martin
A2 - Somasundaran, Swapna
A2 - Glavas, Goran
A2 - Hovy, Eduard
PB - Association for Computational Linguistics
Y2 - 3 August 2017
ER -