Enter your keyword

2-s2.0-84946690676

[vc_empty_space][vc_empty_space]

Predicting latent attributes by extracting lexical and sociolinguistics features from user tweets

Hawari M.A.A.a, Khodra M.L.a

a School of Electrical Engineering and Informatics, Institut Teknologi Bandung, Bandung, Indonesia

[vc_row][vc_column][vc_row_inner][vc_column_inner][vc_separator css=”.vc_custom_1624529070653{padding-top: 30px !important;padding-bottom: 30px !important;}”][/vc_column_inner][/vc_row_inner][vc_row_inner layout=”boxed”][vc_column_inner width=”3/4″ css=”.vc_custom_1624695412187{border-right-width: 1px !important;border-right-color: #dddddd !important;border-right-style: solid !important;border-radius: 1px !important;}”][vc_empty_space][megatron_heading title=”Abstract” size=”size-sm” text_align=”text-left”][vc_column_text]© 2014 IEEE.Twitter user profile information is very useful for various fields such as marketing, HRD, advertising, and personalization. Since user profile provided by Twitter is very limited, some latent attributes such as gender, age, work, or interest should be predicted. In this paper, we aim to predict those four latent attributes using her/his tweet and bio data by employing machine learning techniques. We conduct experiments in order to find the best algorithm, weighting method, minimal frequency number, preprocess, for each latent attribute we predicts. We also compare the accuracy of lexical feature and sociolinguistic feature classification models. Our experiment shows that SVM is the best performer and lexical feature models perform better than sociolinguistic feature models.[/vc_column_text][vc_empty_space][vc_separator css=”.vc_custom_1624528584150{padding-top: 25px !important;padding-bottom: 25px !important;}”][vc_empty_space][megatron_heading title=”Author keywords” size=”size-sm” text_align=”text-left”][vc_column_text]Feature classification,latent attribute,Lexical features,Machine learning techniques,Personalizations,sociolinguistic feature,Twitter,Weighting methods[/vc_column_text][vc_empty_space][vc_separator css=”.vc_custom_1624528584150{padding-top: 25px !important;padding-bottom: 25px !important;}”][vc_empty_space][megatron_heading title=”Indexed keywords” size=”size-sm” text_align=”text-left”][vc_column_text]classification,latent attribute,lexical feature,machine learning,sociolinguistic feature,Twitter[/vc_column_text][vc_empty_space][vc_separator css=”.vc_custom_1624528584150{padding-top: 25px !important;padding-bottom: 25px !important;}”][vc_empty_space][megatron_heading title=”Funding details” size=”size-sm” text_align=”text-left”][vc_column_text][/vc_column_text][vc_empty_space][vc_separator css=”.vc_custom_1624528584150{padding-top: 25px !important;padding-bottom: 25px !important;}”][vc_empty_space][megatron_heading title=”DOI” size=”size-sm” text_align=”text-left”][vc_column_text]https://doi.org/10.1109/ICODSE.2014.7062666[/vc_column_text][/vc_column_inner][vc_column_inner width=”1/4″][vc_column_text]Widget Plumx[/vc_column_text][/vc_column_inner][/vc_row_inner][/vc_column][/vc_row][vc_row][vc_column][vc_separator css=”.vc_custom_1624528584150{padding-top: 25px !important;padding-bottom: 25px !important;}”][/vc_column][/vc_row]