Filter hashtag context through an original data cleaning method

Abstract : Nowadays, social networks are one of the most used means of communication. For example, the social network Twitter has nearly 100 million active users who post about 500 million messages per day. Sharing information on this platform is unique because messages are limited in characters number. Faced with this limitation, users express themselves briefly and use sometimes a hashtag that summarizes the general idea of the message. Nevertheless, hashtags are noisy data because they do not respect any linguistic rule, may have several meanings, and their use is not under control. In this work, we tackle the problem of hashtag context which may have useful applications in several fields like information recommendation or information classification. In this paper, we propose an original data cleaning method to extract the most relevant neighbor hashtags of a hashtag. We test our method with a dataset containing hashtags related to several topics (such as sport, music, technology, etc.) in order to show the efficacy and the robustness of our approach.
Liste complète des métadonnées

Cited literature [23 references]  Display  Hide  Download

https://hal.archives-ouvertes.fr/hal-01806156
Contributor : Didier Henry <>
Submitted on : Friday, June 1, 2018 - 4:38:49 PM
Last modification on : Tuesday, February 19, 2019 - 9:56:02 PM
Document(s) archivé(s) le : Sunday, September 2, 2018 - 4:27:36 PM

File

ANT2018.pdf
Publisher files allowed on an open archive

Identifiers

Collections

Citation

Didier Henry, Erick Stattner, Martine Collard. Filter hashtag context through an original data cleaning method. Procedia Computer Science, Elsevier, 2018, The 9th International Conference on Ambient Systems, Networks and Technologies (ANT 2018) / The 8th International Conference on Sustainable Energy Information Technology (SEIT-2018), 130, pp.464-471. 〈10.1016/j.procs.2018.04.050〉. 〈hal-01806156〉

Share

Metrics

Record views

81

Files downloads

78