期刊文献+
共找到1篇文章
< 1 >
每页显示 20 50 100
Modeling Chinese Microblogs with Five Ws for Topic Hashtags Extraction
1
作者 Zhibin Zhao Jiahong Sun +4 位作者 Lan Yao Xun Wang Jiahong Chu Huan Liu Ge Yu 《Tsinghua Science and Technology》 SCIE EI CAS CSCD 2017年第2期135-148,共14页
Hashtags are important metadata in microblogs and are used to mark topics or index messages. However,statistics show that hashtags are absent from most microblogs. This poses great challenges for the retrieval and ana... Hashtags are important metadata in microblogs and are used to mark topics or index messages. However,statistics show that hashtags are absent from most microblogs. This poses great challenges for the retrieval and analysis of these tagless microblogs. In this paper, we summarize the similarity between microblogs and shortmessage-style news, and then propose an algorithm, named 5WTAG, for detecting microblog topics based on a model of five Ws(When, Where, Who, What, ho W). As five-W attributes are the core components in event description, it is guaranteed theoretically that 5WTAG can properly extract semantic topics from microblogs. We introduce the detailed procedure of the algorithm in this paper including spam microblog identification, microblog segmentation, and candidate hashtag construction. In addition, we propose a novel recommendation computing method for ranking candidate hashtags, which combines syntax and semantic analysis and observes the distribution of artificial topic hashtags. Finally, we conduct comprehensive experiments to verify the semantic correctness and completeness of the candidate hashtags, as well as the accuracy of the recommendation method using real data from Sina Weibo. 展开更多
关键词 hashtag microblog topic detection short-message-style news five Ws
原文传递
上一页 1 下一页 到第
使用帮助 返回顶部