科技报告详细信息
Event Detection in Twitter
Weng, Jianshu ; Yao, Yuxia ; Leonardi, Erwin ; Lee, Francis
HP Development Company
关键词: twitter;    event detection;    wavelet;   
RP-ID  :  HPL-2011-98
学科分类:计算机科学(综合)
美国|英语
来源: HP Labs
PDF
【 摘 要 】

Twitter, as a form of social media, is fast emerging in recent years. Users are using Twitter to report real-life events. This paper focuses on detecting those events by analyzing the text stream in Twitter. Although event detection has long been a research topic, the characteristics of Twitter make it a non- trivial task. Tweets reporting such events are usually overwhelmed by high flood of meaningless "babbles". Moreover, event detection algorithm needs to be scalable given the sheer amount of tweets. This paper attempts to tackle these challenges with EDCoW (Event Detection with Clustering of Wavelet-based Signals). EDCoW builds signals for individual words by applying wavelet analysis on the frequency-based raw signals of the words. It then filters away the trivial words by looking at their corresponding signal auto- correlations. The remaining words are then clustered to form events with a modularity-based graph partitioning technique. Experimental studies show promising result of EDCoW. We also present the design of a proof-of-concept system, which was used to analyze netizens' online discussion about Singapore General Election 2011.

【 预 览 】
附件列表
Files Size Format View
RO201804100002877LZ 399KB PDF download
  文献评价指标  
  下载次数:26次 浏览次数:45次