-
- Vasileios Lampos
- University of Bristol, UK
-
- Nello Cristianini
- University of Bristol, UK
書誌事項
- 公開日
- 2012-09
- 権利情報
-
- https://www.acm.org/publications/policies/copyright_policy#Background
- DOI
-
- 10.1145/2337542.2337557
- 公開者
- Association for Computing Machinery (ACM)
この論文をさがす
説明
<jats:p> We present a general methodology for inferring the occurrence and magnitude of an event or phenomenon by exploring the rich amount of unstructured textual information on the social part of the Web. Having geo-tagged user posts on the microblogging service of <jats:italic>Twitter</jats:italic> as our input data, we investigate two case studies. The first consists of a benchmark problem, where actual levels of rainfall in a given location and time are inferred from the content of <jats:italic>tweets</jats:italic> . The second one is a real-life task, where we infer regional Influenza-like Illness rates in the effort of detecting timely an emerging epidemic disease. Our analysis builds on a statistical learning framework, which performs sparse learning via the bootstrapped version of LASSO to select a consistent subset of textual features from a large amount of candidates. In both case studies, selected features indicate close semantic correlation with the target topics and inference, conducted by regression, has a significant performance, especially given the short length --approximately one year-- of Twitter’s data time series. </jats:p>
収録刊行物
-
- ACM Transactions on Intelligent Systems and Technology
-
ACM Transactions on Intelligent Systems and Technology 3 (4), 1-22, 2012-09
Association for Computing Machinery (ACM)