Country
The Prevalence of Political Discourse in Non-Political Blogs
Munson, Sean A. (University of Michigan) | Resnick, Paul (University of Michigan)
Though political theorists have emphasized the importance of political discussion in non-political spaces, past study of online political discussion has focused on primarily political websites. Using a random sample from Blogger.com, we find that 25% of all political posts are from blogs that post about politics less than 20% of the time, because the vast majority of blogs post about politics some of the time but infrequently. Far from being taboo topics in those non- political blogs, political posts got slightly more comments than non-political posts in those same blogs, and the comments overwhelmingly engage the political topics of the post, mostly agreeing but frequently disagreeing as well. We argue that non-political spaces devoted primarily to personal diaries, hobbies, and other topics represent a substantial place of online political discussion and should be a site for further study.
Memes Online: Extracted, Subtracted, Injected, and Recollected
Simmons, Matthew P. (University of Michigan) | Adamic, Lada A. (Universiry of Michigan) | Adar, Eytan (University of Michigan)
Social media is playing an increasingly vital role in information dissemination. But with dissemination being more distributed, content often makes multiple hops, and consequently has opportunity to change. In this paper we focus on content that should be changing the least, namely quoted text. We find changes to be frequent, with their likelihood depending on the authority of the copied source and the type of site that is copying. We uncover patterns in the rate of appearance of new variants, their length, and popularity, and develop a simple model that is able to capture them. These patterns are distinct from ones produced when all copies are made from the same source, suggesting that information is evolving as it is being processed collectively in online social media.
Beyond Trending Topics: Real-World Event Identification on Twitter
Becker, Hila (Columbia University) | Naaman, Mor (Rutgers University) | Gravano, Luis (Columbia University)
User-contributed messages on social media sites such as Twitter have emerged aspowerful, real-time means of information sharing on the Web. These short messages tend to reflect a variety of events in real time, making Twitter particularly well suited as a source of real-time event content. In this paper, we explore approaches for analyzing the stream of Twitter messages to distinguish between messages about real-world events andnon-event messages. Our approach relies on a rich family of aggregatestatistics of topically similar message clusters. Large-scale experiments over millions of Twitter messages show the effectiveness of our approach for surfacing real-world event content on Twitter.
Event Summarization Using Tweets
Chakrabarti, Deepayan (Yahoo! Research) | Punera, Kunal (Yahoo! Research)
Twitter has become exceedingly popular, with hundreds of millions of tweets being posted every day on a wide variety of topics. This has helped make real-time search applications possible with leading search engines routinely displaying relevant tweets in response to user queries. Recent research has shown that a considerable fraction of these tweets are about "events," and the detection of novel events in the tweet-stream has attracted a lot of research interest. However, very little research has focused on properly displaying this real-time information about events. For instance, the leading search engines simply display all tweets matching the queries in reverse chronological order. In this paper we argue that for some highly structured and recurring events, such as sports, it is better to use more sophisticated techniques to summarize the relevant tweets. We formalize the problem of summarizing event-tweets and give a solution based on learning the underlying hidden state representation of the event via Hidden Markov Models. In addition, through extensive experiments on real-world data we show that our model significantly outperforms some intuitive and competitive baselines.
TweetTrader.net: Leveraging Crowd Wisdom in a Stock Microblogging Forum
Sprenger, Timm Oliver (Technische Universität München)
TweetTrader.net is a stock microblogging forum that leverages the wisdom of crowds to aggregate the information contained in stock-related tweets. Based on insights from academic research on stock microblogs, the application integrates inputs from text classification, user voting and a proprietary Stock Game in order to extract the sentiment (i.e., the bullishness) of online investors with respect to all publicly traded companies of the S&P 500.
Asked and Answered: On Qualities and Quantities of Answers in Online Q&A Sites
Logie, John (University of Minnesota) | Weinberg, Joseph (University of Minnesota) | Harper, F. Maxwell (University of Minnesota) | Konstan, Joseph A. (University of Minnesota)
This paper builds upon several recent research efforts that have explored the nature and qualities of questions asked on these social Q&A sites by offering a focused examination of answers posted to three of the most popular Q&A sites. Specifically, this paper examines sets of answers responding to specific types of questions and explores the degree to which question types are predictive of answer quantity and answer quality. Blending qualitative and quantitative methods, the paper builds upon rich coding of a representative sets of real questions โ drawn from Answerbag, (Ask) MetaFilter, and Yahoo! Answers โ in order to better understand whether the explicit and implicit theories and predictions drawn from coding of these questions were borne out in the corresponding answer sets found on these sites. Quantitative findings include data underscoring the general overall success of social Q&A sites in producing answers that can satisfy the needs of those who pose questions. Additionally, this paper presents a predictive model that can anticipate the archival value of answers based on the category and qualities of questions asked. Qualitative findings include an analysis of the variation in responses to questions that are primarily seeking objective, grounded information relative to those seeking subjective opinions.
Limits of Electoral Predictions Using Twitter
Gayo-Avello, Daniel (Universidad de Oviedo) | Metaxas, Panagiotis Takis (Wellesley College) | Mustafaraj, Eni (Wellesley College)
Using social media for political discourse is becoming common practice, especially around election time. One interesting aspect of this trend is the possibility of pulsing the publicโs opinion about the elections, and that has attracted the interest of many researchers and the press. Allegedly, predicting electoral outcomes from social media data can be feasible and even simple. Positive results have been reported, but without an analysis on what principle enables them. Our work puts to test the purported predictive power of socialmedia metrics against the 2010 US congressional elections. Here, we applied techniques that had reportedly led to positive election predictions in the past, on the Twitter data collected from the 2010 US congressional elections. Unfortunately, we find no correlation between the analysis results and the electoral outcomes, contradicting previous reports. Observing that 80 years of polling research would support our findings, we argue that one should not be accepting predictions about events using social media data as a black box. Instead, scholarly research should be accompanied by a model explaining the predictive power of social media, when there is one.
Online Identification and Tracking of Subspaces from Highly Incomplete Information
Balzano, Laura, Nowak, Robert, Recht, Benjamin
This work presents GROUSE (Grassmanian Rank-One Update Subspace Estimation), an efficient online algorithm for tracking subspaces from highly incomplete observations. GROUSE requires only basic linear algebraic manipulations at each iteration, and each subspace update can be performed in linear time in the dimension of the subspace. The algorithm is derived by analyzing incremental gradient descent on the Grassmannian manifold of subspaces. With a slight modification, GROUSE can also be used as an online incremental algorithm for the matrix completion problem of imputing missing entries of a low-rank matrix. GROUSE performs exceptionally well in practice both in tracking subspaces and as an online algorithm for matrix completion.
Rule-based query answering method for a knowledge base of economic crimes
We present a description of the PhD thesis which aims to propose a rule-based query answering method for relational data. In this approach we use an additional knowledge which is represented as a set of rules and describes the source data at concept (ontological) level. Queries are posed in the terms of abstract level. We present two methods. The first one uses hybrid reasoning and the second one exploits only forward chaining. These two methods are demonstrated by the prototypical implementation of the system coupled with the Jess engine. Tests are performed on the knowledge base of the selected economic crimes: fraudulent disbursement and money laundering.
Semantic-ontological combination of Business Rules and Business Processes in IT Service Management
Sellner, Alexander, Schwarz, Christopher, Zinser, Erwin
IT Service Management deals with managing a broad range of items related to complex system environments. As there is both, a close connection to business interests and IT infrastructure, the application of semantic expressions which are seamlessly integrated within applications for managing ITSM environments, can help to improve transparency and profitability. This paper focuses on the challenges regarding the integration of semantics and ontologies within ITSM environments. It will describe the paradigm of relationships and inheritance within complex service trees and will present an approach of ontologically expressing them. Furthermore, the application of SBVR-based rules as executable SQL triggers will be discussed. Finally, the broad range of topics for further research, derived from the findings, will be presented.