Telecommunications
Predicting SLA Violations in Real Time using Online Machine Learning
Ahmed, Jawwad, Johnsson, Andreas, Yanggratoke, Rerngvit, Ardelius, John, Flinta, Christofer, Stadler, Rolf
Detecting faults and SLA violations in a timely manner is critical for telecom providers, in order to avoid loss in business, revenue and reputation. At the same time predicting SLA violations for user services in telecom environments is difficult, due to time-varying user demands and infrastructure load conditions. In this paper, we propose a service-agnostic online learning approach, whereby the behavior of the system is learned on the fly, in order to predict client-side SLA violations. The approach uses device-level metrics, which are collected in a streaming fashion on the server side. Our results show that the approach can produce highly accurate predictions (>90% classification accuracy and < 10% false alarm rate) in scenarios where SLA violations are predicted for a video-on-demand service under changing load patterns. The paper also highlight the limitations of traditional offline learning methods, which perform significantly worse in many of the considered scenarios.
Estimating an Activity Driven Hidden Markov Model
Meyer, David A., Shakeel, Asif
We define a Hidden Markov Model (HMM) in which each hidden state has time-dependent $\textit{activity levels}$ that drive transitions and emissions, and show how to estimate its parameters. Our construction is motivated by the problem of inferring human mobility on sub-daily time scales from, for example, mobile phone records.
Joint community and anomaly tracking in dynamic networks
Baingana, Brian, Giannakis, Georgios B.
Most real-world networks exhibit community structure, a phenomenon characterized by existence of node clusters whose intra-edge connectivity is stronger than edge connectivities between nodes belonging to different clusters. In addition to facilitating a better understanding of network behavior, community detection finds many practical applications in diverse settings. Communities in online social networks are indicative of shared functional roles, or affiliation to a common socio-economic status, the knowledge of which is vital for targeted advertisement. In buyer-seller networks, community detection facilitates better product recommendations. Unfortunately, reliability of community assignments is hindered by anomalous user behavior often observed as unfair self-promotion, or "fake" highly-connected accounts created to promote fraud. The present paper advocates a novel approach for jointly tracking communities while detecting such anomalous nodes in time-varying networks. By postulating edge creation as the result of mutual community participation by node pairs, a dynamic factor model with anomalous memberships captured through a sparse outlier matrix is put forth. Efficient tracking algorithms suitable for both online and decentralized operation are developed. Experiments conducted on both synthetic and real network time series successfully unveil underlying communities and anomalous nodes.
Kernel-Based Adaptive Online Reconstruction of Coverage Maps With Side Information
Kasparick, Martin, Cavalcante, Renato L. G., Valentin, Stefan, Stanczak, Slawomir, Yukawa, Masahiro
In this paper, we address the problem of reconstructing coverage maps from path-loss measurements in cellular networks. We propose and evaluate two kernel-based adaptive online algorithms as an alternative to typical offline methods. The proposed algorithms are application-tailored extensions of powerful iterative methods such as the adaptive projected subgradient method and a state-of-the-art adaptive multikernel method. Assuming that the moving trajectories of users are available, it is shown how side information can be incorporated in the algorithms to improve their convergence performance and the quality of the estimation. The complexity is significantly reduced by imposing sparsity-awareness in the sense that the algorithms exploit the compressibility of the measurement data to reduce the amount of data which is saved and processed. Finally, we present extensive simulations based on realistic data to show that our algorithms provide fast, robust estimates of coverage maps in real-world scenarios. Envisioned applications include path-loss prediction along trajectories of mobile users as a building block for anticipatory buffering or traffic offloading.
Semantic Enrichment of Mobile Phone Data Records Using Background Knowledge
Dashdorj, Zolzaya, Sobolevsky, Stanislav, Serafini, Luciano, Antonelli, Fabrizio, Ratti, Carlo
Every day, billions of mobile network events (i.e. CDRs) are generated by cellular phone operator companies. Latent in this data are inspiring insights about human actions and behaviors, the discovery of which is important because context-aware applications and services hold the key to user-driven, intelligent services, which can enhance our everyday lives such as social and economic development, urban planning, and health prevention. The major challenge in this area is that interpreting such a big stream of data requires a deep understanding of mobile network events' context through available background knowledge. This article addresses the issues in context awareness given heterogeneous and uncertain data of mobile network events missing reliable information on the context of this activity. The contribution of this research is a model from a combination of logical and statistical reasoning standpoints for enabling human activity inference in qualitative terms from open geographical data that aimed at improving the quality of human behaviors recognition tasks from CDRs. We use open geographical data, Openstreetmap (OSM), as a proxy for predicting the content of human activity in the area. The user study performed in Trento shows that predicted human activities (top level) match the survey data with around 93% overall accuracy. The extensive validation for predicting a more specific economic type of human activity performed in Barcelona, by employing credit card transaction data. The analysis identifies that appropriately normalized data on points of interest (POI) is a good proxy for predicting human economical activities, with 84% accuracy on average. So the model is proven to be efficient for predicting the context of human activity, when its total level could be efficiently observed from cell phone data records, missing contextual information however.
Inferring Social Status and Rich Club Effects in Enterprise Communication Networks
Dong, Yuxiao, Tang, Jie, Chawla, Nitesh, Lou, Tiancheng, Yang, Yang, Wang, Bai
Social status, defined as the relative rank or position that an individual holds in a social hierarchy, is known to be among the most important motivating forces in social behaviors. In this paper, we consider the notion of status from the perspective of a position or title held by a person in an enterprise. We study the intersection of social status and social networks in an enterprise. We study whether enterprise communication logs can help reveal how social interactions and individual status manifest themselves in social networks. To that end, we use two enterprise datasets with three communication channels --- voice call, short message, and email --- to demonstrate the social-behavioral differences among individuals with different status. We have several interesting findings and based on these findings we also develop a model to predict social status. On the individual level, high-status individuals are more likely to be spanned as structural holes by linking to people in parts of the enterprise networks that are otherwise not well connected to one another. On the community level, the principle of homophily, social balance and clique theory generally indicate a "rich club" maintained by high-status individuals, in the sense that this community is much more connected, balanced and dense. Our model can predict social status of individuals with 93% accuracy.
Country-scale Exploratory Analysis of Call Detail Records through the Lens of Data Grid Models
Guigourès, Romain, Gay, Dominique, Boullé, Marc, Clérot, Fabrice, Rossi, Fabrice
Call Detail Records (CDRs) are data recorded by telecommunications companies, consisting of basic informations related to several dimensions of the calls made through the network: the source, destination, date and time of calls. CDRs data analysis has received much attention in the recent years since it might reveal valuable information about human behavior. It has shown high added value in many application domains like e.g., communities analysis or network planning. In this paper, we suggest a generic methodology for summarizing information contained in CDRs data. The method is based on a parameter-free estimation of the joint distribution of the variables that describe the calls. We also suggest several well-founded criteria that allows one to browse the summary at various granularities and to explore the summary by means of insightful visualizations. The method handles network graph data, temporal sequence data as well as user mobility data stemming from original CDRs data. We show the relevance of our methodology for various case studies on real-world CDRs data from Ivory Coast.
Target-Dependent Churn Classification in Microblogs
Amiri, Hadi (University of Maryland) | III, Hal Daume (University of Maryland)
In particular, we investigate demographic business. Banks, telecommunication companies, airlines, Internet churn indicators (obtained from users of microposts), service providers, pay TV companies, and insurance content churn indicators (obtained from the textual firms etc., utilize customer churn or attrition rates as one of content of micro-posts), and context churn indicators (obtained their key business metrics. This metric is important as the from threads containing the micro-posts). We examine churn rate of a business is a good indicator of customer response factors that make this problem more challenging and investigate to services, pricing, and competitions. The ability to the performance of several state-of-the-art machine identify churny contents / behaviors can enable early intervention learning techniques on this problem. A challenging aspect processes (as part of retention campaigns) and ultimately of such classification task is that churny contents can be expressed a reduction in customer churn.
The Network Data Repository with Interactive Graph Analytics and Visualization
Rossi, Ryan (Purdue University) | Ahmed, Nesreen (Purdue University)
NetworkRepository (NR) is the first interactive data repository with a web-based platform for visual interactive analytics. Unlike other data repositories (e.g., UCI ML Data Repository, and SNAP), the network data repository (networkrepository.com) allows users to not only download, but to interactively analyze and visualize such data using our web-based interactive graph analytics platform. Users can in real-time analyze, visualize, compare, and explore data along many different dimensions. The aim of NR is to make it easy to discover key insights into the data extremely fast with little effort while also providing a medium for users to share data, visualizations, and insights. Other key factors that differentiate NR from the current data repositories is the number of graph datasets, their size, and variety. While other data repositories are static, they also lack a means for users to collaboratively discuss a particular dataset, corrections, or challenges with using the data for certain applications. In contrast, NR incorporates many social and collaborative aspects that facilitate scientific research, e.g., users can discuss each graph, post observations, and visualizations.
' 7 '/ - 0/ THE DESIGN OF LARGE MULTI-MICROPROCESSOR NEIVORKS Kjell G. Fnutsen1 Computer Science Department Stanford University August, 1977
The low cost of microprocessors today, and the future trend in both cost and performence, makes large microprocessor-networks very interesting. A net of a thousend processors or more can be built using present techniques. However, there are several problems in utilizing such a horee of processors in a rcesonebly efficient way. There also seem to be restrictions to the kinds of opplications that can be mapped on to such a system. One control mechanism that seems to be very useful, at least for some Artificial Intelligence type problems, is the CONTRACT NET, ISmith77]. In this paper we will first look at some of the desirable characteristics of a lerge multi-microprocessor net. Than we will describe several different organizations together with their advantages and problems. We are also discussing broadcasting in lattices using circuit switched minimum spanning trees.