Goto

Collaborating Authors

 Education


Continual learning: A comparative study on how to defy forgetting in classification tasks

arXiv.org Machine Learning

Artificial neural networks thrive in solving the classification problem for a particular rigid task, where the network resembles a static entity of knowledge, acquired through generalized learning behaviour from a distinct training phase. However, endeavours to extend this knowledge without targeting the original task usually result in a catastrophic forgetting of this task. Continual learning shifts this paradigm towards a network that can continually accumulate knowledge over different tasks without the need for retraining from scratch, with methods in particular aiming to alleviate forgetting. We focus on task-incremental classification, where tasks arrive in a batch-like fashion, and are delineated by clear boundaries. Our main contributions concern 1) a taxonomy and extensive overview of the state-of-the-art, 2) a novel framework to continually determine stability-plasticity trade-off of the continual learner, 3) a comprehensive experimental comparison of 10 state-of-the-art continual learning methods and 4 baselines. We empirically scrutinize which method performs best, both on balanced Tiny Imagenet and a large-scale unbalanced iNaturalist datasets. We study the influence of model capacity, weight decay and dropout regularization, and the order in which the tasks are presented, and qualitatively compare methods in terms of required memory, computation time and storage.


Advancing subgroup fairness via sleeping experts

arXiv.org Machine Learning

We study methods for improving fairness to subgroups in settings with overlapping populations and sequential predictions. Classical notions of fairness focus on the balance of some property across different populations. However, in many applications the goal of the different groups is not to be predicted equally but rather to be predicted well. We demonstrate that the task of satisfying this guarantee for multiple overlapping groups is not straightforward and show that for the simple objective of unweighted average of false negative and false positive rate, satisfying this for overlapping populations can be statistically impossible even when we are provided predictors that perform well separately on each subgroup. On the positive side, we show that when individuals are equally important to the different groups they belong to, this goal is achievable; to do so, we draw a connection to the sleeping experts literature in online learning. Motivated by the one-sided feedback in natural settings of interest, we extend our results to such a feedback model. We also provide a game-theoretic interpretation of our results, examining the incentives of participants to join the system and to provide the system full information about predictors they may possess. We end with several interesting open problems concerning the strength of guarantees that can be achieved in a computationally efficient manner.


Distributed Machine Learning on Mobile Devices: A Survey

arXiv.org Machine Learning

In recent years, mobile devices have gained increasingly development with stronger computation capability and larger storage. Some of the computation-intensive machine learning and deep learning tasks can now be run on mobile devices. To take advantage of the resources available on mobile devices and preserve users' privacy, the idea of mobile distributed machine learning is proposed. It uses local hardware resources and local data to solve machine learning sub-problems on mobile devices, and only uploads computation results instead of original data to contribute to the optimization of the global model. This architecture can not only relieve computation and storage burden on servers, but also protect the users' sensitive information. Another benefit is the bandwidth reduction, as various kinds of local data can now participate in the training process without being uploaded to the server. In this paper, we provide a comprehensive survey on recent studies of mobile distributed machine learning. We survey a number of widely-used mobile distributed machine learning methods. We also present an in-depth discussion on the challenges and future directions in this area. We believe that this survey can demonstrate a clear overview of mobile distributed machine learning and provide guidelines on applying mobile distributed machine learning to real applications.


Evaluating Effects of Tuition Fees: Lasso for the Case of Germany

arXiv.org Machine Learning

We study the effect of the introduction of university tuition fees on the enrollment behavior of students in Germany. For this, an appropriate Lasso-technique is crucial in order to identify the magnitude and significance of the effect due to potentially many relevant controlling factors and only a short time frame where fees existed. We show that a post-double selection strategy combined with stability selection determines a significant negative impact of fees on student enrollment and identifies relevant variables. This is in contrast to previous empirical studies and a plain linear panel regression which cannot detect any effect of tuition fees in this case. In our study, we explicitly deal with data challenges in the response variable in a transparent way and provide respective robust results. Moreover, we control for spatial cross-effects capturing the heterogeneity in the introduction scheme of fees across federal states ("Bundesl\"ander"), which can set their own educational policy. We also confirm the validity of our Lasso approach in a comprehensive simulation study.


Visualizing Movement Control Optimization Landscapes

arXiv.org Machine Learning

A large body of animation research focuses on optimization of movement control, either as action sequences or policy parameters. However, as closed-form expressions of the objective functions are often not available, our understanding of the optimization problems is limited. Building on recent work on analyzing neural network training, we contribute novel visualizations of high-dimensional control optimization landscapes; this yields insights into why control optimization is hard and why common practices like early termination and spline-based action parameterizations make optimization easier. For example, our experiments show how trajectory optimization can become increasingly ill-conditioned with longer trajectories, but parameterizing control as partial target states - e.g., target angles converted to torques using a PD-controller - can act as an efficient preconditioner. Both our visualizations and quantitative empirical data also indicate that neural network policy optimization scales better than trajectory optimization for long planning horizons. Our work advances the understanding of movement optimization and our visualizations should also provide value in educational use.


Causal Modeling for Fairness in Dynamical Systems

arXiv.org Artificial Intelligence

In this work, we present causal directed acyclic graphs (DAGs) as a unifying framework for the recent literature on fairness in dynamical systems. We advocate for the use of causal DAGs as a tool in both designing equitable policies and estimating their impacts. By visualizing models of dynamic unfairness graphically, we expose implicit causal assumptions which can then be more easily interpreted and scrutinized by domain experts. We demonstrate that this method of reinterpretation can be used to critique the robustness of an existing model/policy, or uncover new policy evaluation questions. Causal models also enable a rich set of options for evaluating a new candidate policy without incurring the risk of implementing the policy in the real world. We close the paper with causal analyses of several models from the recent literature, and provide an in-depth case study to demonstrate the utility of causal DAGs for modeling fairness in dynamical systems.


A literature review on current approaches and applications of fuzzy expert systems

arXiv.org Artificial Intelligence

The main purposes of this study are to distinguish the trends of research in publication exits for the utilisations of the fuzzy expert and knowledge-based systems that is done based on the classification of studies in the last decade. The present investigation covers 60 articles from related scholastic journals, International conference proceedings and some major literature review papers. Our outcomes reveal an upward trend in the up-to-date publications number, that is evidence of growing notoriety on the various applications of fuzzy expert systems. This raise in the reports is mainly in the medical neuro-fuzzy and fuzzy expert systems. Moreover, another most critical observation is that many modern industrial applications are extended, employing knowledge-based systems by extracting the experts' knowledge.


The Animal-AI Environment: Training and Testing Animal-Like Artificial Cognition

arXiv.org Artificial Intelligence

Recent advances in artificial intelligence have been strongly driven by the use of game environments for training and evaluating agents. Games are often accessible and versatile, with well-defined state-transitions and goals allowing for intensive training and experimentation. However, agents trained in a particular environment are usually tested on the same or slightly varied distributions, and solutions do not necessarily imply any understanding. If we want AI systems that can model and understand their environment, we need environments that explicitly test for this. Inspired by the extensive literature on animal cognition, we present an environment that keeps all the positive elements of standard gaming environments, but is explicitly designed for the testing of animal-like artificial cognition. All source-code is publicly available (see appendix).


TC3 SPONSOR SERIES: Finding Needles in a Haystack with Graph Databases and Machine Learning Telecom Council Blog

#artificialintelligence

You know a technology has reached a tipping point when your kids ask about it. This happened recently when my eighth grade daughter asked, "What is Machine Learning and why is it so important?". Answering her question, I explained how Machine Learning is part of AI, where we teach machines to reason and learn like human beings. I used the example of fraud detection. In many ways catching fraud is like finding needles in a haystack โ€“ you must sort and make sense of massive amounts of data in order to find your "needles" or in this case, your fraudsters.


Why we need to rethink education in the artificial intelligence age

#artificialintelligence

Artificial intelligence (AI) and emerging technologies (ET) are poised to transform modern society in profound ways. As with electricity in the last century, AI is an enabling technology that will animate everyday products and communications, endowing everything from cars to cameras with the ability to interact with the world around them, and with each other. These developments are just the beginning, and as AI/ET matures, it will have sweeping impacts on our work, security, politics, and very lives.1 These technologies are already impacting the world around us, as Darrell West and I wrote in our April 2018 piece "How artificial intelligence is transforming the world," and I highly recommend that anyone just discovering the topic of AI policy read it thoroughly. There, Darrell and I describe several important implications related to AI/ET, but chief among them is that these technology developments are on the cusp of ushering in a true revolution in human affairs at an increasingly fast pace. As AI continues to influence and shape existing industries and allows new ones to take root, its macro-level impact, particularly in the realm of economics, will become more and more apparent.