Goto

Collaborating Authors

 Technology


Uber in the market for a fleet of self-driving cars, source says

#artificialintelligence

Ride-hailing service Uber has sounded out car companies about placing a large order for self-driving cars, an auto industry source has said. "They wanted autonomous cars," the source, who declined to be named, said. "It seemed like they were shopping around." Loss-making Uber would make drastic savings on its biggest cost -- drivers -- if it were able to incorporate self-driving cars into its fleet. Volkswagen's Audi, Daimler's Mercedes-Benz, BMW and car industry suppliers Bosch and Continental are all working on technologies for autonomous or semi-autonomous cars.


Uber in the market for a fleet of self-driving cars, source says

#artificialintelligence

Shedding drivers would save Uber a lot of money. Ride-hailing service Uber has sounded out car companies about placing a large order for self-driving cars, an auto industry source has said. "They wanted autonomous cars," the source, who declined to be named, said. "It seemed like they were shopping around." Loss-making Uber would make drastic savings on its biggest cost -- drivers -- if it were able to incorporate self-driving cars into its fleet.


Oculus founder: Compared to sci-fi, future is 'going to be a lot more boring'

#artificialintelligence

When we imagine a future where humans and robots coexist, it doesn't take long for us to arrive at a conclusion where the human race tragically ends. A robot takeover usually occurs, followed by the inevitable enslavement of all humankind. But when it comes to the future and what will actually unfold, Palmer Luckey, founder of Oculus VR (which Facebook now owns) and inventor of the Oculus Rift virtual reality headset, isn't sweating it. "The reason I'm not creeped out is pretty simple," said Luckey, who sat down with Apple co-founder Steve Wozniak and Re/code journalist Kara Swisher on Saturday at the Silicon Valley Comic Con in San Jose, California. "A lot of people look to science-fiction for representations of technology. It can also be flawed."


Big data analytics and artificial intelligence come to the SMB as MasterCard integrates IBM's Watson

#artificialintelligence

Small and medium sized businesses are being targeted by IBM and MasterCard as they look to bring big data analytics insights to better understand their markets and consumers. A partnership has been formed by the two companies that sees MasterCard integrate IBM Watson Analytics into its platform, along with its own anonymised transaction data that is gathered through the payment company's Local Market Intelligence. This combination will bring artificial intelligence to its payments platform. The aim is to be able to offer SMBs insights on revenue, market share, customer demographics and competitors in a particular location and across multiple locations. The problem being tackled is that smaller merchants often don't have the resources to maximise data insights.


A Comparison Study of Nonlinear Kernels

arXiv.org Machine Learning

In this paper, we compare 5 different nonlinear kernels: min-max, RBF, fRBF (folded RBF), acos, and acos-$\chi^2$, on a wide range of publicly available datasets. The proposed fRBF kernel performs very similarly to the RBF kernel. Both RBF and fRBF kernels require an important tuning parameter ($\gamma$). Interestingly, for a significant portion of the datasets, the min-max kernel outperforms the best-tuned RBF/fRBF kernels. The acos kernel and acos-$\chi^2$ kernel also perform well in general and in some datasets achieve the best accuracies. One crucial issue with the use of nonlinear kernels is the excessive computational and memory cost. These days, one increasingly popular strategy is to linearize the kernels through various randomization algorithms. In our study, the randomization method for the min-max kernel demonstrates excellent performance compared to the randomization methods for other types of nonlinear kernels, measured in terms of the number of nonzero terms in the transformed dataset. Our study provides evidence for supporting the use of the min-max kernel and the corresponding randomized linearization method (i.e., the so-called "0-bit CWS"). Furthermore, the results motivate at least two directions for future research: (i) To develop new (and linearizable) nonlinear kernels for better accuracies; and (ii) To develop better linearization algorithms for improving the current linearization methods for the RBF kernel, the acos kernel, and the acos-$\chi^2$ kernel. One attempt is to combine the min-max kernel with the acos kernel or the acos-$\chi^2$ kernel. The advantages of these two new and tuning-free nonlinear kernels are demonstrated vias our extensive experiments.


Data Augmentation via Levy Processes

arXiv.org Machine Learning

If a document is about travel, we may expect that short snippets of the document should also be about travel. We introduce a general framework for incorporating these types of invariances into a discriminative classifier. The framework imagines data as being drawn from a slice of a Lรฉvy process. If we slice the Lรฉvy process at an earlier point in time, we obtain additional pseudo-examples, which can be used to train the classifier. We show that this scheme has two desirable properties: it preserves the Bayes decision boundary, and it is equivalent to fitting a generative model in the limit where we rewind time back to 0. Our construction captures popular schemes such as Gaussian feature noising and dropout training, as well as admitting new generalizations. Black-box discriminative classifiers such as logistic regression, neural networks, and SVMs are the go-to solution in machine learning: they are simple to apply and often perform well. However, an expert may have additional knowledge to exploit, often taking the form of a certain family of transformations that should usually leave labels fixed. For example, in object recognition, an image of a cat rotated, translated, and peppered with a small amount of noise is probably still a cat.


Analysis of Crowdsourced Sampling Strategies for HodgeRank with Sparse Random Graphs

arXiv.org Machine Learning

Crowdsourcing enables researchers to conduct social experiments on a heterogenous set of participants and at a lower economic cost than conventional laboratory studies. For example, researchers can harness internet users to conduct user studies on their personal computers. Among various approaches to conduct subjective tests, pairwise comparisons are expected to yield more reliable results. However, in crowdsourced studies, the individuals performing the ratings are diverse compared to more controlled settings, which is difficult to control for using traditional experimental designs; researchers have recently proposed several randomized methods to conduct user studies [1, 2, 3], which accommodate incomplete and imbalanced data. HodgeRank, as an application of combinatorial Hodge theory to the preference or rank aggregation problem from pairwise comparison data, possibly being incomplete and imbalanced, was first introduced by [4], and inspired a series of studies in statistical ranking [5, 6, 7, 8]. Hodge theory has also found applications in game theory [9] and computer vision [10, 11], in addition to traditional applications in fluid mechanics [12] etc. HodgeRank formulates the ranking problem in terms of the discrete Hodge decomposition of the pairwise data and shows that it can be decomposed into three orthogonal components: a gradient flow representing a global rating (optimal in the L


Predictive Interval Models for Non-parametric Regression

arXiv.org Machine Learning

Having a regression model, we are interested in finding two-sided intervals that are guaranteed to contain at least a desired proportion of the conditional distribution of the response variable given a specific combination of predictors. We name such intervals predictive intervals. This work presents a new method to find two-sided predictive intervals for non-parametric least squares regression without the homoscedasticity assumption. Our predictive intervals are built by using tolerance intervals on prediction errors in the query point's neighborhood. We proposed a predictive interval model test and we also used it as a constraint in our hyper-parameter tuning algorithm. This gives an algorithm that finds the smallest reliable predictive intervals for a given dataset. We also introduce a measure for comparing different interval prediction methods yielding intervals having different size and coverage. These experiments show that our methods are more reliable, effective and precise than other interval prediction methods.


Variational Autoencoders for Feature Detection of Magnetic Resonance Imaging Data

arXiv.org Machine Learning

Independent component analysis (ICA), as an approach to the blind source-separation (BSS) problem, has become the de-facto standard in many medical imaging settings. Despite successes and a large ongoing research effort, the limitation of ICA to square linear transformations have not been overcome, so that general INFOMAX is still far from being realized. As an alternative, we present feature analysis in medical imaging as a problem solved by Helmholtz machines, which include dimensionality reduction and reconstruction of the raw data under the same objective, and which recently have overcome major difficulties in inference and learning with deep and nonlinear configurations. We demonstrate one approach to training Helmholtz machines, variational auto-encoders (VAE), as a viable approach toward feature extraction with magnetic resonance imaging (MRI) data.


How Robust are Reconstruction Thresholds for Community Detection?

arXiv.org Machine Learning

The stochastic block model is one of the oldest and most ubiquitous models for studying clustering and community detection. In an exciting sequence of developments, motivated by deep but non-rigorous ideas from statistical physics, Decelle et al. conjectured a sharp threshold for when community detection is possible in the sparse regime. Mossel, Neeman and Sly and Massoulie proved the conjecture and gave matching algorithms and lower bounds. Here we revisit the stochastic block model from the perspective of semirandom models where we allow an adversary to make `helpful' changes that strengthen ties within each community and break ties between them. We show a surprising result that these `helpful' changes can shift the information-theoretic threshold, making the community detection problem strictly harder. We complement this by showing that an algorithm based on semidefinite programming (which was known to get close to the threshold) continues to work in the semirandom model (even for partial recovery). This suggests that algorithms based on semidefinite programming are robust in ways that any algorithm meeting the information-theoretic threshold cannot be. These results point to an interesting new direction: Can we find robust, semirandom analogues to some of the classical, average-case thresholds in statistics? We also explore this question in the broadcast tree model, and we show that the viewpoint of semirandom models can help explain why some algorithms are preferred to others in practice, in spite of the gaps in their statistical performance on random models.