Supervised Learning
Linearly Independent Sets in Vector Spaces induced by Kernels • /r/MachineLearning
I hope this post is okay (if not let me know). I'm attaching a pdf which rigorously defines my question. Briefly, what I'm wondering is this - for the set of data points {x1,...,xp} in a vector space, (say, Rn) under what conditions is the set {k(x1,),...,k(xp,)} (where k(,) is a kernel function) independent? What conditions must the set {x1,...,xp} and the kernel function have to ensure independence? If there isn't an immediate answer to this question I'll happily take recommendations for mathematical reading towards trying to answer this question.
Fearless Frenchman breaks hoverboard record, sets sights on the clouds
A fearless Frenchman, Franky Zapata, thinks one day people will be able to ride his hoverboard to pick up bread in the morning (it's a French thing). The jet ski champion on Saturday set a new Guinness World Record for the farthest hoverboard flight – yes, just like in the movies – off the coast of Sausset-les-Pins in the south of France. Mr. Zapata rode the 1,000 horsepower drone, standing on top of it, for 7,388 feet, or more than a mile. He hovered 165 feet above the surface of the water, "trailed by a fleet of boats and jet skis," as Guinness reports. His feat shattered the previous hoverboard travel record of 905 feet and 2 inches, set last year by Canadian inventor Catalin Alexandru Duru.
How To Extract Feature Vectors From Deep Neural Networks In Python Caffe
Convolutional Neural Networks are great at identifying all the information that makes an image distinct. When we train a deep neural network in Caffe to classify images, we specify a multilayered neural network with different types of layers like convolution, rectified linear unit, softmax loss, and so on. The last layer is the output layer that gives us the output tag with the corresponding confidence value. But sometimes it's useful for us to extract the feature vectors from various layers and use it for other purposes. Let's see how to do it in Python Caffe, shall we?
Multi-Instance Multi-Label Class Discovery: A Computational Approach for Assessing Bird Biodiversity
Briggs, Forrest (Facebook, Inc.) | Fern, Xiaoli Z. (Oregon State University) | Raich, Raviv (Oregon State University) | Betts, Matthew (Oregon State University)
Briggs et al. (2012b) proposed to represent audio Bioacoustic monitoring is a rapidly growing field, where the recordings of bird sound in the multi-instance multi-label goal is to learn about organisms such as birds and marine (MIML) framework (Zhou et al. 2012). In this formulation, mammals, by applying signal processing and machine learning an audio recording is transformed to a spectrogram, to audio recordings. In this paper, we consider the problem then automatically segmented into a collection of regions of class discovery from bird bioacoustics data. Given believed to be distinct utterances of bird sound. Each segment a large collection of audio recordings of birds (and other is then described by a feature vector that characterizes sounds in the environment), our goal is to automatically select its shape, texture, and time/frequency profiles. A recording a subset of recordings to be manually labeled by human is represented as a set of segment feature vectors (instances).
Creating Images by Learning Image Semantics Using Vector Space Models
Heath, Derrall (Brigham Young University) | Ventura, Dan (Brigham Young University)
When dealing with images and semantics, most computational systems attempt to automatically extract meaning from images. Here we attempt to go the other direction and autonomously create images that communicate concepts. We present an enhanced semantic model that is used to generate novel images that convey meaning. We employ a vector space model and a large corpus to learn vector representations of words and then train the semantic model to predict word vectors that could describe a given image. Once trained, the model autonomously guides the process of rendering images that convey particular concepts. A significant contribution is that, because of the semantic associations encoded in these word vectors, we can also render images that convey concepts on which the model was not explicitly trained. We evaluate the semantic model with an image clustering technique and demonstrate that the model is successful in creating images that communicate semantic relationships.
Aggregating Inter-Sentence Information to Enhance Relation Extraction
Zheng, Hao (Beihang University) | Li, Zhoujun (Beihang University) | Wang, Senzhang (Beihang University) | Yan, Zhao ( Beihang University ) | Zhou, Jianshe ( Capital Normal University )
Previous work for relation extraction from free text is mainly based on intra-sentence information. As relations might be mentioned across sentences, inter-sentence information can be leveraged to improve distantly supervised relation extraction. To effectively exploit inter-sentence information, we propose a ranking based approach, which first learns a scoring function based on a listwise learning-to-rank model and then uses it for multi-label relation extraction. Experimental results verify the effectiveness of our method for aggregating information across sentences. Additionally, to further improve the ranking of high-quality extractions, we propose an effective method to rank relations from different entity pairs. This method can be easily integrated into our overall relation extraction framework, and boosts the precision significantly.
A Generative Model of Words and Relationships from Multiple Sources
Hyland, Stephanie L. (Weill Cornell Graduate School of Medical Sciences/Memorial Sloan Kettering Cancer Center) | Karaletsos, Theofanis (Memorial Sloan Kettering Cancer Center) | Rätsch, Gunnar (Memorial Sloan Kettering Cancer Center)
Neural language models are a powerful tool to embed words into semantic vector spaces. However, learning such models generally relies on the availability of abundant and diverse training examples. In highly specialised domains this requirement may not be met due to difficulties in obtaining a large corpus, or the limited range of expression in average use. Such domains may encode prior knowledge about entities in a knowledge base or ontology. We propose a generative model which integrates evidence from diverse data sources, enabling the sharing of semantic information. We achieve this by generalising the concept of co-occurrence from distributional semantics to include other relationships between entities or words, which we model as affine transformations on the embedding space. We demonstrate the effectiveness of this approach by outperforming recent models on a link prediction task and demonstrating its ability to profit from partially or fully unobserved data training labels. We further demonstrate the usefulness of learning from different data sources with overlapping vocabularies.
Robustness of Bayesian Pool-Based Active Learning Against Prior Misspecification
Cuong, Nguyen Viet (National University of Singapore) | Ye, Nan (Queensland University of Technology) | Lee, Wee Sun (National University of Singapore)
We study the robustness of active learning (AL) algorithms against prior misspecification: whether an algorithm achieves similar performance using a perturbed prior as compared to using the true prior. In both the average and worst cases of the maximum coverage setting, we prove that all alpha-approximate algorithms are robust (i.e., near alpha-approximate) if the utility is Lipschitz continuous in the prior. We further show that robustness may not be achieved if the utility is non-Lipschitz. This suggests we should use a Lipschitz utility for AL if robustness is required. For the minimum cost setting, we can also obtain a robustness result for approximate AL algorithms. Our results imply that many commonly used AL algorithms are robust against perturbed priors. We then propose the use of a mixture prior to alleviate the problem of prior misspecification. We analyze the robustness of the uniform mixture prior and show experimentally that it performs reasonably well in practice.
Who are alike? Use BigObject feature vector to find similarities
Cluster Analysis is a common technique to group a set of objects in the way that the objects in the same group share certain attributes. It's commonly used in marketing and sales planning to define market segmentations. Here at BigObject we adopt a simple approach to exploring the similarities between objects. We simply calculate the "Feature Vector" based on given attributes and use the score to determine which objects are "alike." This is a simple example to show how to use BigObject to extract product features and then find similar products in your retail data.
Human dominoes record broken
On Thursday, Aaron's Inc., a Maryland-based appliance and electronics company, set a new Guinness World Record for the "largest human mattress dominoes" chain with 1,200 participants taking a total of 13 minutes and 38 seconds to complete the larger than life feat. Event organizers used two exhibit halls covering 70,000 square feet to set up 34 rows of mattresses. The first mattress was pushed over by Aaron's CEO John Robinson. "Breaking a Guinness World Records title has been a great team building event for the associates we have attending our National Managers meeting," said Robinson at the event. The event not only broke a world record but supported a great cause.