Oceania
AI reveals what the Kardashians would look like without cosmetic work
Artificial intelligence has predicted what the Kardashian-Jenner family would look like if they had aged naturally. The famous family, who are known for their love of cosmetic enhancements, appear very different in a digitally altered video that recently went viral on TikTok. The clip, created by popular Australian streamers Vandahood Live, estimates what Kim Kardashian, Kylie Jenner, Khloe Kardashian, Kris Jenner and Kourtney Kardashian would look like today without cosmetic intervention. A viral TikTok video has revealed using artificial intelligence (AI) what the Kardashian-Jenner family would look like if they had aged naturally. In the video, footage from last year's Keeping Up with the Kardashians finale special is played alongside doctored versions of the same clip.
Why the sci-fi dream of cryonics never died
The environment was something of a shift for Drake, who had spent the previous seven years as the medical response director of the Alcor Life Extension Foundation. Though it was the longtime leader in cryonics, Alcor was still a small nonprofit. It had been freezing the bodies and brains of its members, with the idea of one day bringing them back to life, since 1976. The foundation, and cryonics in general, had long survived outside of mainstream acceptance. Typically shunned by the scientific community, cryonics is best known for its appearance in sci-fi films like 2001: A Space Odyssey.
PseudoReasoner: Leveraging Pseudo Labels for Commonsense Knowledge Base Population
Fang, Tianqing, Do, Quyet V., Zhang, Hongming, Song, Yangqiu, Wong, Ginny Y., See, Simon
Commonsense Knowledge Base (CSKB) Population aims at reasoning over unseen entities and assertions on CSKBs, and is an important yet hard commonsense reasoning task. One challenge is that it requires out-of-domain generalization ability as the source CSKB for training is of a relatively smaller scale (1M) while the whole candidate space for population is way larger (200M). We propose PseudoReasoner, a semi-supervised learning framework for CSKB population that uses a teacher model pre-trained on CSKBs to provide pseudo labels on the unlabeled candidate dataset for a student model to learn from. The teacher can be a generative model rather than restricted to discriminative models as previous works. In addition, we design a new filtering procedure for pseudo labels based on influence function and the student model's prediction to further improve the performance. The framework can improve the backbone model KG-BERT (RoBERTa-large) by 3.3 points on the overall performance and especially, 5.3 points on the out-of-domain performance, and achieves the state-of-the-art. Codes and data are available at https://github.com/HKUST-KnowComp/PseudoReasoner.
Boosting Performance of a Baseline Visual Place Recognition Technique by Predicting the Maximally Complementary Technique
Malone, Connor, Hausler, Stephen, Fischer, Tobias, Milford, Michael
One recent promising approach to the Visual Place Recognition (VPR) problem has been to fuse the place recognition estimates of multiple complementary VPR techniques using methods such as SRAL and multi-process fusion. These approaches come with a substantial practical limitation: they require all potential VPR methods to be brute-force run before they are selectively fused. The obvious solution to this limitation is to predict the viable subset of methods ahead of time, but this is challenging because it requires a predictive signal within the imagery itself that is indicative of high performance methods. Here we propose an alternative approach that instead starts with a known single base VPR technique, and learns to predict the most complementary additional VPR technique to fuse with it, that results in the largest improvement in performance. The key innovation here is to use a dimensionally reduced difference vector between the query image and the top-retrieved reference image using this baseline technique as the predictive signal of the most complementary additional technique, both during training and inference. We demonstrate that our approach can train a single network to select performant, complementary technique pairs across datasets which span multiple modes of transportation (train, car, walking) as well as to generalise to unseen datasets, outperforming multiple baseline strategies for manually selecting the best technique pairs based on the same training data.
Attention Regularized Laplace Graph for Domain Adaptation
Luo, Lingkun, Chen, Liming, Hu, Shiqiang
In leveraging manifold learning in domain adaptation (DA), graph embedding-based DA methods have shown their effectiveness in preserving data manifold through the Laplace graph. However, current graph embedding DA methods suffer from two issues: 1). they are only concerned with preservation of the underlying data structures in the embedding and ignore sub-domain adaptation, which requires taking into account intra-class similarity and inter-class dissimilarity, thereby leading to negative transfer; 2). manifold learning is proposed across different feature/label spaces separately, thereby hindering unified comprehensive manifold learning. In this paper, starting from our previous DGA-DA, we propose a novel DA method, namely Attention Regularized Laplace Graph-based Domain Adaptation (ARG-DA), to remedy the aforementioned issues. Specifically, by weighting the importance across different sub-domain adaptation tasks, we propose the Attention Regularized Laplace Graph for class-aware DA, thereby generating the attention regularized DA. Furthermore, using a specifically designed FEEL strategy, our approach dynamically unifies alignment of the manifold structures across different feature/label spaces, thus leading to comprehensive manifold learning. Comprehensive experiments are carried out to verify the effectiveness of the proposed DA method, which consistently outperforms the state-of-the-art DA methods on 7 standard DA benchmarks, i.e., 37 cross-domain image classification tasks including object, face, and digit images. An in-depth analysis of the proposed DA method is also discussed, including sensitivity, convergence, and robustness.
A situated agent-based model to reveal irrigators' options behind their actions under institutional arrangements in Southern France
Richard, Bastien, Bontรฉ, Bruno, Barreteau, Olivier, Braud, Isabelle
There has been little exploration of the explicit simulation of the set of options of actors in agent-based models and its evolution over time. This study proposes to use affordances as intermediate entities between agents' environment and agent actions. We illustrated the approach on a typical gravity-fed network in the South-East of France to explore how the abandonment of traditional sharing of water changes the irrigators' options to irrigate. We simulated a typical dry year irrigation season under two institutional arrangements (i.e. traditional coordination through daily slots and its abandonment). Simulation results are consistent with field surveys, and reveal an increase in the number of internal conflicts among irrigators as the counterpart of the abandonment of traditional sharing of water. They also highlight the consequences of the heterogeneity of the irrigators' interests within the collective institution. The sensitivity analysis of the model allowed identification of optimal modalities of coordination, and a potential compromise between past and current institutional arrangements. The key benefits of using affordances in ABM lie in the study of their population dynamics for characterizing the interaction situations between actors and their environment and for better understanding the model dynamics.
MKIS-Net: A Light-Weight Multi-Kernel Network for Medical Image Segmentation
Khan, Tariq M., Arsalan, Muhammad, Robles-Kelly, Antonio, Meijering, Erik
Image segmentation is an important task in medical imaging. It constitutes the backbone of a wide variety of clinical diagnostic methods, treatments, and computer-aided surgeries. In this paper, we propose a multi-kernel image segmentation net (MKIS-Net), which uses multiple kernels to create an efficient receptive field and enhance segmentation performance. As a result of its multi-kernel design, MKIS-Net is a light-weight architecture with a small number of trainable parameters. Moreover, these multi-kernel receptive fields also contribute to better segmentation results. We demonstrate the efficacy of MKIS-Net on several tasks including segmentation of retinal vessels, skin lesion segmentation, and chest X-ray segmentation. The performance of the proposed network is quite competitive, and often superior, in comparison to state-of-the-art methods. Moreover, in some cases MKIS-Net has more than an order of magnitude fewer trainable parameters than existing medical image segmentation alternatives and is at least four times smaller than other light-weight architectures.
FLUTE: Figurative Language Understanding through Textual Explanations
Chakrabarty, Tuhin, Saakyan, Arkadiy, Ghosh, Debanjan, Muresan, Smaranda
Figurative language understanding has been recently framed as a recognizing textual entailment (RTE) task (a.k.a. natural language inference, or NLI). However, similar to classical RTE/NLI datasets, the current benchmarks suffer from spurious correlations and annotation artifacts. To tackle this problem, work on NLI has built explanation-based datasets such as e-SNLI, allowing us to probe whether language models are right for the right reasons.Yet no such data exists for figurative language, making it harder to assess genuine understanding of such expressions. To address this issue, we release FLUTE, a dataset of 9,000 figurative NLI instances with explanations, spanning four categories: Sarcasm, Simile, Metaphor, and Idioms. We collect the data through a model-in-the-loop framework based on GPT-3, crowd workers, and expert annotators. We show how utilizing GPT-3 in conjunction with human annotators (novices and experts) can aid in scaling up the creation of datasets even for such complex linguistic phenomena as figurative language. The baseline performance of the T5 model fine-tuned on FLUTE shows that our dataset can bring us a step closer to developing models that understand figurative language through textual explanations.
Neural Routing in Meta Learning
Cai, Jicang, Vahidian, Saeed, Wang, Weijia, Joneidi, Mohsen, Lin, Bill
Meta-learning often referred to as learning-to-learn is a promising notion raised to mimic human learning by exploiting the knowledge of prior tasks but being able to adapt quickly to novel tasks. A plethora of models has emerged in this context and improved the learning efficiency, robustness, etc. The question that arises here is can we emulate other aspects of human learning and incorporate them into the existing meta learning algorithms? Inspired by the widely recognized finding in neuroscience that distinct parts of the brain are highly specialized for different types of tasks, we aim to improve the model performance of the current meta learning algorithms by selectively using only parts of the model conditioned on the input tasks. In this work, we describe an approach that investigates task-dependent dynamic neuron selection in deep convolutional neural networks (CNNs) by leveraging the scaling factor in the batch normalization (BN) layer associated with each convolutional layer. The problem is intriguing because the idea of helping different parts of the model to learn from different types of tasks may help us train better filters in CNNs, and improve the model generalization performance. We find that the proposed approach, neural routing in meta learning (NRML), outperforms one of the well-known existing meta learning baselines on few-shot classification tasks on the most widely used benchmark datasets.
CLASP: Few-Shot Cross-Lingual Data Augmentation for Semantic Parsing
Rosenbaum, Andy, Soltan, Saleh, Hamza, Wael, Saffari, Amir, Damonte, Marco, Groves, Isabel
A bottleneck to developing Semantic Parsing (SP) models is the need for a large volume of human-labeled training data. Given the complexity and cost of human annotation for SP, labeled data is often scarce, particularly in multilingual settings. Large Language Models (LLMs) excel at SP given only a few examples, however LLMs are unsuitable for runtime systems which require low latency. In this work, we propose CLASP, a simple method to improve low-resource SP for moderate-sized models: we generate synthetic data from AlexaTM 20B to augment the training set for a model 40x smaller (500M parameters). We evaluate on two datasets in low-resource settings: English PIZZA, containing either 348 or 16 real examples, and mTOP cross-lingual zero-shot, where training data is available only in English, and the model must generalize to four new languages. On both datasets, we show significant improvements over strong baseline methods.