Goto

Collaborating Authors

 Government


Adversarial Auto-Augment with Label Preservation: A Representation Learning Principle Guided Approach

arXiv.org Artificial Intelligence

Data augmentation is a critical contributing factor to the success of deep learning but heavily relies on prior domain knowledge which is not always available. Recent works on automatic data augmentation learn a policy to form a sequence of augmentation operations, which are still pre-defined and restricted to limited options. In this paper, we show that a prior-free autonomous data augmentation's objective can be derived from a representation learning principle that aims to preserve the minimum sufficient information of the labels. Given an example, the objective aims at creating a distant "hard positive example" as the augmentation, while still preserving the original label. We then propose a practical surrogate to the objective that can be optimized efficiently and integrated seamlessly into existing methods for a broad class of machine learning tasks, e.g., supervised, semi-supervised, and noisy-label learning. Unlike previous works, our method does not require training an extra generative model but instead leverages the intermediate layer representations of the end-task model for generating data augmentations. In experiments, we show that our method consistently brings non-trivial improvements to the three aforementioned learning tasks from both efficiency and final performance, either or not combined with strong pre-defined augmentations, e.g., on medical images when domain knowledge is unavailable and the existing augmentation techniques perform poorly.


Natural Language Deduction with Incomplete Information

arXiv.org Artificial Intelligence

A growing body of work studies how to answer a question or verify a claim by generating a natural language "proof": a chain of deductive inferences yielding the answer based on a set of premises. However, these methods can only make sound deductions when they follow from evidence that is given. We propose a new system that can handle the underspecified setting where not all premises are stated at the outset; that is, additional assumptions need to be materialized to prove a claim. By using a natural language generation model to abductively infer a premise given another premise and a conclusion, we can impute missing pieces of evidence needed for the conclusion to be true. Our system searches over two fringes in a bidirectional fashion, interleaving deductive (forward-chaining) and abductive (backward-chaining) generation steps. We sample multiple possible outputs for each step to achieve coverage of the search space, at the same time ensuring correctness by filtering low-quality generations with a round-trip validation procedure. Results on a modified version of the EntailmentBank dataset and a new dataset called Everyday Norms: Why Not? show that abductive generation with validation can recover premises across in- and out-of-domain settings.


When Bioprocess Engineering Meets Machine Learning: A Survey from the Perspective of Automated Bioprocess Development

arXiv.org Artificial Intelligence

Machine learning (ML) is becoming increasingly crucial in many fields of engineering but has not yet played out its full potential in bioprocess engineering. While experimentation has been accelerated by increasing levels of lab automation, experimental planning and data modeling are still largerly depend on human intervention. ML can be seen as a set of tools that contribute to the automation of the whole experimental cycle, including model building and practical planning, thus allowing human experts to focus on the more demanding and overarching cognitive tasks. First, probabilistic programming is used for the autonomous building of predictive models. Second, machine learning automatically assesses alternative decisions by planning experiments to test hypotheses and conducting investigations to gather informative data that focus on model selection based on the uncertainty of model predictions. This review provides a comprehensive overview of ML-based automation in bioprocess development. On the one hand, the biotech and bioengineering community should be aware of the potential and, most importantly, the limitation of existing ML solutions for their application in biotechnology and biopharma. On the other hand, it is essential to identify the missing links to enable the easy implementation of ML and Artificial Intelligence (AI) tools in valuable solutions for the bio-community.


Topology-aware Graph Neural Networks for Learning Feasible and Adaptive ac-OPF Solutions

arXiv.org Artificial Intelligence

Solving the optimal power flow (OPF) problem is a fundamental task to ensure the system efficiency and reliability in real-time electricity grid operations. We develop a new topology-informed graph neural network (GNN) approach for predicting the optimal solutions of real-time ac-OPF problem. To incorporate grid topology to the NN model, the proposed GNN-for-OPF framework innovatively exploits the locality property of locational marginal prices and voltage magnitude. Furthermore, we develop a physics-aware (ac-)flow feasibility regularization approach for general OPF learning. The advantages of our proposed designs include reduced model complexity, improved generalizability and feasibility guarantees. By providing the analytical understanding on the graph subspace stability under grid topology contingency, we show the proposed GNN can quickly adapt to varying grid topology by an efficient re-training strategy. Numerical tests on various test systems of different sizes have validated the prediction accuracy, improved flow feasibility, and topology adaptivity capability of our proposed GNN-based learning framework.


Towards Inter-character Relationship-driven Story Generation

arXiv.org Artificial Intelligence

In this paper, we introduce the task of modeling interpersonal relationships for story generation. For addressing this task, we propose Relationships as Latent Variables for Story Generation, (ReLiSt). ReLiSt generates stories sentence by sentence and has two major components - a relationship selector and a story continuer. The relationship selector specifies a latent variable to pick the relationship to exhibit in the next sentence and the story continuer generates the next sentence while expressing the selected relationship in a coherent way. Our automatic and human evaluations demonstrate that ReLiSt is able to generate stories with relationships that are more faithful to desired relationships while maintaining the content quality. The relationship assignments to sentences during inference bring interpretability to ReLiSt.


FRSUM: Towards Faithful Abstractive Summarization via Enhancing Factual Robustness

arXiv.org Artificial Intelligence

Despite being able to generate fluent and grammatical text, current Seq2Seq summarization models still suffering from the unfaithful generation problem. In this paper, we study the faithfulness of existing systems from a new perspective of factual robustness which is the ability to correctly generate factual information over adversarial unfaithful information. We first measure a model's factual robustness by its success rate to defend against adversarial attacks when generating factual information. The factual robustness analysis on a wide range of current systems shows its good consistency with human judgments on faithfulness. Inspired by these findings, we propose to improve the faithfulness of a model by enhancing its factual robustness. Specifically, we propose a novel training strategy, namely FRSUM, which teaches the model to defend against both explicit adversarial samples and implicit factual adversarial perturbations. Extensive automatic and human evaluation results show that FRSUM consistently improves the faithfulness of various Seq2Seq models, such as T5, BART.


Machine learning can guide experimental approaches for protein digestibility estimations

arXiv.org Artificial Intelligence

Food protein digestibility and bioavailability are critical aspects in addressing human nutritional demands, particularly when seeking sustainable alternatives to animal-based proteins. In this study, we propose a machine learning approach to predict the true ileal digestibility coefficient of food items. The model makes use of a unique curated dataset that combines nutritional information from different foods with FASTA sequences of some of their protein families. We extracted the biochemical properties of the proteins and combined these properties with embeddings from a Transformer-based protein Language Model (pLM). In addition, we used SHAP to identify features that contribute most to the model prediction and provide interpretability. This first AI-based model for predicting food protein digestibility has an accuracy of 90% compared to existing experimental techniques. With this accuracy, our model can eliminate the need for lengthy in-vivo or in-vitro experiments, making the process of creating new foods faster, cheaper, and more ethical.


Machine Learning Training in Jersey

#artificialintelligence

Currently, computer and technology occupations rank the highest and most favorable jobs and are expected to grow by 11% by 2029 as per US Bureau of Labor Statistics. As per Glassdoor, Machine Learning salary on average in the US is around $118,875 per year, and some professionals earning more than $150k annually. Indeed, another popular job search portal entitles Machine learning engineering a highly popular and among the best jobs to look out for in the coming decade. Some famous names hiring Machine Learning professionals are Amazon, Adobe, Apple, Google, Lockheed Martin, Spotify, Zoom, Bank of America, and PayPal. So, considering a career through a Machine Learning bootcamp can be all you need to step on the career ladder.


Insights: The role of artificial intelligence and blockchain in Web3

#artificialintelligence

How do you see the growth of artificial intelligence in the Middle East? The Middle East region has always delivered on ambitious visions. This region has realised that artificial intelligence (AI) will be ubiquitous in every part of life in the next five to 10 years, will bring huge cost savings and/or efficiency gains and the governments are looking at harnessing the power of AI as quickly as possible. What are the new trends that you have seen? Upskilling Industry wide upskilling programmes to have the basic mathematical skills of data science to understand AI.


Analysis: What's behind Iran's alleged drone deal with Russia?

Al Jazeera

Russia's recent missile and drone strikes on Ukraine have highlighted three war developments. Firstly, Iran appears to be playing a significant role in arming Russia. And finally, this new level of intensity means Ukraine is going to need help from the West if this new influx of missiles is to be stopped. In recent weeks, despite Iranian and Russian denials, images of the Shahed-136 drones, their distinctive delta wings silhouetted against the sky, have circulated around the global media. In Kyiv, tower block residents watched in horror as the drones flew below their windows, the lawnmower-like whine of their engines clearly heard, as they purposefully made their way to intended targets.