Overview
More Than Reading Comprehension: A Survey on Datasets and Metrics of Textual Question Answering
Textual Question Answering (QA) aims to provide precise answers to user's questions in natural language using unstructured data. One of the most popular approaches to this goal is machine reading comprehension(MRC). In recent years, many novel datasets and evaluation metrics based on classical MRC tasks have been proposed for broader textual QA tasks. In this paper, we survey 47 recent textual QA benchmark datasets and propose a new taxonomy from an application point of view. In addition, We summarize 8 evaluation metrics of textual QA tasks. Finally, we discuss current trends in constructing textual QA benchmarks and suggest directions for future work.
Rethinking Crowd Sourcing for Semantic Similarity
Solomon, Shaul, Cohn, Adam, Rosenblum, Hernan, Hershkovitz, Chezi, Yamshchikov, Ivan P.
Estimation of semantic similarity is crucial for a variety of natural language processing (NLP) tasks. In the absence of a general theory of semantic information, many papers rely on human annotators as the source of ground truth for semantic similarity estimation. This paper investigates the ambiguities inherent in crowd-sourced semantic labeling. It shows that annotators that treat semantic similarity as a binary category (two sentences are either similar or not similar and there is no middle ground) play the most important role in the labeling. The paper offers heuristics to filter out unreliable annotators and stimulates further discussions on human perception of semantic similarity.
Named Entity Recognition and Classification on Historical Documents: A Survey
Ehrmann, Maud, Hamdi, Ahmed, Pontes, Elvys Linhares, Romanello, Matteo, Doucet, Antoine
After decades of massive digitisation, an unprecedented amount of historical documents is available in digital format, along with their machine-readable texts. While this represents a major step forward with respect to preservation and accessibility, it also opens up new opportunities in terms of content mining and the next fundamental challenge is to develop appropriate technologies to efficiently search, retrieve and explore information from this 'big data of the past'. Among semantic indexing opportunities, the recognition and classification of named entities are in great demand among humanities scholars. Yet, named entity recognition (NER) systems are heavily challenged with diverse, historical and noisy inputs. In this survey, we present the array of challenges posed by historical documents to NER, inventory existing resources, describe the main approaches deployed so far, and identify key priorities for future developments.
Deep Learning for Ultrasound Beamforming
van Sloun, Ruud JG, Ye, Jong Chul, Eldar, Yonina C
Diagnostic imaging plays a critical role in healthcare, serving as a fundamental asset for timely diagnosis, disease staging and management as well as for treatment choice, planning, guidance, and follow-up. Among the diagnostic imaging options, ultrasound imaging is uniquely positioned, being a highly cost-effective modality that offers the clinician an unmatched and invaluable level of interaction, enabled by its real-time nature. Ultrasound probes are becoming increasingly compact and portable, with the market demand for low-cost pocket-sized and (in-body) miniaturized devices expanding. At the same time, there is a strong trend towards 3D imaging and the use of high-frame-rate imaging schemes; both accompanied by dramatically increasing data rates that pose a heavy burden on the probe-system communication and subsequent image reconstruction algorithms. With the demand for high-quality image reconstruction and signal extraction from less (e.g unfocused or parallel) transmissions that facilitate fast imaging, and a push towards compact probes, modern ultrasound imaging leans heavily on innovations in powerful digital receive channel processing. Beamforming, the process of mapping received ultrasound echoes to the spatial image domain, naturally lies at the heart of the ultrasound image formation chain. In this chapter on Deep Learning for Ultrasound Beamforming, we discuss why and when deep learning methods can play a compelling role in the digital beamforming pipeline, and then show how these data-driven systems can be leveraged for improved ultrasound image reconstruction.
A survey of Bayesian Network structure learning
Kitson, Neville K., Constantinou, Anthony C., Guo, Zhigao, Liu, Yang, Chobtham, Kiattikun
Bayesian Networks (BNs) have become increasingly popular over the last few decades as a tool for reasoning under uncertainty in fields as diverse as medicine, biology, epidemiology, economics and the social sciences. This is especially true in real-world areas where we seek to answer complex questions based on hypothetical evidence to determine actions for intervention. However, determining the graphical structure of a BN remains a major challenge, especially when modelling a problem under causal assumptions. Solutions to this problem include the automated discovery of BN graphs from data, constructing them based on expert knowledge, or a combination of the two. This paper provides a comprehensive review of combinatoric algorithms proposed for learning BN structure from data, describing 61 algorithms including prototypical, well-established and state-of-the-art approaches. The basic approach of each algorithm is described in consistent terms, and the similarities and differences between them highlighted. Methods of evaluating algorithms and their comparative performance are discussed including the consistency of claims made in the literature. Approaches for dealing with data noise in real-world datasets and incorporating expert knowledge into the learning process are also covered.
A Survey on Cost Types, Interaction Schemes, and Annotator Performance Models in Selection Algorithms for Active Learning in Classification
Herde, Marek, Huseljic, Denis, Sick, Bernhard, Calma, Adrian
Pool-based active learning (AL) aims to optimize the annotation process (i.e., labeling) as the acquisition of annotations is often time-consuming and therefore expensive. For this purpose, an AL strategy queries annotations intelligently from annotators to train a high-performance classification model at a low annotation cost. Traditional AL strategies operate in an idealized framework. They assume a single, omniscient annotator who never gets tired and charges uniformly regardless of query difficulty. However, in real-world applications, we often face human annotators, e.g., crowd or in-house workers, who make annotation mistakes and can be reluctant to respond if tired or faced with complex queries. Recently, a wide range of novel AL strategies has been proposed to address these issues. They differ in at least one of the following three central aspects from traditional AL: (1) They explicitly consider (multiple) human annotators whose performances can be affected by various factors, such as missing expertise. (2) They generalize the interaction with human annotators by considering different query and annotation types, such as asking an annotator for feedback on an inferred classification rule. (3) They take more complex cost schemes regarding annotations and misclassifications into account. This survey provides an overview of these AL strategies and refers to them as real-world AL. Therefore, we introduce a general real-world AL strategy as part of a learning cycle and use its elements, e.g., the query and annotator selection algorithm, to categorize about 60 real-world AL strategies. Finally, we outline possible directions for future research in the field of AL.
A review of deep learning methods for MRI reconstruction
Following the success of deep learning in a wide range of applications, neural network-based machine-learning techniques have received significant interest for accelerating magnetic resonance imaging (MRI) acquisition and reconstruction strategies. A number of ideas inspired by deep learning techniques for computer vision and image processing have been successfully applied to nonlinear image reconstruction in the spirit of compressed sensing for accelerated MRI. Given the rapidly growing nature of the field, it is imperative to consolidate and summarize the large number of deep learning methods that have been reported in the literature, to obtain a better understanding of the field in general. This article provides an overview of the recent developments in neural-network based approaches that have been proposed specifically for improving parallel imaging. A general background and introduction to parallel MRI is also given from a classical view of k-space based reconstruction methods.
Artificial intelligence (AI) at the edge: 3 key facts
Artificial Intelligence (AI) is moving from the realm of science fiction to widespread enterprise scalability. Even ten years ago, AI workloads were almost exclusively utilized by a small number of very profitable companies that had the resources to experiment and hire an extensive team of data scientists. Today, AI is used in a number of everyday tools, from language recognition to health care prediction and nearly every industry in between. AI is also now deployed at the edge, not just inside massive data processing facilities. That trend will continue in the coming years.
Reports of the Association for the Advancement of Artificial Intelligence's 2020 Fall Symposium Series
The Association for the Advancement of Artificial Intelligence's 2020 Fall Symposium Series was held virtually from November 11-14, 2020, and was collocated with three symposia postponed from March 2020 due to the COVID-19 Pandemic. There were five symposia in the fall program: AI for Social Good, Artificial Intelligence in Government and Public Sector, Conceptual Abstraction and Analogy in Natural and Artificial Intelligence, Physics-Guided AI to Accelerate Scientific Discovery, and Trust and Explainability in Artificial Intelligence for Human-Robot Interaction. Additionally, there were three symposia delayed from spring: AI Welcomes Systems Engineering: Towards the Science of Interdependence for Autonomous Human-Machine Teams, Deep Models and Artificial Intelligence for Defense Applications: Potentials, Theories, Practices, Tools, and Risks, and Towards Responsible AI in Surveillance, Media, and Security through Licensing. Recent developments in big data and computational power are revolutionizing several domains, opening up new opportunities and challenges. In this symposium, we highlighted two specific themes, namely humanitarian relief, and healthcare, where AI could be used for social good to achieve the United Nations (UN) sustainable development goals (SDGs) in those areas, which touch every aspect of human, social, and economic development. The talks at the symposium were focused on identifying the critical needs and pathways for responsible AI solutions to achieve SDGs, which demand holistic thinking on optimizing the trade-off between automation benefits and their potential side-effects, especially in a year that has upended societies globally due to the COVID-19 pandemic. Riding on the success of the AI for Social Good symposium that was held in Washington, DC, in November 2019, we organized the 2020 version of the symposium.
89% of tech execs see synthetic data as a key to staying ahead
The Transform Technology Summits start October 13th with Low-Code/No Code: Enabling Enterprise Agility. Nearly nine in ten (89%) technology decision makers who use vision data agree synthetic data is a new and innovative technology and believe that organizations that fail to adopt synthetic data are at risk of falling behind the curve, according to new research by Synthesis AI in conjunction with Vanson Bourne. Technology leaders agree that synthetic data will be an essential enabling technology and key to staying ahead. Above: Synthetic Data could be a solution to the time consuming and cost prohibitive nature of supervised learning. AI is driven by the speed, diversity, and quality of data.