Goto

Collaborating Authors

 Education


How to prepare students for the rise of artificial intelligence in the workforce

#artificialintelligence

The future impacts of artificial intelligence (AI) on society and the labour force have been studied and reported extensively. In a recent book, AI Superpowers, Kai-Fu Lee, former president of Google China, wrote that 40 to 50 per cent of current jobs will be technically and economically viable with AI and automation over the next 15 years. Artificial intelligence refers to computer systems that collect, interpret and learn from external data to achieve specific goals and tasks. Unlike natural intelligence displayed by humans and animals, it is an artificial form of intelligence demonstrated by machines. This has raised questions about the ethics of AI decision-making and impacts of AI in the workplace.


Collaborative Interactive Learning -- A clarification of terms and a differentiation from other research fields

arXiv.org Artificial Intelligence

The field of collaborative interactive learning (CIL) aims at developing and investigating the technological foundations for a new generation of smart systems that support humans in their everyday life. While the concept of CIL has already been carved out in detail (including the fields of dedicated CIL and opportunistic CIL) and many research objectives have been stated, there is still the need to clarify some terms such as information, knowledge, and experience in the context of CIL and to differentiate CIL from recent and ongoing research in related fields such as active learning, collaborative learning, and others. Both aspects are addressed in this paper.


Cluster, Classify, Regress: A General Method For Learning Discountinous Functions

arXiv.org Machine Learning

This paper presents a method for solving the supervised learning problem in which the output is highly nonlinear and discontinuous. It is proposed to solve this problem in three stages: (i) cluster the pairs of input-output data points, resulting in a label for each point; (ii) classify the data, where the corresponding label is the output; and finally (iii) perform one separate regression for each class, where the training data corresponds to the subset of the original input-output pairs which have that label according to the classifier. It has not yet been proposed to combine these 3 fundamental building blocks of machine learning in this simple and powerful fashion. This can be viewed as a form of deep learning, where any of the intermediate layers can itself be deep. The utility and robustness of the methodology is illustrated on some toy problems, including one example problem arising from simulation of plasma fusion in a tokamak.


Sampling Clustering

arXiv.org Artificial Intelligence

We propose an efficient linear-time graph-based divisive cluster analysis approach called Sampling Clustering. It constructs a lite informative dendrogram by recursively dividing a graph into subgraphs. In each recursive call, a graph is sampled first with a set of vertices being removed to disconnect latent clusters, then condensed by adding edges to the remaining vertices to avoid graph fragmentation caused by vertex removals. We also present some sampling and condensing methods and discuss the effectiveness in this paper. Our implementations run in linear time and achieve outstanding performance on various types of datasets. Experimental results show that they outperform state-of-the-art clustering algorithms with significantly less computing resource requirements.


SFI CRT-AI: Science Foundation Ireland Centre for Research Training in Artificial Intelligence

#artificialintelligence

The typical student in our CRT will follow a structured PhD training programming comprising research methods training, supervisor-initiated research-specific training, a CRT-organised training module in artificial intelligence methods, work placements, and international scientific research laboratories. In terms of specific CRT-organised training in artificial intelligence, we will focus on six thematic areas in which the co-applicant team and supervisor group have particular expertise. In our choice of areas, we are making no attempt at comprehensive coverage of the field of AI. Instead, we are selecting strands that (a) will relate to likely student PhD topics, and (b) afford scope for the students to acquire a broad range of relevant skills. A typical student in our CRT will follow a structured PhD training programme that comprises four main elements: (i) Host-based research methods training; (ii) Supervisor-initiated research-specific training; (iii) CRT-organized training in Artificial Intelligence methods; and (iv) Work placements.


The Best Public Datasets for Machine Learning and Data Science

#artificialintelligence

Google Dataset Search: Similar to how Google Scholar works, Dataset Search lets you find datasets wherever they're hosted, whether it's a publisher's site, a digital library, or an author's personal web page. Kaggle: A data science site that contains a variety of externally contributed to interesting datasets. You can find all kinds of niche datasets in its master list, from ramen ratings to basketball data to and even Seattle pet licenses. Although the data sets are user-contributed and thus have varying levels of cleanliness, the vast majority are clean. VisualData: Discover computer vision datasets by category, it allows searchable queries.


Towards Predicting Difficulty of Reading Comprehension Questions

AAAI Conferences

We present a corpus and approach to deduce the difficulty of questions asked in a reading comprehension test. A feature-driven model is designed that associates each question with a difficulty level. This would eliminate the laborious task of manually annotating questions in a computerized testing environment. Experiments performed on our corpus show that our model can classify questions with a micro F-score of 0.68.


Ignorance-Aware Approaches and Algorithms for Prototype Selection in Machine Learning

arXiv.org Machine Learning

Operating with ignorance is an important concern of the Machine Learning research, especially when the objective is to discover knowledge from the imperfect data. Data mining (driven by appropriate knowledge discovery tools) is about processing available (observed, known and understood) samples of data aiming to build a model (e.g., a classifier) to handle data samples, which are not yet observed, known or understood. These tools traditionally take samples of the available data (known facts) as an input for learning. We want to challenge the indispensability of this approach and we suggest considering the things the other way around. What if the task would be as follows: how to learn a model based on our ignorance, i.e. by processing the shape of 'voids' within the available data space? Can we improve traditional classification by modeling also the ignorance? In this paper, we provide some algorithms for the discovery and visualizing of the ignorance zones in two-dimensional data spaces and design two ignorance-aware smart prototype selection techniques (incremental and adversarial) to improve the performance of the nearest neighbor classifiers. We present experiments with artificial and real datasets to test the concept of the usefulness of ignorance discovery in machine learning.


FLAIRS-32 Poster Abstracts

AAAI Conferences

The FLAIRS poster track is designed to promote discussion of emerging ideas and work in order to encourage and help guide researchers — especially new researchers — who are able to present a full poster in the conference poster session and receive that critical work-shaping feedback that helps guide good work into great work. Abstracts of those posters appear here, which we hope to see fully developed into future FLAIRS papers..


A Conversational Intelligent Agent for Career Guidance and Counseling

AAAI Conferences

Navigating a career constitutes one of life’s most enduring challenges, particularly within a unique organization like the US Navy. While the Navy has numerous resources for guidance, accessing and identifying key information sources across the many existing platforms can be challenging for sailors (e.g., determining the appropriate program or point of contact, developing an accurate understanding of the process, and even recognizing the need for planning itself). Focusing on intermediate goals, evaluations, education, certifications, and training is quite demanding, even before considering their cumulative long-term implications. These are on top of generic personal issues, such as financial difficulties and homesickness when at sea for prolonged periods. We present the preliminary construction of a conversational intelligent agent designed to provide a user-friendly, adaptive environment that recognizes user input pertinent to these issues and provides guidance to appropriate resources within the Navy. User input from “counseling sessions” is linked, using advanced natural language processing techniques, to our framework of Navy training and education standards, promotion protocols, and organizational structure, producing feedback on resources and recommendations sensitive to user history and stated career goals. The proposed innovative technology monitors sailors’ career progress, proactively triggering sessions before major career milestones or when performance drops below Navy expectations, by using a mixed-initiative design. System-triggered sessions involve positive feedback and informative dialogues (using existing Navy career guidance protocols). The intelligent agent also offers counseling for personal problems, triggering targeted dialogues designed to gather more information, offer tailored suggestions, and provide referrals to appropriate resources or to a human counselor when in-depth counseling is warranted. This software, currently in alpha testing, has the potential to serve as a centralized information hub, engaging and encouraging sailors to take ownership of their career paths in the most efficient way possible, benefiting both individuals and the Navy as a whole.