Ontologies
Invited Talks
Aleven, Vincent (Carnegie Mellon University) | Freuder, Eugene C. (University College Cork) | Graesser, Arthur C. (The University of Memphis) | Pustejovsky, James (Brandeis University) | Wiebe, Jan (University of Pittsburgh)
Vincent Aleven Intelligent tutoring systems (ITS) are highly effective in supporting student learning, but are difficult to build. The Cognitive Tutor Authoring Tools (CTAT) project started over 6 years ago with the goals of making it easier for experienced programmers, and possible for non-programmers to create an ITS. CTAT supports tutor building through programming by demonstration, an approach that has been successful in a range of application areas, but that has been applied to only a very limited degree to ITS authoring. Using CTAT, an author creates a tutor by demonstrating correct and incorrect problem solving behaviors, rather than by writing code. The resulting tutors, called exampletracing tutors, evaluate student behavior by flexibly comparing it against the demonstrated problem-solving examples.
Interpretations of the Web of Data
The emerging Web of Data utilizes the web infrastructure to represent and interrelate data. The foundational standards of the Web of Data include the Uniform Resource Identifier (URI) and the Resource Description Framework (RDF). URIs are used to identify resources and RDF is used to relate resources. While RDF has been posited as a logic language designed specifically for knowledge representation and reasoning, it is more generally useful if it can conveniently support other models of computing. In order to realize the Web of Data as a general-purpose medium for storing and processing the world's data, it is necessary to separate RDF from its logic language legacy and frame it simply as a data model. Moreover, there is significant advantage in seeing the Semantic Web as a particular interpretation of the Web of Data that is focused specifically on knowledge representation and reasoning. By doing so, other interpretations of the Web of Data are exposed that realize RDF in different capacities and in support of different computing models.
Mining Meaning from Wikipedia
Medelyan, Olena, Milne, David, Legg, Catherine, Witten, Ian H.
Wikipedia is a goldmine of information; not just for its many readers, but also for the growing community of researchers who recognize it as a resource of exceptional scale and utility. It represents a vast investment of manual effort and judgment: a huge, constantly evolving tapestry of concepts and relations that is being applied to a host of tasks. This article provides a comprehensive description of this work. It focuses on research that extracts and makes use of the concepts, relations, facts and descriptions found in Wikipedia, and organizes the work into four broad categories: applying Wikipedia to natural language processing; using it to facilitate information retrieval and information extraction; and as a resource for ontology building. The article addresses how Wikipedia is being used as is, how it is being improved and adapted, and how it is being combined with other structures to create entirely new resources. We identify the research groups and individuals involved, and how their work has developed in the last few years. We provide a comprehensive list of the open-source software they have produced.
Semantic Social Network Analysis
Erรฉtรฉo, Guillaume, Gandon, Fabien, Corby, Olivier, Buffa, Michel
Since its birth, the web provided many ways of interacting between us [6], revealing huge social network structures [17], a phenomenon amplified by web 2.0 applications [11]. Researchers extracted social networks from emails, mailinglist archives, hyperlink structure of homepages, cooccurrence of names in documents and from the digital traces created by web 2.0 application usages [9]. Facebook, LinkedIn or Myspace provide huge amounts of structured network data. The emergence of the semantic web approaches led researchers to build models of such online interactions using ontologies like FOAF, SIOC or SCOT. This paper starts with a brief state of the art on these enhanced RDF-based representations. We will see that the graphs built using these ontologies have a great potential that is not fully exploited so far. Then, we present a new framework for applying SNA to RDF representations of social data. In particular, the use of graph models underlying RDF and SPARQL extensions enables us to extract efficiently and to parameterize the classic SNA features directly from these representations.
Safe Reasoning Over Ontologies
Grabarnik, Genady, Kershenbaum, Aaron
As ontologies proliferate and automatic reasoners become more powerful, the problem of protecting sensitive information becomes more serious. In particular, as facts can be inferred from other facts, it becomes increasingly likely that information included in an ontology, while not itself deemed sensitive, may be able to be used to infer other sensitive information. We first consider the problem of testing an ontology for safeness defined as its not being able to be used to derive any sensitive facts using a given collection of inference rules. We then consider the problem of optimizing an ontology based on the criterion of making as much useful information as possible available without revealing any sensitive facts.
Tagging multimedia stimuli with ontologies
Horvat, Marko, Popovic, Sinisa, Bogunovic, Nikola, Cosic, Kresimir
Successful management of emotional stimuli is a pivotal issue concerning Affective Computing (AC) and the related research. As a subfield of Artificial Intelligence, AC is concerned not only with the design of computer systems and the accompanying hardware that can recognize, interpret, and process human emotions, but also with the development of systems that can trigger human emotional response in an ordered and controlled manner. This requires the maximum attainable precision and efficiency in the extraction of data from emotionally annotated databases While these databases do use keywords or tags for description of the semantic content, they do not provide either the necessary flexibility or leverage needed to efficiently extract the pertinent emotional content. Therefore, to this extent we propose an introduction of ontologies as a new paradigm for description of emotionally annotated data. The ability to select and sequence data based on their semantic attributes is vital for any study involving metadata, semantics and ontological sorting like the Semantic Web or the Social Semantic Desktop, and the approach described in the paper facilitates reuse in these areas as well.
Embedding Data within Knowledge Spaces
Myers, James D., Futrelle, Joe, Gaynor, Jeff, Plutchak, Joel, Bajcsy, Peter, Kastner, Jason, Kotwani, Kailash, Lee, Jong Sung, Marini, Luigi, Kooper, Rob, McGrath, Robert E., McLaren, Terry, Rodriguez, Alejandro, Liu, Yong
The promise of e-Science will only be realized when data is discoverable, accessible, and comprehensible within distributed teams, across disciplines, and over the long-term - without reliance on out-of-band (non-digital) means. We have developed the open-source Tupelo semantic content management framework and are employing it to manage a wide range of e-Science entities (including data, documents, workflows, people, and projects) and a broad range of metadata (including provenance, social networks, geospatial relationships, temporal relations, and domain descriptions). Tupelo couples the use of global identifiers and resource description framework (RDF) statements with an aggregatable content repository model to provide a unified space for securely managing distributed heterogeneous content and relationships. The Tupelo framework includes an HTTPbased data/metadata management protocol, application programming interfaces, and user interface widgets which have been incorporated into NCSA's portal and workflow tools and is a key component in recent work creating dynamic digital observatories (digital watersheds) that combine observational and modeled information. Tupelo also supports specialized indexes and inference logic (computation) relevant to metadata including geospatial location and provenance. This additional capability creates a powerful knowledge space that can map between disciplinary conceptual models and between the storage and data organization choices made by different e-Science organizations.
Geospatial semantics: beyond ontologies, towards an enactive approach
Current approaches to semantics in the geospatial domain are mainly based on ontologies, but ontologies, since continue to build entirely on the symbolic methodology, suffers from the classical problems, e.g. the symbol grounding problem, affecting representational theories. We claim for an enactive approach to semantics, where meaning is considered to be an emergent feature arising context-dependently in action. Since representational theories are unable to deal with context, a new formalism is required toward a contextual theory of concepts. SCOP is considered a promising formalism in this sense and is briefly described.
On Introspection, Metacognitive Control and Augmented Data Mining Live Cycles
We discuss metacognitive modelling as an enhancement to cognitive modelling and computing. Metacognitive control mechanisms should enable AI systems to self-reflect, reason about their actions, and to adapt to new situations. In this respect, we propose implementation details of a knowledge taxonomy and an augmented data mining life cycle which supports a live integration of obtained models.
Edhibou: a Customizable Interface for Decision Support in a Semantic Portal
Badra, Fadi, D'Aquin, Mathieu, Lieber, Jean, Meilender, Thomas
The Semantic Web is becoming more and more a reality, as the required technologies have reached an appropriate level of maturity. However, at this stage, it is important to provide tools facilitating the use and deployment of these technologies by end-users. In this paper, we describe EdHibou, an automatically generated, ontology-based graphical user interface that integrates in a semantic portal. The particularity of EdHibou is that it makes use of OWL reasoning capabilities to provide intelligent features, such as decision support, upon the underlying ontology. We present an application of EdHibou to medical decision support based on a formalization of clinical guidelines in OWL and show how it can be customized thanks to an ontology of graphical components.