Case-Based Reasoning
Four different ways to solve a data science problem - case study
Based on the theory of stochastic processes (Poisson processes) and the Erlang distribution, the estimated number of postings per time unit is indeed 2x the time since last posting. The theory will also give you the variance for this estimator (infinite) and will tell you that it's much more robust to use time to 2nd or 3rd or 4th previous posting, which have finite and known variances. Now if the group is inactive, the time to previous posting itself can be infinite, but in practice this is not an issue. Note that the Poisson assumption would be violated in this case. The theory will also suggest how to combine time to 2nd, time to 3rd and time to 4th previous posting to get a better estimator, read my paper Estimation of the Intensity of a Poisson process by means of neares... for details.
Data science versus statistics, to solve problems: case study
In this article, I compare two approaches (with their advantages and drawbacks) to compute a simple metric: the number of unique visitors ("uniques") per year for a website. I use the word user or visitor interchangeably. The problem seems straightforward at first glance, but it is not. It is a complex big data problem because the naive approach involves sorting hundreds of billions of observations - called transactions or page views here. It is also complicated because there's no 100% sure way to identify and track a user over long time periods: cookies and IP addresses / browser combinations both have drawbacks.
The RatioLog Project: Rational Extensions of Logical Reasoning
Furbach, Ulrich, Schon, Claudia, Stolzenburg, Frieder, Weis, Karl-Heinz, Wirth, Claus-Peter
Higher-level cognition includes logical reasoning and the ability of question answering with common sense. The RatioLog project addresses the problem of rational reasoning in deep question answering by methods from automated deduction and cognitive computing. In a first phase, we combine techniques from information retrieval and machine learning to find appropriate answer candidates from the huge amount of text in the German version of the free encyclopedia "Wikipedia". In a second phase, an automated theorem prover tries to verify the answer candidates on the basis of their logical representations. In a third phase - because the knowledge may be incomplete and inconsistent -, we consider extensions of logical reasoning to improve the results. In this context, we work toward the application of techniques from human reasoning: We employ defeasible reasoning to compare the answers w.r.t. specificity, deontic logic, normative reasoning, and model construction. Moreover, we use integrated case-based reasoning and machine learning techniques on the basis of the semantic structure of the questions and answer candidates to learn giving the right answers.
Report on the Twenty-Second International Conference on Case-Based Reasoning
Bridge, Derek (University College Cork) | Lamontagne, Luc (Université Laval) | Plaza, Enric (IIIA, Artificial Intelligence Research Institute CSIC, Spanish National Research Council)
In cooperation with the Association for the Advancement of Artificial Intelligence (AAAI), the Twenty-Second International Conference on Case-Based Reasoning (ICCBR), the premier international meeting on research and applications in case-based reasoning (CBR), was held from Monday September 29 to Wednesday October 1, 2014, in Cork, Ireland. ICCBR is the annual meeting of the CBR community and the leading conference on this topic. Started in 1993 as the European Conference on CBR and 1995 as ICCBR, the two conferences alternated biennially until their merger in 2010.
Report on the Twenty-Second International Conference on Case-Based Reasoning
Bridge, Derek (University College Cork) | Lamontagne, Luc (Université Laval) | Plaza, Enric (IIIA, Artificial Intelligence Research Institute CSIC, Spanish National Research Council)
ICCBR is the annual meeting of the CBR community and the leading conference on this topic. Started in 1993 as the European Conference on CBR and 1995 as ICCBR, the two conferences alternated biennially until their merger in 2010. The main conference track featured 19 research paper presentations, 16 posters, and two invited speakers. The papers and posters reflected the state of the art of case-based reasoning, dealing both with open problems at the core of casebased reasoning (especially in similarity assessment, case adaptation, and case-based maintenance), as well as trending applications of CBR. Minor, Goethe University, Germany, and Emmanuel The first invited speaker, Tony Veale from University Nauer, LORIA, France.
Trust-Guided Behavior Adaptation Using Case-Based Reasoning
Floyd, Michael (Knexus Research) | Drinkwater, Michael (Knexus Research) | Aha, David (Naval Research Laboratory)
We propose an approach that allows a robot to evaluate its trustworthiness and adapt its behavior accordingly. The The addition of a robot to a team can be difficult if trust estimate, which we refer to as an inverse trust estimate, the human teammates do not trust the robot. This differs from traditional computational trust metrics in that it can result in underutilization or disuse of the robot, measures how much trust other agents have in the robot rather even if the robot has skills or abilities that are necessary than how much trust the robot has in other agents. Since the to achieve team goals or reduce risk. To robot can only use observable information and not information help a robot integrate itself with a human team, we that is internal to the teammates' reasoning, the inverse present an agent algorithm that allows a robot to estimate trust estimate relies on evaluating the standard interactions its trustworthiness and adapt its behavior accordingly.
Narrative Hermeneutic Circle: Improving Character Role Identification from Natural Language Text via Feedback Loops
Valls-Vargas, Josep (Drexel University) | Zhu, Jichen (Drexel University) | Ontanon, Santiago (Drexel University)
While most natural language understanding systems rely on a pipeline-based architecture, certain human text interpretation methods are based on a cyclic process between the whole text and its parts: the hermeneutic circle. In the task of automatically identifying characters and their narrative roles, we propose a feedback-loop-based approach where the output of later modules of the pipeline is fed back to earlier ones. We analyze this approach using a corpus of 21 Russian folktales. Initial results show that feeding back high-level narrative information improves the performance of some NLP tasks.
Computational Invention of Cadences and Chord Progressions by Conceptual Chord-Blending
Eppe, Manfred (IIIA-CSIC, ICSI) | Confalonieri, Roberto (IIIA-CSIC) | MacLean, Ewen (University of Edinburgh) | Kaliakatsos, Maximos (Uniersity of Thessaloniki) | Cambouropoulos, Emilios (University of Thessaloniki) | Schorlemmer, Marco (IIIA-CSIC) | Codescu, Mihai (University of Magdeburg) | Kühnberger, Kai-Uwe (University of Osnabrück)
We present a computational framework for chord invention based on a cognitive-theoretic perspective on conceptual blending. The framework builds on algebraic specifications, and solves two musicological problems. It automatically finds transitions between chord progressions of different keys or idioms, and it substitutes chords in a chord progression by other chords of a similar function, as a means to create novel variations. The approach is demonstrated with several examples where jazz cadences are invented by blending chords in cadences from earlier idioms, and where novel chord progressions are generated by inventing transition chords.
Nonparametric Nearest Neighbor Random Process Clustering
Tschannen, Michael, Bölcskei, Helmut
We consider the problem of clustering noisy finite-length observations of stationary ergodic random processes according to their nonparametric generative models without prior knowledge of the model statistics and the number of generative models. Two algorithms, both using the L1-distance between estimated power spectral densities (PSDs) as a measure of dissimilarity, are analyzed. The first algorithm, termed nearest neighbor process clustering (NNPC), to the best of our knowledge, is new and relies on partitioning the nearest neighbor graph of the observations via spectral clustering. The second algorithm, simply referred to as k-means (KM), consists of a single k-means iteration with farthest point initialization and was considered before in the literature, albeit with a different measure of dissimilarity and with asymptotic performance results only. We show that both NNPC and KM succeed with high probability under noise and even when the generative process PSDs overlap significantly, all provided that the observation length is sufficiently large. Our results quantify the tradeoff between the overlap of the generative process PSDs, the noise variance, and the observation length. Finally, we present numerical performance results for synthetic and real data.
A Case-Based Reasoning Framework to Choose Trust Models for Different E-Marketplace Environments
A. Irissappane, Athirai, Zhang, Jie
The performance of trust models highly depend on the characteristics of the environments where they are applied. Thus, it becomes challenging to choose a suitable trust model for a given e-marketplace environment, especially when ground truth about the agent (buyer and seller) behavior is unknown (called unknown environment). We propose a case-based reasoning framework to choose suitable trust models for unknown environments, based on the intuition that if a trust model performs well in one environment, it will do so in another similar environment. Firstly, we build a case base with a number of simulated environments (with known ground truth) along with the trust models most suitable for each of them. Given an unknown environment, case-based retrieval algorithms retrieve the most similar case(s), and the trust model of the most similar case(s) is chosen as the most suitable model for the unknown environment. Evaluation results confirm the effectiveness of our framework in choosing suitable trust models for different e-marketplace environments.