Learning Representations for Reasoning: Generalizing Across Diverse Structures

Oct-16-2024–arXiv.org Artificial Intelligence

Reasoning, the ability to logically draw conclusions from existing knowledge, is a hallmark of human. Together with perception, they constitute the two major themes of artificial intelligence. While deep learning has pushed the limit of perception beyond human-level performance, the progress in reasoning domains is way behind. One fundamental reason is that reasoning problems usually have flexible structures for both knowledge and queries, and many existing models only perform well on structures seen during training. Here we aim to push the boundary of reasoning models by devising algorithms that generalize across knowledge and query structures, as well as systems that accelerate development on structured data. This thesis consists of three parts. In Part I, we study models that can inductively generalize to unseen knowledge graphs with new entity and relation vocabularies. For new entities, we propose a framework that learns neural operators in a dynamic programming algorithm computing path representations. For relations, we construct a relation graph to capture the interactions between relations, thereby converting new relations into new entities. In Part II, we propose two solutions for generalizing across multi-step queries on knowledge graphs and text respectively. For knowledge graphs, we show that multi-step queries can be solved by multiple calls of graph neural networks and fuzzy logic operations. For text, we devise an algorithm to learn explicit knowledge as textual rules to improve large language models on multi-step queries. In Part III, we propose two systems to facilitate machine learning development on structured data. Our library treats structured data as first-class citizens and removes the barrier for developing algorithms on structured data. Our node embedding system solves the GPU memory bottleneck of embedding matrices and scales to graphs with billion nodes.

large language model, machine learning, natural language, (23 more...)

arXiv.org Artificial Intelligence

Oct-16-2024

arXiv.org PDF

Add feedback

Country:
- South America > Chile (0.04)
- North America
  - United States
    - New York (0.04)
    - Utah (0.04)
    - New Jersey (0.04)
    - Pennsylvania > Allegheny County
      - Pittsburgh (0.04)
    - New Mexico > Los Alamos County
      - Los Alamos (0.04)
    - California > Santa Clara County
      - Palo Alto (0.04)
  - Canada
    - Quebec > Montreal (0.04)
    - Ontario > Toronto (0.04)
- Europe
  - Italy (0.04)
  - Greece (0.04)
  - United Kingdom > England
    - Cambridgeshire > Cambridge (0.04)
  - Netherlands > North Holland
    - Amsterdam (0.04)
- Asia
  - Japan (0.14)
  - Middle East > Jordan (0.04)
  - China > Liaoning Province
    - Shenyang (0.04)
- Africa > Middle East
  - Algeria (0.04)

Genre:
- Overview (1.00)
- Workflow (0.92)
- Research Report
  - New Finding (1.00)
  - Experimental Study (0.67)

Industry:
- Leisure & Entertainment > Sports (1.00)
- Information Technology (0.92)
- Health & Medicine
  - Pharmaceuticals & Biotechnology (1.00)
  - Therapeutic Area
    - Infections and Infectious Diseases (0.67)
    - Immunology (0.67)

Technology:
- Information Technology > Artificial Intelligence
  - Cognitive Science > Problem Solving (1.00)
  - Representation & Reasoning
    - Search (1.00)
    - Rule-Based Reasoning (1.00)
    - Expert Systems (1.00)
    - Semantic Networks (0.92)
  - Natural Language
    - Large Language Model (1.00)
    - Chatbot (1.00)
    - Information Retrieval > Query Processing (0.92)
  - Machine Learning
    - Statistical Learning (1.00)
    - Neural Networks > Deep Learning (1.00)
    - Inductive Learning (0.92)

Duplicate Docs Excel Report

Title
None found

Similar Docs Excel Report more

Title	Similarity	Source
None found