Neural Probabilistic Model for Non-projective MST Parsing

Sep-3-2017–arXiv.org Machine Learning

In this paper, we propose a probabilistic parsing model that defines a proper conditional probability distribution over non-projective dependency trees for a given sentence, using neural representations as inputs. The neural network architecture is based on bidirectional LSTM-CNNs, which automatically benefits from both word-and character-level representations, by using a combination of bidirectional LSTMs and CNNs. On top of the neural network, we introduce a probabilistic structured layer, defining a conditional log-linear model over non-projective trees. By exploiting Kirchhoff's Matrix-Tree Theorem (Tutte, 1984), the partition functions and marginals can be computed efficiently, leading to a straightforward end-to-end model training procedure via back-propagation. We evaluate our model on 17 different datasets, across 14 different languages. Our parser achieves state-of-the-art parsing performance on nine datasets.

artificial intelligence, machine learning, proceedings, (18 more...)

arXiv.org Machine Learning

Sep-3-2017

arXiv.org PDF

Add feedback

Country:
- North America > United States
  - Texas (0.14)
  - Maryland (0.14)
  - Colorado (0.14)
- Europe > United Kingdom
  - Scotland (0.14)
- Asia > Middle East
  - Qatar (0.14)

Genre:
- Research Report (0.82)

Technology:
- Information Technology > Artificial Intelligence > Machine Learning > Neural Networks > Deep Learning (1.00)

Duplicate Docs Excel Report

Title
None found

Similar Docs Excel Report more

Title	Similarity	Source
None found