Structured Attention Networks
Explore this paper's citation graph
Summary
This work shows that structured attention networks are simple extensions of the basic attention procedure, and that they allow for extending attention beyond the standard soft-selection approach, such as attending to partial segmentations or to subtrees.
- Type
- article
- Published
- 2017-02-03
- Cited by
- 498
- References
- 66
- Access
- Open access
- OpenAlex
- https://openalex.org/W2586050494
- Semantic Scholar
- https://api.semanticscholar.org/CorpusID:6961760
Keywords
Computer science, Artificial intelligence, Inference, Machine translation, Machine learning
References
- Attention-Based Models for Speech Recognition
- Show, Attend and Tell: Neural Image Caption Generation with Visual Attention
- Teaching Machines to Read and Comprehend
- Learning to Transduce with Unbounded Memory
- Gradient Estimation Using Stochastic Computation Graphs
- End-To-End Memory Networks
- Deep Structured Output Learning for Unconstrained Text Recognition
- Listen, attend and spell: A neural network for large vocabulary conversational speech recognition
- Effective Approaches to Attention-based Neural Machine Translation
- Describing Multimedia Content Using Attention-Based Encoder-Decoder Networks
- Approximation-Aware Dependency Parsing by Belief Propagation
- Dependency Parsing by Belief Propagation
- Parameter learning with truncated message-passing
- Trainable grammars for speech recognition
- Three New Probabilistic Models for Dependency Parsing: An Exploration
- First- and Second-Order Expectation Semirings with Applications to Minimum-Risk Training on Translation Forests
- Gradient-based learning applied to document recognition
- Reasoning about Entailment with Neural Attention
- Conditional Neural Fields
- Pointwise Prediction for Robust, Adaptable Japanese Morphological Analysis
Cited by
- Neural Machine Translation and Sequence-to-sequence Models: A Tutorial
- Recurrent Models for Situation Recognition
- Pose-conditioned Spatio-Temporal Attention for Human Action Recognition
- Baselines and Test Data for Cross-Lingual Inference
- A Regularized Framework for Sparse and Structured Neural Attention
- Jointly learning sentence embeddings and syntax with unsupervised Tree-LSTMs
- Frame-Semantic Parsing with Softmax-Margin Segmental RNNs and a Syntactic Scaffold
- Character-level Intra Attention Network for Natural Language Inference
- Improved Neural Machine Translation with a Syntax-Aware Encoder and Decoder
- Sequential Attention: A Context-Aware Alignment Function for Machine Reading
- Human Action Recognition: Pose-Based Attention Draws Focus to Hands
- Towards Neural Machine Translation with Latent Tree Attention
- Do latent tree learning models identify meaningful structure in sentences?
- Learning to Compose Task-Specific Tree Structures
- Aspect-level Sentiment Classification with HEAT (HiErarchical ATtention) Network
- Weighted Transformer Network for Machine Translation
- Dialogue Act Recognition via CRF-Attentive Structured Network
- Attention networks for image-to-text
- Beyond saliency: understanding convolutional neural networks from saliency prediction on layer-wise relevance propagation
- Enhanced Image Classification With Data Augmentation Using Position Coordinates