[1406.1078] Learning Phrase Representations using RNN Encoder-Decoder for Statistical Machine Translation
arXivLabs is a framework that allows collaborators to develop and share new arXiv features directly on our website. Both individuals and organizations that work with arXivLabs have embraced and accepted our values of openness, community, excellence, and user data privacy. arXiv is committed to these values and only works with partners that adhere to them. Have an idea for a project that will add value for arXiv's community? Learn more about arXivLabs. arXiv Operational Status Get status notifications via email or slack
[1406.1078] Learning Phrase Representations using RNN Encoder-Decoder for Statistical Machine Translation Skip to main content arXiv is now an independent nonprofit! Learn more × Search arXiv Press Enter to search · Advanced search --> Computer Science > Computation and Language arXiv:1406.1078 (cs) [Submitted on 3 Jun 2014 ( v1 ), last revised 3 Sep 2014 (this version, v3)] Title: Learning Phrase Representations using RNN Encoder-Decoder for Statistical Machine Translation Authors: Kyunghyun Cho , Bart van Merrienboer , Caglar Gulcehre , Dzmitry Bahdanau , Fethi Bougares , Holger
Explore this link on the map →related reading
- [1609.08144] Google’s Neural Machine Translation System: Bridging the Gap between Human and Machine Translationarxiv-vanity.com
- Visualizing A Neural Machine Translation Model (Mechanics of Seq2seq Models With Attention) – Jay Alammar – Visualizing machine learning one concept at a time.jalammar.github.io
- [1706.03762] Attention Is All You Needarxiv.org
- Aman's AI Journal • Primers • Ilya Sutskever's Top 30aman.ai
- Towards Monosemanticity: Decomposing Language Models With Dictionary Learningtransformer-circuits.pub
- [1409.3215] Sequence to Sequence Learning with Neural Networksarxiv.org
- transformer_attention.pdfarxiv.org
- Encoder-Decoder Seq2Seq Models, Clearly Explained!! | by Kriz Moses | Analytics Vidhya | Mediummedium.com
- The Illustrated Transformer – Jay Alammar – Visualizing machine learning one concept at a time.jalammar.github.io
- The Unreasonable Effectiveness of Recurrent Neural Networkskarpathy.github.io
- 1706.03762arxiv.org
- Transformer: A Novel Neural Network Architecture for Language Understandingblog.research.google