Attention Is All You Need research paper by Google, 2017
Google · Jun 12, 2017 · Architectures · 193,621 citations · 142 upvotes
What it shows
Introduced the Transformer, the attention-only architecture behind nearly every large language model since.
Summarised by hand from the abstract.
More from Google
All 4Other architectures papers
TopicAbout this paper
- Authors
- Ashish Vaswani, Noam Shazeer, Niki Parmar and 5 more
- arXiv
- 1706.03762 · PDF
- Venue
- Neural Information Processing Systems
- Citations
- 193,621, 20,774 influential · Semantic Scholar
- Upvotes
- 142 · Hugging Face
- Lab
- Google · on Companies · on Acquisitions · on Paydays · on TechConf · on Releases
Changes
| What changed | |
|---|---|
| Sep 24, 2026 | New paperAdded to the listSep 24, 2026 |
Sources: each lab's own papers and arXiv, with citation and upvote counts from Semantic Scholar and Hugging Face. One-line summaries are for orientation, not a substitute for the paper. Logos via logo.dev; trademarks belong to their owners.