Who Cited It

Attention Is All You Need

2025 · 7,133 citations · 12 from inside this corpus

Ashish Vaswani, Noam Shazeer, Niki Parmar, Jakob Uszkoreit, Llion Jones, Aidan N. Gomez, Łukasz Kaiser, Illia Polosukhin

The source holds an abstract for this work, but its best open-access copy is under no open licence, which does not permit us to republish the text. Read it at the source below.

Attention Is All You Need (2025)Attention Is All You Need[No title in the source record — DROPS (Schloss Dagstuhl – Leibniz Center for Informatics… (2015)[No title in the source recor…Effective Approaches to Attention-based Neural Machine Translation (2015)Effective Approaches to Atten…Building a Large Annotated Corpus of English: The Penn Treebank (1993)Building a Large Annotated Co…[No title in the source record — Edinburgh Research Explorer (University of Edinburgh)][No title in the source recor…Google's Neural Machine Translation System: Bridging the Gap between Human and Machine Tr… (2016)Google's Neural Machine Trans…Convolutional Sequence to Sequence Learning (2017)Convolutional Sequence to Seq…Explainable Artificial Intelligence (XAI): Concepts, taxonomies, opportunities and challe… (2019)Explainable Artificial Intell…Graph neural networks: A review of methods and applications (2020)Graph neural networks: A revi…DistilBERT, a distilled version of BERT: smaller, faster, cheaper and lighter (2019)DistilBERT, a distilled versi…A review of uncertainty quantification in deep learning: Techniques, applications and cha… (2021)A review of uncertainty quant…Relational inductive biases, deep learning, and graph networks (2018)Relational inductive biases, …BERTScore: Evaluating Text Generation with BERT (2019)BERTScore: Evaluating Text Ge…Shortcut learning in deep neural networks (2020)Shortcut learning in deep neu…Cross-lingual Language Model Pretraining (2019)Cross-lingual Language Model …Unsupervised Data Augmentation for Consistency Training (2019)Unsupervised Data Augmentatio…Pre-trained models for natural language processing: A survey (2020)Pre-trained models for natura…Heterogeneous Graph Transformer (2020)Heterogeneous Graph Transform…How to Fine-Tune BERT for Text Classification? (2019)How to Fine-Tune BERT for Tex…
18 of 18 neighbouring works in this corpus. Blue is what this paper cites; orange is what cites it, and a dashed line is one neighbour citing another. Only the largest labels are drawn — every node carries its full title on hover.
this paper works it cites works citing it node size = global citations · hover for the full title

What this paper cites, inside the corpus

What cites it, inside the corpus

Topics

Natural Language Processing TechniquesComputer Science
Topic ModelingComputer Science
Multimodal Machine Learning ApplicationsComputer Science

Is this record sound?

complete

Nothing in this record contradicts itself and no field we check is missing.

  • supports8 author record(s) attached.
  • supports28 reference(s) recorded.
  • neutralThe DOI carries no year to check against.
  • supportsA title is present.

Provenance

Everything above was read from one stored OpenAlex payload, fetched 2026-09-04T03:58:43+00:00.

sha256 5cad55ac5d4d41e1…