Who Cited It

The citation and collaboration graph of open research, drawn and linked to its sources.

The most-cited works in OpenAlex's Artificial Intelligence subfield, with every author and every citation edge that falls inside the set. Every page is precomputed and every graph is drawn here rather than in your browser, so there is no login, no per-month graph allowance and nothing to wait for.

3,000papers
6,906authors
13,915citation edges
61,645collaboration edges
4,521institutions

What this corpus is, exactly

The 3,000 most-cited works in OpenAlex's Artificial Intelligence subfield, every author on them, and every citation between two works that are both inside the set. That bound is committed to the repository as corpus.json, so “why is this paper here and not that one” has a diffable answer. 3,000 is where this starts, not the shape of the thing.

34% of author records here are marked low confidence, and 101 of 3,000 paper records contradict themselves. Author identity in every open bibliographic database is produced by an algorithm that splits one researcher across several records and merges several researchers into one. We do not fix that silently. Each author page states how confident it is and shows you the signals, and nothing is ever merged away. How that is judged.

Works by publication year (93 years)

1873: 1 work(s)1922: 1 work(s)1927: 1 work(s)1928: 1 work(s)1929: 1 work(s)1931: 1 work(s)1932: 1 work(s)1936: 2 work(s)1937: 1 work(s)1940: 1 work(s)1944: 1 work(s)1945: 1 work(s)1946: 2 work(s)1947: 2 work(s)1948: 3 work(s)1949: 1 work(s)1950: 3 work(s)1951: 4 work(s)1952: 6 work(s)1953: 1 work(s)1954: 1 work(s)1955: 3 work(s)1956: 6 work(s)1957: 1 work(s)1958: 6 work(s)1959: 6 work(s)1960: 8 work(s)1961: 3 work(s)1962: 5 work(s)1963: 7 work(s)1964: 8 work(s)1965: 8 work(s)1966: 7 work(s)1967: 13 work(s)1968: 10 work(s)1969: 14 work(s)1970: 6 work(s)1971: 15 work(s)1972: 7 work(s)1973: 12 work(s)1974: 11 work(s)1975: 14 work(s)1976: 14 work(s)1977: 14 work(s)1978: 19 work(s)1979: 11 work(s)1980: 17 work(s)1981: 14 work(s)1982: 26 work(s)1983: 26 work(s)1984: 22 work(s)1985: 29 work(s)1986: 37 work(s)1987: 36 work(s)1988: 49 work(s)1989: 50 work(s)1990: 41 work(s)1991: 45 work(s)1992: 57 work(s)1993: 55 work(s)1994: 65 work(s)1995: 79 work(s)1996: 76 work(s)1997: 77 work(s)1998: 81 work(s)1999: 79 work(s)2000: 74 work(s)2001: 79 work(s)2002: 100 work(s)2003: 71 work(s)2004: 83 work(s)2005: 64 work(s)2006: 75 work(s)2007: 78 work(s)2008: 56 work(s)2009: 81 work(s)2010: 54 work(s)2011: 62 work(s)2012: 57 work(s)2013: 72 work(s)2014: 89 work(s)2015: 69 work(s)2016: 90 work(s)2017: 115 work(s)2018: 107 work(s)2019: 111 work(s)2020: 75 work(s)2021: 50 work(s)2022: 19 work(s)2023: 18 work(s)2024: 10 work(s)2025: 7 work(s)2026: 2 work(s) 187320171152026

Is each author record one person?

high: 3,643 (53%)medium: 890 (13%)low: 2,373 (34%)
high 3,643 53%medium 890 13%low 2,373 34%

Does each paper record agree with itself?

complete: 2,089 (70%)partial: 810 (27%)suspect: 101 (3%)
complete 2,089 70%partial 810 27%suspect 101 3%

Most cited

PaperYearCitedIn corpus
Random Forests
Leo Breiman
2001131,10945
Long Short-Term Memory
Sepp Hochreiter, Jürgen Schmidhuber
1997101,35989
Fitting Linear Mixed-Effects Models Using lme4
Douglas M. Bates, Martin Mächler, Benjamin M. Bolker, Steve Walker
201588,2960
Adam: A Method for Stochastic Optimization
Diederik P. Kingma, Jimmy Ba
201484,69856
Exploiting Generative AI to Scale up Intelligent Tutoring Systems
202379,07129
Fuzzy sets
L. A. Zadeh
196567,09321
Scikit-learn: Machine Learning in Python
Fabián Pedregosa, Gaël Varoquaux, Alexandre Gramfort, Vincent Michel and 12 more
201264,02523
XGBoost
201652,30410
[No title in the source record — DROPS (Schloss Dagstuhl – Leibniz Center for Informatics)]
201550,39652
Genetic algorithms in search, optimization, and machine learning
198949,34456
Particle swarm optimization
James Kennedy, R.C. Eberhart
200248,44749
Fuzzy sets
Lotfi A. Zadeh
199647,1476
AI-Assisted Pipeline for Dynamic Generation of Trustworthy Health Supplement Content at Scale
201846,03627
Model Selection and Multimodel Inference: A Practical Information-Theoretic Approach
Fred S. Guthery, Kenneth P. Burnham, David Anderson
200342,1872
Support-vector networks
Corinna Cortes, Vladimir Vapnik
199541,02821
Matplotlib: A 2D Graphics Environment
John D. Hunter
200740,82211
SciPy 1.0: fundamental algorithms for scientific computing in Python
202039,6422
The Nature of Statistical Learning Theory
Vladimir Vapnik
199539,40959
Adaptation in Natural and Artificial Systems
John H. Holland
199235,56967
Biostatistical Analysis
William H. Baltosser, Jerrold H. Zar
199635,4623
Dropout: a simple way to prevent neural networks from overfitting
Nitish Srivastava, Geoffrey E. Hinton, Alex Krizhevsky, Ilya Sutskever and 1 more
201434,23657
Glove: Global Vectors for Word Representation
Jeffrey Pennington, Richard Socher, Christopher D. Manning
201434,06740
ggplot2
Hadley Wickham
201633,9622
Support-Vector Networks
Corinna Cortes, Vladimir Vapnik
199533,91436
BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding
Jacob Devlin, Ming‐Wei Chang, Kenton Lee, Kristina Toutanova
201933,41639
SMOTE: Synthetic Minority Over-sampling Technique
Nitesh V. Chawla, Kevin W. Bowyer, Lawrence Hall, W. Philip Kegelmeyer
200232,40225
Learning representations by back-propagating errors
David E. Rumelhart, Geoffrey E. Hinton, Ronald J. Williams
198631,68860
A New Approach to Linear Filtering and Prediction Problems
R. E. Kalman
196031,48811
The Genome Analysis Toolkit: A MapReduce framework for analyzing next-generation DNA sequencing data
Aaron McKenna, Matthew G. Hanna, Eric Banks, Andrey Sivachenko and 7 more
201030,5270
Greedy function approximation: A gradient boosting machine.
Jerome H. Friedman
200130,13719
Neural Networks: A Comprehensive Foundation
Simon Haykin
199829,84014
Reinforcement Learning: An Introduction
Richard S. Sutton, Andy Barto
199827,40913
Latent dirichlet allocation
David M. Blei, Andrew Y. Ng, Michael I. Jordan
200327,07831
Statistical Learning Theory
Yuhai Wu, Vladimir Vapnik
199926,95744
Random sample consensus
Martin A. Fischler, Robert C. Bolles
198125,8701
Reinforcement Learning: An Introduction
200525,75827
Learning Phrase Representations using RNN Encoder–Decoder for Statistical Machine Translation
Kyunghyun Cho, Bart van Merriënboer, Çağlar Gülçehre, Dzmitry Bahdanau and 3 more
201425,04827
Batch Normalization: Accelerating Deep Network Training by Reducing Internal Covariate Shift
Sergey Ioffe, Christian Szegedy
201524,40432
C4.5: Programs for Machine Learning
J. R. Quinlan
199223,70448
A Survey on Transfer Learning
Sinno Jialin Pan, Qiang Yang
200923,67028

All 3,000 papers · All 6,906 authors