Who Cited It

SMOTE for Learning from Imbalanced Data: Progress and Challenges, Marking the 15-year Anniversary

2018 · Journal of Artificial Intelligence Research · 2,226 citations · 0 from inside this corpus

Alberto Fernández, Salvador García, Francisco Herrera, Nitesh V. Chawla

The Synthetic Minority Oversampling Technique (SMOTE) preprocessing algorithm is considered "de facto" standard in the framework of learning from imbalanced data. This is due to its simplicity in the design of the procedure, as well as its robustness when applied to different type of problems. Since its publication in 2002, SMOTE has proven successful in a variety of applications from several different domains. SMOTE has also inspired several approaches to counter the issue of class imbalance, and has also significantly contributed to new supervised learning paradigms, including multilabel classification, incremental learning, semi-supervised learning, multi-instance learning, among others. It is standard benchmark for learning from imbalanced data. It is also featured in a number of different software packages - from open source to commercial. In this paper, marking the fifteen year anniversary of SMOTE, we reflect on the SMOTE journey, discuss the current state of affairs with SMOTE, its applications, and also identify the next set of challenges to extend SMOTE for Big Data problems.

SMOTE for Learning from Imbalanced Data: Progress and Challenges, Marking the 15-year Ann… (2018)SMOTE for Learning from Imbal…SMOTE: Synthetic Minority Over-sampling Technique (2002)SMOTE: Synthetic Minority Ove…Learning from Imbalanced Data (2009)Learning from Imbalanced DataADASYN: Adaptive synthetic sampling approach for imbalanced learning (2008)ADASYN: Adaptive synthetic sa…A study of the behavior of several methods for balancing machine learning training data (2004)A study of the behavior of se…Borderline-SMOTE: A New Over-Sampling Method in Imbalanced Data Sets Learning (2005)Borderline-SMOTE: A New Over-…A Review on Ensembles for the Class Imbalance Problem: Bagging-, Boosting-, and Hybrid-Ba… (2011)A Review on Ensembles for the…Learning from imbalanced data: open challenges and future directions (2016)Learning from imbalanced data…Learning from class-imbalanced data: Review of methods and applications (2016)Learning from class-imbalance…Multiobjective evolutionary algorithms: A survey of the state of the art (2011)Multiobjective evolutionary a…Editorial (2004)EditorialIntroduction to Semi-Supervised Learning (2009)Introduction to Semi-Supervis…CLASSIFICATION OF IMBALANCED DATA: A REVIEW (2009)CLASSIFICATION OF IMBALANCED …SMOTEBoost: Improving Prediction of the Minority Class in Boosting (2003)SMOTEBoost: Improving Predict…An insight into classification with imbalanced data: Empirical results and current trends… (2013)An insight into classificatio…Imbalanced-learn: A Python Toolbox to Tackle the Curse of Imbalanced Datasets in Machine … (2016)Imbalanced-learn: A Python To…Imbalanced-learn: A Python Toolbox to Tackle the Curse of Imbalanced\n Datasets in Machin… (2016)Imbalanced-learn: A Python To…
16 of 16 neighbouring works in this corpus. Blue is what this paper cites; orange is what cites it, and a dashed line is one neighbour citing another. Only the largest labels are drawn — every node carries its full title on hover.
this paper works it cites works citing it node size = global citations · hover for the full title

What this paper cites, inside the corpus

Topics

Imbalanced Data Classification TechniquesComputer Science
Text and Document Classification TechnologiesComputer Science
Electricity Theft Detection TechniquesEngineering

Is this record sound?

complete

Nothing in this record contradicts itself and no field we check is missing.

  • supports4 author record(s) attached.
  • supports271 reference(s) recorded.
  • neutralThe DOI carries no year to check against.
  • supportsA title is present.

Provenance

Everything above was read from one stored OpenAlex payload, fetched 2026-09-04T03:58:50+00:00.

sha256 86a6d6715ff540a0…