Data Envelopment Learning: An Approach for Feature Selection

Vol 57, 2025 - 340380
Complete Articles (CA)
Favorite this paper
How to cite this paper?
Abstract

This study addresses feature selection in Data Envelopment Analysis (DEA) to mitigate the curse of dimensionality, which causes efficiency score degeneration. We first establish an Exhaustive Combinatorial Strategy as a benchmark to maximize efficiency variance; however, this approach is computationally prohibitive for high-dimensional datasets. To solve this, we propose Data Envelopment Learning (DEL), a hybrid framework integrating Autoencoders for unsupervised feature selection, TreeSHAP for interpretability, and CCR-DEA for evaluation. DEL identifies nonredundant, informative subsets while preserving original features. Experimental results show that DEL achieves near-equivalent performance to the exhaustive search, recovering 99.3% of efficiency variance while reducing computational time by over 104×. Furthermore, the method demonstrates strong ranking stability and superior generalization in cross-validation. These findings indicate that DEL offers a scalable, interpretable, and effective solution for DEA feature selection, enabling its application to complex scenarios that were previously computationally intractable.

Share your ideas or questions with the authors!

Did you know that the greatest stimulus in scientific and cultural development is curiosity? Leave your questions or suggestions to the author!

Sign in to interact

Have a question or suggestion? Share your feedback with the authors!

Institutions
  • 1 Universidade Estadual do Ceará - UECE
  • 2 Instituto Federal de Educação, Ciência e Tecnologia do Ceará (IFCE - campus Caucaia)
  • 3 Universidade Estadual do Ceará
Track
  • AD&GP – Operations Research in Production Management and Administration
Keywords
Data Envelopment Analysis.
Results Degeneration.
Feature Selection.