Using Different Convergence Behaviors Across Samples to Detect Minority Subgroups in Data
ISEF · 2026 Robotics and Intelligent Machines
Overview
As computer vision systems become increasingly integrated into everyday life, it is imperative to consider the deficiencies and biases present in our computer algorithms. One such bias is spurious correlation caused by subgroup underrepresentation in datasets, which is a major issue in machine learning. These biases occur when a spurious feature is predictive of a specific label during training and is learned instead of the core feature (feature predictive of a specific label in real-world applications). When trained using empirical risk minimization (ERM) on unbalanced datasets, convolutional neural networks (CNNs) struggle to perform well on bias-conflicting subgroups. In this paper, we propose a novel method for identifying spuriously correlated samples in image datasets. Specifically, we trained different CNN architectures on a variety of datasets and plotted each step of one training epoch to observe model convergence behavior, and potentially find evidence for overfitting to the bias-conflicting (minority) samples in the training set. To verify this, we evaluated the trained model on noisy versions of the datasets for model robustness evaluation. Throughout the process, we used subgroup labels to validate our results. We found that our approach successfully identifies spuriously correlated samples, as convergence behaviors and model performance differs across subgroups.
Competition history
- ISEF 2026
Resources
Related projects
ISEF · 2024
Improving the Fairness of Artificially Intelligent Skin Disease Detectors Using Stable Diffusion
ISEF · 2023
BrainTrain: Aligning Deep Neural Networks to Human Behavior to Improve Robustness and Generalization
ISEF · 2022
Neural Networks Learn Lazily: Improving Generalization and Adversarial Robustness via Learning Capacity-Complexity Constraints
JSHS · 2024
Increasing Machine Learning Training Data to Reduce Selection Bias in AI
ISEF · 2024
Advancing Bias Mitigation in Machine Learning Models: The Use of Feature-Wise Mixing Across Diverse Classifiers for Contextual Bias Mitigation
ISEF · 2023
Increasing Computer Vision Models Interpretability
JSHS · 2024
Fairness in Autonomous Driving: Towards Understanding Confounding Factors in Object Detection Under Challenging Weather
ISEF · 2026
DermEquity: A Universal Safety Framework for Medical AI Using Worst-Case Bias Certification and Influence-Based Attribution Validated in Skin Cancer Screening
Closest projects by meaning, across every fair and year in the corpus.
Browse more like this
Source: Regeneron International Science and Engineering Fair