Detalhes do Documento

Mitigating false negatives in imbalanced datasets: an ensemble approach

Autor(es): Cavique, Luís ; Vasconcelos, Marcelo

Data: 2025

Identificador Persistente: http://hdl.handle.net/10400.2/20357

Origem: Repositório Aberto da Universidade Aberta

Assunto(s): Imbalanced dataset; False negative rate; Ensemble algorithms; Fraud detection; Set covering problem


Descrição

Imbalanced datasets present a challenge in machine learning, especially in binary classification scenarios where one class significantly outweighs the other. This imbalance often leads to models favoring the majority class, resulting in inadequate predictions for the minority class, specifically in false negatives. In response to this issue, this work introduces the MinFNR ensemble algorithm, designed to minimize False Negative Rates (FNR) in imbalanced datasets. The new approach strategically combines data-level, algorithmic-level, and hybrid-level approaches to enhance overall predictive capabilities while minimizing computational resources using the Set Covering Problem (SCP) formulation. Through a comprehensive evaluation of diverse datasets, MinFNR consistently outperforms individual algorithms, showing its potential for applications where the cost of false negatives is substantial, such as fraud detection and medical diagnosis. This work also contributes to ongoing efforts to improve the reliability and effectiveness of machine learning algorithms in real imbalanced scenarios.

Tipo de Documento Artigo científico
Idioma Inglês
Contribuidor(es) Repositório Aberto
Licença CC
facebook logo  linkedin logo  twitter logo 
mendeley logo

Documentos Relacionados

Não existem documentos relacionados.