• español
    • English
  • Login
  • español 
    • español
    • English

UniversidaddeCádiz

Área de Biblioteca, Archivo y Publicaciones
Comunidades y colecciones
Ver ítem 
  •   RODIN Principal
  • Producción Científica
  • Artículos Científicos
  • Ver ítem
  •   RODIN Principal
  • Producción Científica
  • Artículos Científicos
  • Ver ítem
JavaScript is disabled for your browser. Some features of this site may not work without it.

REDIBAGG: Reducing the training set size in ensemble machine learning-based prediction models

Identificadores

URI: http://hdl.handle.net/10498/35835

DOI: 10.1016/j.engappai.2025.110382

URL: https://authors.elsevier.com/a/1klJL3OWJ9CRyl

ISSN: ISSN: 0952-1976

Ficheros
Artículo principal (1.008Mb)
Estadísticas
Ver estadísticas
Métricas y Citas
 
Compartir
Exportar a
Exportar a MendeleyRefworksEndNoteBibTexRIS
Metadatos
Mostrar el registro completo del ítem
Autor/es
Silva Ramírez, Esther LydiaAutoridad UCA; Cabrera Sánchez, Juan FranciscoAutoridad UCA; López Coello, ManuelAutoridad UCA
Fecha
2025-06-01
Departamento/s
Ingeniería Informática
Fuente
Engineering Applications of Artificial Intelligence Volume 149, 1 110382
Resumen
Machine learning-based algorithms have gained wide acceptance over the years due to their high generalization capabilities for a wide range of classification applications. Although these algorithms demonstrate potential and promising performance, they are often limited by speed, particularly when training on large databases. Big data sets with many instances have storage requirements and execution times that can be excessive. This study proposes reducing the sample size generated by the bootstrap resampling method in Ensemble Machine Learning-based models and evaluates their generalization capability on unknown data. Reduced bootstrap samples are employed in the training phase of the Bagging ensemble model. This approach reduces execution times and, consequently, storage requirements. The proposed method was tested on classification tasks, effectively reducing training subset size without compromising performance. Experimental results demonstrate that this approach achieves execution times reductions of up to 70% for some data sets. This reduction has no impact on accuracy, whereas maintaining levels comparable to classical Bagging and its variants. On average, the training subset size was reduced by 25% compared to the original size.
Materias
Artificial intelligence; Machine Learning; Big data; Bagging ensemble; Bootstrap; Reduced sample
Colecciones
  • Artículos Científicos [11777]
Attribution-NonCommercial-NoDerivatives 4.0 Internacional
Esta obra está bajo una Licencia Creative Commons Attribution-NonCommercial-NoDerivatives 4.0 Internacional

Listar

Todo RODINComunidades y ColeccionesPor fecha de publicaciónAutoresTítulosMateriasEsta colecciónPor fecha de publicaciónAutoresTítulosMaterias

Mi cuenta

AccederRegistro

Estadísticas

Ver Estadísticas de uso

Información adicional

Acerca de...Deposita en RODINPolíticasNormativasDerechos de autorEnlaces de interésEstadísticasNovedadesPreguntas frecuentes

RODIN está accesible a través de

OpenAIREOAIsterRecolectaHispanaEuropeanaBaseDARTOATDGoogle Académico

Enlaces de interés

Sherpa/RomeoDulcineaROAROpenDOARCreative CommonsORCID

RODIN está gestionado por el Área de Biblioteca, Archivo y Publicaciones de la Universidad de Cádiz

ContactoSugerenciasAtención al Usuario