A Comparative Study of Machine Learning Algorithms for Predictive Data Analysis



EOI: 10.11242/viva-tech.01.09.26

Download Full Text here



Citation

Sonia Dubey, Satyam Chavan, Vignesh Asolkar,"A Comparative Study of Machine Learning Algorithms for Predictive Data Analysis" VIVA-IJRI Volume 1, Issue 9, Article 19, pp. 1-6, 2026. Published by MCA Department, VIVA Institute of Technology, Virar, India.

Abstract

Machine learning has become essential for predictive data analysis across diverse domains including healthcare, finance, and manufacturing. Despite the availability of numerous algorithms, practitioners face challenges in selecting appropriate methods for specific analytical tasks. This paper presents a qualitative analysis of machine learning algorithm selection criteria and proposes a structured decision framework based on data characteristics, computational constraints, and interpretability requirements. Through comprehensive literature review, we identify key factors influencing algorithm performance and provide practical guidance for algorithm selection without requiring extensive experimental comparison

Keywords

Machine learning, algorithm selection, predictive analytics, decision framework, data analysis

References

  1. [1] C. Fernรกndez-Delgado, E. Cernadas, S. Barro, and D. Amorim, "Do we need hundreds of classifiers to solve real world classification problems?" Journal of Machine Learning Research, vol. 15, pp. 3133-3181, 2014. ๐Ÿ”—https://jmlr.org/papers/v15/delgado14a.html
  2. [2] R. Caruana and A. Niculescu-Mizil, "An empirical comparison of supervised learning algorithms," Proceedings of the 23rd International Conference on Machine Learning, pp. 161-168, 2006. ๐Ÿ”—https://dl.acm.org/doi/10.1145/1143844.1143865
  3. [3] D. Wolpert and W. Macready, "No free lunch theorems for optimization," IEEE Transactions on Evolutionary Computation, vol. 1, no. 1, pp. 67-82, 1997. https://ieeexplore.ieee.org/document/585893
  4. [4] P. Domingos, "A few useful things to know about machine learning," Communications of the ACM, vol. 55, no. 10, pp. 78-87, 2012. https://dl.acm.org/doi/10.1145/2347736.2347755
  5. [5] D. Hand, "Classifier technology and the illusion of progress," Statistical Science, vol. 21, no. 1, pp. 1-14, 2006. ๐Ÿ”— https://projecteuclid.org/euclid.ss/1149600839
  6. [6] S. Kotsiantis, I. Zaharakis, and P. Pintelas, "Supervised machine learning: A review of classification techniques,"Emerging Artificial Intelligence Applications in Computer Engineering, vol. 160, pp. 3-24, 2007. ๐Ÿ”—https://www.researchgate.net/publication/228084860
  7. [7] R. Caruana, A. Niculescu-Mizil, G. Crew, and A. Ksikes, "Ensemble selection from libraries of models," Proceedings of the 21st International Conference on Machine Learning, 2004. ๐Ÿ”— https://dl.acm.org/doi/10.1145/1015330.1015432
  8. [8] R. Kohavi, "A study of cross-validation and bootstrap for accuracy estimation and model selection," International JointConference on Artificial Intelligence, vol. 14, pp. 1137-1145, 1995. ๐Ÿ”— https://www.ijcai.org/Proceedings/95-2/Papers/016.pdf
  9. [9] P. Brazdil, C. Giraud-Carrier, C. Soares, and R. Vilalta, "Metalearning: Applications to data mining," Springer Science & Business Media, 2009. ๐Ÿ”— https://link.springer.com/book/10.1007/978-3-540-73263-1
  10. [10] C. Lemke, M. Budka, and B. Gabrys, "Metalearning: A survey of trends and technologies," Artificial Intelligence Review, vol. 44, no. 1, pp. 117-130, 2015. ๐Ÿ”— https://link.springer.com/article/10.1007/s10462-013-9406-y
  11. [11] S. Ali and K. Smith-Miles, "A meta-learning approach to automatic kernel selection for support vector machines," Neurocomputing, vol. 70, no. 1-3, pp. 173-186, 2006. ๐Ÿ”— https://www.sciencedirect.com/science/article/abs/pii/S0925231206002773
  12. [12] T. Hastie, R. Tibshirani, and J. Friedman, "The Elements of Statistical Learning: Data Mining, Inference, and Prediction," 2nd ed., Springer, 2009. ๐Ÿ”— https://hastie.su.domains/ElemStatLearn/
  13. [13] G. James, D. Witten, T. Hastie, and R. Tibshirani, "An Introduction to Statistical Learning," Springer, 2013. ๐Ÿ”— https://www.statlearning.com/
  14. [14] C. Molnar, "Interpretable Machine Learning: A Guide for Making Black Box Models Explainable," 2020. ๐Ÿ”—https://christophm.github.io/interpretable-ml-book/
  15. [15] C. Rudin, "Stop explaining black box machine learning models for high stakes decisions and use interpretable models instead," Nature Machine Intelligence, vol. 1, no. 5, pp. 206-215, 2019. ๐Ÿ”— https://www.nature.com/articles/s42256-019-0048-x
  16. [16] M. Ribeiro, S. Singh, and C. Guestrin, "Why should I trust you? Explaining the predictions of any classifier,"Proceedings of the 22nd ACM SIGKDD International Conference on Knowledge Discovery and Data Mining, pp. 1135-1144, 2016. ๐Ÿ”— https://dl.acm.org/doi/10.1145/2939672.2939778
  17. [17] I. Goodfellow, Y. Bengio, and A. Courville, "Deep Learning," MIT Press, 2016. ๐Ÿ”—https://www.deeplearningbook.org/
  18. [18] Y. Bengio, A. Courville, and P.Vincent. Vincent, "Representation learning: A review and new perspectives," IEEE Transactions on Pattern Analysis and Machine Intelligence, vol. 35, no. 8, pp. 1798-1828, 2013. ๐Ÿ”—https://ieeexplore.ieee.org/document/6472238
  19. [19] C. Cortes and V. Vapnik, "Support-vector networks," Machine Learning, vol. 20, no. 3, pp. 273-297, 1995. ๐Ÿ”—https://link.springer.com/article/10.1007/BF00994018
  20. [20] L. Breiman, "Random forests," Machine Learning, vol. 45, no. 1, pp. 5-32, 2001. ๐Ÿ”— https://link.springer.com/article/10.1023/A:1010933404324