A Comparative Study of Machine Learning Algorithms for Predictive Data Analysis
EOI: 10.11242/viva-tech.01.09.26
Citation
Sonia Dubey, Satyam Chavan, Vignesh Asolkar,"A Comparative Study of Machine Learning Algorithms for Predictive Data Analysis" VIVA-IJRI Volume 1, Issue 9, Article 19, pp. 1-6, 2026. Published by MCA Department, VIVA Institute of Technology, Virar, India.
Abstract
Machine learning has become essential for predictive data analysis across diverse domains including healthcare, finance, and manufacturing. Despite the availability of numerous algorithms, practitioners face challenges in selecting appropriate methods for specific analytical tasks. This paper presents a qualitative analysis of machine learning algorithm selection criteria and proposes a structured decision framework based on data characteristics, computational constraints, and interpretability requirements. Through comprehensive literature review, we identify key factors influencing algorithm performance and provide practical guidance for algorithm selection without requiring extensive experimental comparison
Keywords
Machine learning, algorithm selection, predictive analytics, decision framework, data analysis
References
- [1] C. Fernรกndez-Delgado, E. Cernadas, S. Barro, and D. Amorim, "Do we need hundreds of classifiers to solve real world classification problems?" Journal of Machine Learning Research, vol. 15, pp. 3133-3181, 2014. ๐https://jmlr.org/papers/v15/delgado14a.html
- [2] R. Caruana and A. Niculescu-Mizil, "An empirical comparison of supervised learning algorithms," Proceedings of the 23rd International Conference on Machine Learning, pp. 161-168, 2006. ๐https://dl.acm.org/doi/10.1145/1143844.1143865
- [3] D. Wolpert and W. Macready, "No free lunch theorems for optimization," IEEE Transactions on Evolutionary Computation, vol. 1, no. 1, pp. 67-82, 1997. https://ieeexplore.ieee.org/document/585893
- [4] P. Domingos, "A few useful things to know about machine learning," Communications of the ACM, vol. 55, no. 10, pp. 78-87, 2012. https://dl.acm.org/doi/10.1145/2347736.2347755
- [5] D. Hand, "Classifier technology and the illusion of progress," Statistical Science, vol. 21, no. 1, pp. 1-14, 2006. ๐ https://projecteuclid.org/euclid.ss/1149600839
- [6] S. Kotsiantis, I. Zaharakis, and P. Pintelas, "Supervised machine learning: A review of classification techniques,"Emerging Artificial Intelligence Applications in Computer Engineering, vol. 160, pp. 3-24, 2007. ๐https://www.researchgate.net/publication/228084860
- [7] R. Caruana, A. Niculescu-Mizil, G. Crew, and A. Ksikes, "Ensemble selection from libraries of models," Proceedings of the 21st International Conference on Machine Learning, 2004. ๐ https://dl.acm.org/doi/10.1145/1015330.1015432
- [8] R. Kohavi, "A study of cross-validation and bootstrap for accuracy estimation and model selection," International JointConference on Artificial Intelligence, vol. 14, pp. 1137-1145, 1995. ๐ https://www.ijcai.org/Proceedings/95-2/Papers/016.pdf
- [9] P. Brazdil, C. Giraud-Carrier, C. Soares, and R. Vilalta, "Metalearning: Applications to data mining," Springer Science & Business Media, 2009. ๐ https://link.springer.com/book/10.1007/978-3-540-73263-1
- [10] C. Lemke, M. Budka, and B. Gabrys, "Metalearning: A survey of trends and technologies," Artificial Intelligence Review, vol. 44, no. 1, pp. 117-130, 2015. ๐ https://link.springer.com/article/10.1007/s10462-013-9406-y
- [11] S. Ali and K. Smith-Miles, "A meta-learning approach to automatic kernel selection for support vector machines," Neurocomputing, vol. 70, no. 1-3, pp. 173-186, 2006. ๐ https://www.sciencedirect.com/science/article/abs/pii/S0925231206002773
- [12] T. Hastie, R. Tibshirani, and J. Friedman, "The Elements of Statistical Learning: Data Mining, Inference, and Prediction," 2nd ed., Springer, 2009. ๐ https://hastie.su.domains/ElemStatLearn/
- [13] G. James, D. Witten, T. Hastie, and R. Tibshirani, "An Introduction to Statistical Learning," Springer, 2013. ๐ https://www.statlearning.com/
- [14] C. Molnar, "Interpretable Machine Learning: A Guide for Making Black Box Models Explainable," 2020. ๐https://christophm.github.io/interpretable-ml-book/
- [15] C. Rudin, "Stop explaining black box machine learning models for high stakes decisions and use interpretable models instead," Nature Machine Intelligence, vol. 1, no. 5, pp. 206-215, 2019. ๐ https://www.nature.com/articles/s42256-019-0048-x
- [16] M. Ribeiro, S. Singh, and C. Guestrin, "Why should I trust you? Explaining the predictions of any classifier,"Proceedings of the 22nd ACM SIGKDD International Conference on Knowledge Discovery and Data Mining, pp. 1135-1144, 2016. ๐ https://dl.acm.org/doi/10.1145/2939672.2939778
- [17] I. Goodfellow, Y. Bengio, and A. Courville, "Deep Learning," MIT Press, 2016. ๐https://www.deeplearningbook.org/
- [18] Y. Bengio, A. Courville, and P.Vincent. Vincent, "Representation learning: A review and new perspectives," IEEE Transactions on Pattern Analysis and Machine Intelligence, vol. 35, no. 8, pp. 1798-1828, 2013. ๐https://ieeexplore.ieee.org/document/6472238
- [19] C. Cortes and V. Vapnik, "Support-vector networks," Machine Learning, vol. 20, no. 3, pp. 273-297, 1995. ๐https://link.springer.com/article/10.1007/BF00994018
- [20] L. Breiman, "Random forests," Machine Learning, vol. 45, no. 1, pp. 5-32, 2001. ๐ https://link.springer.com/article/10.1023/A:1010933404324
