Application of different Model Algorithm in the Prediction of Transfer Fee of Soccer Players
DOI:
https://doi.org/10.61173/d6jv0m98Keywords:
Machine learning algorithms, Data Analysis, football player’s transfer feeAbstract
Soccer is popular worldwide, and the fees for transferring soccer players show how much a player is worth and give us an idea of how well a country’s soccer is developing and how a team is being managed. This study aims to investigate the application of different model algorithms in the prediction of the transfer fee of soccer players and find which model is the most accurate. The dataset used for this research is available on the Kaggle website, from the football player’s transfer fee prediction dataset. By analyzing the (players’ codes and names) football team, position, height, age, the appearance of a player, the number of goals, assists, yellow cards, second yellow cards, red cards, goals conceded, clean sheets, minutes played, days-injured, games-injured, award, current-value, highest-value and position-encoded. Machine learning is commonly used in diverse fields to solve difficult problems that cannot be readily solved in based on computer approaches [1]. This study compares the accuracy of different machine learning algorithms used for predictive analysis of soccer players’ transfer fees. For example: Linear regression, Random Forest, Decision tree, K-Neighbors, and Neural network. When finding the relationship between those 21 factors, could help players and teams to make valuable decisions and accurate prediction for the establishment of soccer player market.
References
[1] Maulud, Dastan, and Adnan M. Abdulazeez. “A review on linear regression comprehensive in machine learning.” Journal of Applied Science and Technology Trends 1.2 (2020): 140-147.
[2] Wang, Guoming. “Quantum algorithm for linear regression.” Physical review A 96.1 (2017): 012335.
[3] McHale, Ian G., and Benjamin Holmes. “Estimating transfer fees of professional footballers using advanced performance metrics and machine learning.” European Journal of Operational Research 306.1 (2023): 389-399.
[4] Bernerth, Jeremy B., and Herman Aguinis. “A critical review and best‐practice recommendations for control variable usage.” Personnel psychology 69.1 (2016): 229-283.
[5] Keller, James M., Michael R. Gray, and James A. Givens. “A fuzzy k-nearest neighbor algorithm.” IEEE transactions on systems, man, and cybernetics 4 (1985): 580-585.
[6] Batista, G. E. A. P. A., and Diego Furtado Silva. “How k-nearest neighbor parameters affect its performance.” Argentine symposium on artificial intelligence. 2009.
[7] Wilamowski, Bogdan M. “Neural network architectures and learning algorithms.” IEEE Industrial Electronics Magazine 3.4 (2009): 56-63.
[8] Charbuty, Bahzad, and Adnan Abdulazeez. “Classification based on decision tree algorithm for machine learning.” Journal of Applied Science and Technology Trends 2.01 (2021): 20-28.
[9] Schonlau, M., & Zou, R. Y. (2020). The random forest algorithm for statistical learning. The Stata Journal, 20(1), 3-29.
[10] Ali, Jehad, et al. “Random forests and decision trees.” International Journal of Computer Science Issues (IJCSI) 9.5 (2012): 272.
Downloads
Published
Issue
Section
License
Copyright (c) 2025 by the authors.

This work is licensed under a Creative Commons Attribution 4.0 International License.
