FACTORS THAT AFFECT IMAGE RECOGNITION EFFICIENCY WITH HUGGING FACE MODELS

Authors

  • Aleksandar Kotevski University Goce Delchev – Shtip, Faculty of Computer Sciences
  • Radmila Mitrevska University “St. Kliment Ohridski” – Bitola, Faculty of Technical Sciences

DOI:

https://doi.org/10.20544/hisj.2025.592

Keywords:

Image Recognition, Image Quality, Computational Hardware, GPU, CPU, Performance Analysis

Abstract

This study investigated the factors influencing the performance of image recognition systems, with particular attention to image quality, image type, and computational hardware. The objective was to identify optimal configurations for real-world applications by examining the trade-offs among accuracy, processing time, and computational efficiency. Pre-trained image classification models were evaluated on datasets comprising high-, low-, and mixed-resolution images, including natural landscapes, medical scans, and complex scenes. Each experimental condition was assessed in both CPU and GPU environments to measure recognition accuracy and processing time. The results indicated that high-resolution images generally improved recognition accuracy but also increased processing time and computational demands. Image complexity was found to affect both accuracy and inference time, particularly for deeper models. GPUs consistently outperformed CPUs, reducing inference time by more than 72 percent in some cases, especially when processing high-resolution or complex images. Nevertheless, CPUs provided acceptable performance for simpler tasks or in hardware-constrained environments. Overall, these findings provide practical guidance for selecting appropriate hardware and image configurations to optimize the performance of image recognition systems.

Downloads

Download data is not yet available.

References

1. Amerini, I., L. Galteri, R. Caldelli, and A. Del Bimbo. 2019. "Deepfake Video Detection Through Optical Flow Based CNN." In Proceedings of the IEEE/CVF International Conference on Computer Vision Workshops.

https://doi.org/10.1109/ICCVW.2019.00152

2. Armstrong, S., N. Bostrom, and C. Shulman. 2016. "Racing to the Precipice: A Model of Artificial Intelligence Development." AI & Society 31 (2): 201-206.

https://doi.org/10.1007/s00146-015-0590-y

3. Ayache, N. 2019. "AI & Healthcare: Towards a Digital Twin."

4. Chambon, P., C. Bluethgen, C. P. Langlotz, and A. Chaudhari. 2022. "Adapting Pretrained Vision-Language Foundational Models to Medical Imaging Domains." arXiv preprint arXiv:2210.04133.

5. Corvi, R., D. Cozzolino, G. Zingarini, G. Poggi, K. Nagano, and L. Verdoliva. 2022. "On the Detection of Synthetic Images Generated by Diffusion Models." arXiv preprint arXiv:2211.00680.

https://doi.org/10.1109/ICASSP49357.2023.10095167

PMid:33394960

6. Goyal, Y., A. Mohapatra, D. Parikh, and D. Batra. 2016. "Towards Transparent AI Systems: Interpreting Visual Question Answering Models." arXiv preprint arXiv:1608.08974. https://arxiv.org/pdf/1608.08974.

7. Guo, C., Y. Dou, T. Bai, X. Dai, C. Wang, and Y. Wen. 2023. "Artverse: A Paradigm for Parallel Human-Machine Collaborative Painting Creation in Metaverses." IEEE Transactions on Systems, Man, and Cybernetics: Systems.

https://doi.org/10.1109/TSMC.2022.3230406

8. Güera, D., and E. J. Delp. 2018. "Deepfake Video Detection Using Recurrent Neural Networks." In 2018 15th IEEE International Conference on Advanced Video and Signal Based Surveillance (AVSS), 1-6. IEEE.

https://doi.org/10.1109/AVSS.2018.8639163

9. Li, H., B. Li, S. Tan, and J. Huang. 2020. "Identification of Deep Network Generated Images Using Disparities in Color Components." Signal Processing 174: 107616.

https://doi.org/10.1016/j.sigpro.2020.107616

10. Osoba, O. A., and W. Welser IV. 2017. An Intelligence in Our Image: The Risks of Bias and Errors in Artificial Intelligence. Rand Corporation.

https://doi.org/10.7249/RR1744

PMCid:PMC5476816

11. Patil, Abhinav. 2021. "Image Recognition Using Machine Learning." SSRN. http://dx.doi.org/10.2139/ssrn.3835625.

https://doi.org/10.2139/ssrn.3835625

12. Russell, S. J., and P. Norvig. 2016. Artificial Intelligence: A Modern Approach. Malaysia: Pearson Education Limited.

13. Saikia, P., D. Dholaria, P. Yadav, V. Patel, and M. Roy. 2022. "A Hybrid CNN-LSTM Model for Video Deepfake Detection by Leveraging Optical Flow Features." In 2022 International Joint Conference on Neural Networks (IJCNN), 1-7. IEEE.

https://doi.org/10.1109/IJCNN55064.2022.9892905

14. Schneider, F. 2023. "Archisound: Audio Generation with Diffusion." Master's thesis, ETH Zurich.

15. Schneider, F., Z. Jin, and B. Schölkopf. 2023. "Moûsai: Text-to-Music Generation with Long-Context Latent Diffusion." arXiv preprint arXiv:2301.11757.

16. Sha, Z., Z. Li, N. Yu, and Y. Zhang. 2022. "De-Fake: Detection and Attribution of Fake Images Generated by Text-to-Image Diffusion Models." arXiv preprint arXiv:2210.06998.

https://doi.org/10.1145/3576915.3616588

17. Triguero, I., and C. Vens. 2016. "Labelling Strategies for Hierarchical Multi-Label Classification Techniques." Pattern Recognition 56: 170-183.

https://doi.org/10.1016/j.patcog.2016.02.017

18. Wang, J., Z. Wu, W. Ouyang, X. Han, J. Chen, Y.-G. Jiang, and S.-N. Li. 2022. "M2TR: Multi-Modal Multi-Scale Transformers for Deepfake Detection." In Proceedings of the 2022 International Conference on Multimedia Retrieval, 615-623.

https://doi.org/10.1145/3512527.3531415

19. Yi, D., C. Guo, and T. Bai. 2021. "Exploring Painting Synthesis with Diffusion Models." In 2021 IEEE 1st International Conference on Digital Twins and Parallel Intelligence (DTPI), 332-335. IEEE.

https://doi.org/10.1109/DTPI52967.2021.9540115

PMCid:PMC8935952

Downloads

Published

18-12-2025

How to Cite

Kotevski, Aleksandar, and Radmila Mitrevska. 2025. “FACTORS THAT AFFECT IMAGE RECOGNITION EFFICIENCY WITH HUGGING FACE MODELS”. Horizons - International Scientific Journal 2 (1): 98-114. https://doi.org/10.20544/hisj.2025.592.