Intelligent Visual Recognition for Real-Time Object and Color Detection

Authors

  • BVN Prasad Paruchuri Department of Computer Science and Engineering, Koneru Lakshmaiah Education Foundation, Vaddeswaram, India Author
  • Degala Mohana ManiVenkatesh Department of Computer Science and Engineering, Koneru Lakshmaiah Education Foundation, Vaddeswaram, India Author
  • Srinivasa Rao Vankdoth Department of Computer Science and Engineering, Koneru Lakshmaiah Education Foundation, Vaddeswaram, India Author
  • Koppagiri Jyothsna Devi Department of Computer Science and Engineering, Koneru Lakshmaiah Education Foundation, Vaddeswaram, India Author
  • Pulipati Nagaraju Department of Computer Science and Engineering, Vishnu Institute of Technology, Bhimavaram, India Author
  • P. Nagabhushanam Department of Computer Science and Engineering, Sri Vasavi Engineering College, Tadepalligudem, India Author

DOI:

https://doi.org/10.68337/cpsm.v1.i1.2026-013

Keywords:

Object detection, color detection, convolutional neural network, YOLO, HSV, visual recognition, scene understanding

Abstract

Object and color detection is an important part of visual recognition and scene understanding, because recognizing an object together with its color provides the contextual cues needed to interpret real-world situations. Accurate detection is difficult, however, because of changes in lighting, complex backgrounds, and differences in object shape, size, and orientation. The proposed system is designed to address these problems by using deep learning to identify objects and their dominant colors in still images, video streams, and live camera feeds. It uses a pretrained You Only Look Once (YOLO) detector, YOLOv8, trained on a standard benchmark dataset such as Common Objects in Context (COCO). For color analysis, the system extracts the pixel values of the detected object regions and applies preprocessing (resizing, pixel value normalization, noise reduction, and conversion from the RGB to the hue, saturation, and value (HSV) color space) to reduce sensitivity to illumination changes. The system outputs object labels with confidence scores and displays the detections with text annotations and bounding boxes. An easy-to-use graphical user interface (GUI) supports single-image, batch-image, and real-time detection. The system was demonstrated qualitatively on 15 test images; quantitative evaluation metrics were not computed.

References

[1] Joseph Redmon, Santosh Divvala, Ross Girshick, and Ali Farhadi, "You only look once: Unified, real-time object detection," in Proc. IEEE Conf. Comput. Vis. Pattern Recognit. (CVPR), 2016, pp. 779-788, doi: 10.1109/CVPR.2016.91.

[2] Wei Liu, Dragomir Anguelov, Dumitru Erhan, Christian Szegedy, Scott Reed, Cheng-Yang Fu, and Alexander C. Berg, "SSD: Single shot multibox detector," in Computer Vision, ECCV 2016, Lecture Notes in Computer Science, vol. 9905, 2016, pp. 21-37, doi: 10.1007/978-3-319-46448-0_2.

[3] H. D. Cheng, X. H. Jiang, Y. Sun, and Jingli Wang, "Color image segmentation: Advances and prospects," Pattern Recognit., vol. 34, no. 12, pp. 2259-2281, 2001, doi: 10.1016/S0031-3203(00)00149-7.

[4] Shaoqing Ren, Kaiming He, Ross Girshick, and Jian Sun, "Faster R-CNN: Towards real-time object detection with region proposal networks," IEEE Trans. Pattern Anal. Mach. Intell., vol. 39, no. 6, pp. 1137-1149, 2017, doi: 10.1109/TPAMI.2016.2577031.

[5] Srinivasa Rao Vankdoth and Michael Arock, "End-to-end deep learning pipeline for scalable, deployable object detection engine in the traffic system," Signal Image Video Process., vol. 18, no. 2, pp. 1589-1600, 2024, doi: 10.1007/s11760-023-02869-5.

Downloads

Published

2026-09-30 — Updated on 2026-10-01

Versions

Data Availability Statement

The data supporting the findings of this study are available from the corresponding author upon reasonable request.

How to Cite

Intelligent Visual Recognition for Real-Time Object and Color Detection. (2026). Conference Proceedings in Science and Management, 1(1), 53-56. https://doi.org/10.68337/cpsm.v1.i1.2026-013 (Original work published 2026)