AI Object Detection and Smart Response Humanoid Robot
T. Swarna Latha,
Ediga Indu,
Ganaparthi Tribhuvana,
Chintha Vijaya Varshitha,
Edambaku Sai Varshini and
Mohammad Rahath
International Journal of Scientific Research in Science and Technology, 2026, vol. 13, issue 2, 430-443
Abstract:
Vision-based object perception has become a fundamental capability for intelligent humanoid robots, enabling autonomous awareness, adaptive behavior, and natural interaction with dynamic environments. Conventional humanoid as demonstration of robots often rely on the predefined scripts, limited sensor-based on detection, or external computation, which restrict real-time responsiveness and reduce interaction realism. This paper presents the design, implementation, and evaluation of an AI-based object detection and the smart response humanoid robot for autonomous and the interactive operation using the embedded vision intelligence. The proposed system utilizes an onboard camera to acquire real-time visual data, which is processed using deep learning–based computer vision algorithms to detect, classify, and localize objects in the robot’s surroundings. Vision-based preprocessing and AI inference modules operate entirely on an embedded processing platform, enabling low-latency object recognition without external sensors or cloud-based computation. Detected object information, including spatial position and movement, is mapped to intelligent decision logic that generates coordinated humanoid responses. Humanoid motion is achieved through a servo-driven control architecture that enables smooth head and upper-body tracking of detected objects, enhancing visual focus and human-like behavior. In addition to physical response, the system incorporates an integrated text-to-speech mechanism that provides real-time audio has the feedback by announcing recognized objects, creating a multimodal interaction experience. The complete system is implemented on a humanoid robotic platform and evaluated under various indoor conditions to measure detection accuracy, response latency, and motion stability. Experimental results demonstrate as reliable real-time performance, to accurate object detection, smooth tracking behavior, and the effective synchronization between vision, motion, and the voice feedback. The proposed architecture is modular and scalable, supporting future extensions such as multi-object interaction, face and gesture recognition, emotion-aware responses, and intelligent adaptive learning. The system is well suited for applications in robotics education, interactive demonstrations, public exhibitions, and vision-based human–robot interaction research.
Keywords: Humanoid Robot; Object Detection; Computer Vision; Embedded Artificial Intelligence; Vision-Based Perception; Real-Time Object Tracking; Smart Response System; Human–Robot Interaction; Servo Motion Control; Text-to-Speech Interaction (search for similar items in EconPapers)
Date: 2026
References: Add references at CitEc
Citations:
Downloads: (external link)
https://ijsrst.com/home/article/view/IJSRST2613314 Abstract page (text/html)
https://ijsrst.com/home/article/download/IJSRST2613314/IJSRST2613314 Full text (application/pdf)
Related works:
This item may be available elsewhere in EconPapers: Search for items with the same title.
Export reference: BibTeX
RIS (EndNote, ProCite, RefMan)
HTML/Text
Persistent link: https://EconPapers.repec.org/RePEc:etm:ijsrst:v13:y2026:i2:id:1468
DOI: 10.32628/IJSRST2613314
Access Statistics for this article
More articles in International Journal of Scientific Research in Science and Technology from Technoscience Academy
Bibliographic data for series maintained by Pankaj Sharma ().