⚠ Official Notice: www.ijisrt.com is the official website of the International Journal of Innovative Science and Research Technology (IJISRT) Journal for research paper submission and publication. Please beware of fake or duplicate websites using the IJISRT name.



Layer-by-Layer Perception, Sensor Fusion and Closed-Loop Control in a Low-Cost Modular Autonomous Ground Vehicle: Design, Wiring and Simulation-Based Characterisation


Authors : Srikanth Annamareddy; Komal Sai Raj Atmakuri; Sai Revanth Singavarapu; Durga Venkatesh Janaki

Volume/Issue : Volume 11 - 2026, Issue 8 - August


Google Scholar : https://tinyurl.com/2s3vx3dv

DOI : https://doi.org/10.38124/ijisrt/26aug1405

Note : A published paper may take 4-5 working days from the publication date to appear in PlumX Metrics, Semantic Scholar, and ResearchGate.


Abstract : Most papers on autonomous-vehicle software either stop at a literature survey, or describe a full industrial stack that assumes lidar-grade sensing and a rack of GPUs. Not much work actually shows the pin-out and the line of firmware, how the textbook sense-perceive-plan-act loop actually runs on hardware a student can put together in a weekend. That is the gap this paper tries to close. We start by going back through the layered autonomy pipeline, sensing and fusion, perception, localisation, planning, and control, and tying each layer to a decision we actually made while building the vehicle, framed throughout against the SAE J3016 taxonomy of driving-automation levels, operational design domain, and dynamic driving task. We then walk through, wiring diagram in hand, a two-tier prototype built around a Raspberry Pi 4 for vision and decision-making and an Arduino Uno for time-critical sensing and motor control, talking to each other over a 115200- baud serial link. The perception pipeline is given layer by layer, greyscale, blur, Canny, ROI mask, Hough transform, with the OpenCV code alongside it, and the fusion stage is a scalar Kalman filter combining ultrasonic range with inertial motion. Since physical field trials were not something we could run for this paper, we built our own seeded, reproducible Python simulation of the vehicle's sensors and control loops instead of just describing what we expect would happen, and we report the numbers that code actually produced. Every quantitative result in this paper comes from a seeded software simulation, not physical hardware trials. Root-mean-square ultrasonic error drops from 3.80 ± 0.89 cm raw to 1.67 ± 0.72 cm after the two-stage filter (averaged across 50 seeded runs); the lane pipeline detects a lane in every one of 300 synthetic frames with a mean offset error of 0.74 pixels; and, in a forward-collision scenario where a camera-detected lead vehicle brakes hard from 40 km/h, the adaptive-following and collision-warning logic escalates through CAUTION, WARNING, and AUTOBRAKE states and brings the ego vehicle to a stop with 2.7-3.3 m of gap still remaining, no collision, in both a moderate and a harder braking test. We close by tying these results back to the literature and by laying out what it would take to move this platform toward ROS 2, a learned perception model, and CARLA-based and physical testing. Every wiring connection, pin assignment, and piece of simulation code needed to reproduce this paper is reported in full.

Keywords : Autonomous Ground Vehicle, Sensor Fusion, Kalman Filter, OpenCV, Canny Edge Detection, Hough Transform, FiniteState Planning, PID Control, Robot Operating System, Embedded Systems, Controller Area Network, Raspberry Pi, Arduino.

References :

  1. SAE International, "Taxonomy and Definitions for Terms Related to Driving Automation Systems for On-Road Motor Vehicles," SAE Standard J3016_202104, Apr. 2021.
  2. J. Kocic, N. Jovicic, and V. Drndarevic, "Sensors and sensor fusion in autonomous vehicles," in Proc. 26th Telecommunications Forum (TELFOR), Belgrade, Serbia, 2018, pp. 420-425.
  3. J. Redmon, S. Divvala, R. Girshick, and A. Farhadi, "You only look once: Unified, real-time object detection," in Proc. IEEE Conf. Computer Vision and Pattern Recognition (CVPR), Las Vegas, NV, USA, 2016, pp. 779-788.
  4. C. Cadena, L. Carlone, H. Carrillo, Y. Latif, D. Scaramuzza, J. Neira, I. Reid, and J. J. Leonard, "Past, present, and future of simultaneous localization and mapping: Toward the robust-perception age," IEEE Trans. Robotics, vol. 32, no. 6, pp. 1309-1332, Dec. 2016.
  5. G. Welch and G. Bishop, "An Introduction to the Kalman Filter," Univ. North Carolina at Chapel Hill, Chapel Hill, NC, USA, Tech. Rep. TR 95-041, 1995.
  6. B. Paden, M. Cap, S. Z. Yong, D. Yershov, and E. Frazzoli, "A survey of motion planning and control techniques for self-driving urban vehicles," IEEE Trans. Intelligent Vehicles, vol. 1, no. 1, pp. 33-55, Mar. 2016.
  7. A. Dosovitskiy, G. Ros, F. Codevilla, A. Lopez, and V. Koltun, "CARLA: An open urban driving simulator," in Proc. 1st Annu. Conf. Robot Learning (CoRL), Mountain View, CA, USA, 2017, pp. 1-16.
  8. M. Quigley, K. Conley, B. Gerkey, J. Faust, T. Foote, J. Leibs, R. Wheeler, and A. Y. Ng, "ROS: An open-source Robot Operating System," in Proc. ICRA Workshop on Open Source Software, Kobe, Japan, 2009.
  9. S. Kato, S. Tokunaga, Y. Maruyama, S. Maeda, M. Hirabayashi, Y. Kitsukawa, A. Monrroy, T. Ando, Y. Fujii, and T. Azumi, "Autoware on board: Enabling autonomous vehicles with embedded systems," in Proc. ACM/IEEE 9th Int. Conf. Cyber-Physical Systems (ICCPS), Porto, Portugal, 2018, pp. 287-296.
  10. A. Krizhevsky, I. Sutskever, and G. E. Hinton, "ImageNet classification with deep convolutional neural networks," in Advances in Neural Information Processing Systems (NeurIPS), vol. 25, 2012, pp. 1097-1105.
  11. R. Girshick, J. Donahue, T. Darrell, and J. Malik, "Rich feature hierarchies for accurate object detection and semantic segmentation," in Proc. IEEE Conf. Computer Vision and Pattern Recognition (CVPR), Columbus, OH, USA, 2014, pp. 580-587.
  12. J. Long, E. Shelhamer, and T. Darrell, "Fully convolutional networks for semantic segmentation," in Proc. IEEE Conf. Computer Vision and Pattern Recognition (CVPR), Boston, MA, USA, 2015, pp. 3431-3440.
  13. O. Ronneberger, P. Fischer, and T. Brox, "U-Net: Convolutional networks for biomedical image segmentation," in Proc. Int. Conf. Medical Image Computing and Computer-Assisted Intervention (MICCAI), Munich, Germany, 2015, pp. 234-241.
  14. K. He, X. Zhang, S. Ren, and J. Sun, "Deep residual learning for image recognition," in Proc. IEEE Conf. Computer Vision and Pattern Recognition (CVPR), Las Vegas, NV, USA, 2016, pp. 770-778.
  15. M. Bojarski, D. Del Testa, D. Dworakowski, B. Firner, B. Flepp, P. Goyal, L. D. Jackel, M. Monfort, U. Muller, J. Zhang, X. Zhang, J. Zhao, and K. Zieba, "End to end learning for self-driving cars," arXiv preprint arXiv:1604.07316, 2016.
  16. M. Montemerlo, S. Thrun, D. Koller, and B. Wegbreit, "FastSLAM: A factored solution to the simultaneous localization and mapping problem," in Proc. AAAI Conf. Artificial Intelligence, Edmonton, AB, Canada, 2002, pp. 593-598.
  17. S. Thrun, M. Montemerlo, H. Dahlkamp, D. Stavens, A. Aron, J. Diebel, P. Fong, J. Gale, M. Halpenny, G. Hoffmann, et al., "Stanley: The robot that won the DARPA Grand Challenge," J. Field Robotics, vol. 23, no. 9, pp. 661-692, 2006.
  18. C. Urmson, J. Anhalt, D. Bagnell, C. Baker, R. Bittner, M. N. Clark, J. Dolan, D. Duggins, T. Galatali, C. Geyer, et al., "Autonomous driving in urban environments: Boss and the Urban Challenge," J. Field Robotics, vol. 25, no. 8, pp. 425-466, 2008.
  19. A. Geiger, P. Lenz, and R. Urtasun, "Are we ready for autonomous driving? The KITTI vision benchmark suite," in Proc. IEEE Conf. Computer Vision and Pattern Recognition (CVPR), Providence, RI, USA, 2012, pp. 3354-3361.
  20. C. Chen, A. Seff, A. Kornhauser, and J. Xiao, "DeepDriving: Learning affordance for direct perception in autonomous driving," in Proc. IEEE Int. Conf. Computer Vision (ICCV), Santiago, Chile, 2015, pp. 2722-2730.
  21. S. Macenski, T. Foote, B. Gerkey, C. Lalancette, and W. Woodall, "Robot Operating System 2: Design, architecture, and uses in the wild," Science Robotics, vol. 7, no. 66, 2022.
  22. International Organization for Standardization, "ISO 26262-1:2018 Road vehicles - Functional safety" and "ISO/PAS 21448:2019 Road vehicles - Safety of the intended functionality," Geneva, Switzerland, 2018-2019.
  23. P. Koopman and M. Wagner, "Challenges in autonomous vehicle testing and validation," SAE Int. J. Transportation Safety, vol. 4, no. 1, pp. 15-24, 2016.
  24. N. Navet and F. Simonot-Lion, Eds., Automotive Embedded Systems Handbook. Boca Raton, FL, USA: CRC Press, 2008.
  25. S. Ren, K. He, R. Girshick, and J. Sun, "Faster R-CNN: Towards real-time object detection with region proposal networks," in Advances in Neural Information Processing Systems (NeurIPS), vol. 28, 2015, pp. 91-99.

Most papers on autonomous-vehicle software either stop at a literature survey, or describe a full industrial stack that assumes lidar-grade sensing and a rack of GPUs. Not much work actually shows the pin-out and the line of firmware, how the textbook sense-perceive-plan-act loop actually runs on hardware a student can put together in a weekend. That is the gap this paper tries to close. We start by going back through the layered autonomy pipeline, sensing and fusion, perception, localisation, planning, and control, and tying each layer to a decision we actually made while building the vehicle, framed throughout against the SAE J3016 taxonomy of driving-automation levels, operational design domain, and dynamic driving task. We then walk through, wiring diagram in hand, a two-tier prototype built around a Raspberry Pi 4 for vision and decision-making and an Arduino Uno for time-critical sensing and motor control, talking to each other over a 115200- baud serial link. The perception pipeline is given layer by layer, greyscale, blur, Canny, ROI mask, Hough transform, with the OpenCV code alongside it, and the fusion stage is a scalar Kalman filter combining ultrasonic range with inertial motion. Since physical field trials were not something we could run for this paper, we built our own seeded, reproducible Python simulation of the vehicle's sensors and control loops instead of just describing what we expect would happen, and we report the numbers that code actually produced. Every quantitative result in this paper comes from a seeded software simulation, not physical hardware trials. Root-mean-square ultrasonic error drops from 3.80 ± 0.89 cm raw to 1.67 ± 0.72 cm after the two-stage filter (averaged across 50 seeded runs); the lane pipeline detects a lane in every one of 300 synthetic frames with a mean offset error of 0.74 pixels; and, in a forward-collision scenario where a camera-detected lead vehicle brakes hard from 40 km/h, the adaptive-following and collision-warning logic escalates through CAUTION, WARNING, and AUTOBRAKE states and brings the ego vehicle to a stop with 2.7-3.3 m of gap still remaining, no collision, in both a moderate and a harder braking test. We close by tying these results back to the literature and by laying out what it would take to move this platform toward ROS 2, a learned perception model, and CARLA-based and physical testing. Every wiring connection, pin assignment, and piece of simulation code needed to reproduce this paper is reported in full.

Keywords : Autonomous Ground Vehicle, Sensor Fusion, Kalman Filter, OpenCV, Canny Edge Detection, Hough Transform, FiniteState Planning, PID Control, Robot Operating System, Embedded Systems, Controller Area Network, Raspberry Pi, Arduino.

Paper Submission Last Date
30 - September - 2026

SUBMIT YOUR PAPER CALL FOR PAPERS
Video Explanation for Published paper

Never miss an update from Papermashup

Get notified about the latest tutorials and downloads.

Subscribe by Email

Get alerts directly into your inbox after each post and stay updated.
Subscribe
OR

Subscribe by RSS

Add our RSS to your feedreader to get regular updates from us.
Subscribe