Skip to main content
Have a personal or library account? Click to login
Dynamics-aware local trajectory control of a tracked vineyard robot via deep reinforcement learning Cover

Dynamics-aware local trajectory control of a tracked vineyard robot via deep reinforcement learning

Open Access
|Aug 2026

References

  1. R. Siegwart, I. R. Nourbakhsh, and D. Scaramuzza, Introduction to Autonomous Mobile Robots, 2nd ed. Cambridge, MA, USA: MIT Press, 2011.
  2. P. Borges, T. Peynot, S. Liang, B. Arain, M. Wildie, M. Minareci, S. Lichman, G. Samvedi, I. Sa, N. Hudson, M. Milford, P. Moghadam, and P. Corke, “A Survey on Terrain Traversability Analysis for Autonomous Ground Vehicles: Methods, Sensors, and Challenges”, Field Robotics, vol. 2, no. 1, pp. 1567–1627, 2022, DOI: 10.55417/fr.2022049.
  3. C. Stachniss, Robotic Mapping and Exploration. Berlin, Heidelberg: Springer, 2009.
  4. P. E. Hart, N. J. Nilsson, and B. Raphael, “A Formal Basis for the Heuristic Determination of Minimum Cost Paths”, IEEE Transactions on Systems Science and Cybernetics, vol. 4, no. 2, pp. 100–107, 1968, DOI: 10.1109/TSSC.1968.300136.
  5. M. Zucker, N. Ratliff, A. D. Dragan, M. Pivtoraiko, M. Klingensmith, C. M. Dellin, J. A. Bagnell, and S. S. Srinivasa, “CHOMP: Covariant Hamiltonian Optimization for Motion Planning”, The International Journal of Robotics Research, vol. 32, no. 9–10, pp. 1164–1193, 2013, DOI: 10.1177/0278364913488805.
  6. M. Kalakrishnan, S. Chitta, E. Theodorou, P. Pastor, and S. Schaal, “STOMP: Stochastic Trajectory Optimization for Motion Planning”, in 2011 IEEE International Conference on Robotics and Automation (ICRA), pp. 4569–4574, 2011, DOI: 10.1109/ICRA.2011.5980280.
  7. R. S. Sutton and A. G. Barto, Reinforcement Learning: An Introduction, 2nd ed. Cambridge, MA, USA: MIT Press, 2018.
  8. V. Mnih, K. Kavukcuoglu, D. Silver, A. A. Rusu, J. Veness, M. G. Bellemare, A. Graves, M. Riedmiller, A. K. Fidjeland, G. Ostrovski, S. Petersen, C. Beattie, A. Sadik, I. Antonoglou, H. King, D. Kumaran, D. Wierstra, S. Legg, and D. Hassabis, “Human-Level Control through Deep Reinforcement Learning”, Nature, vol. 518, no. 7540, pp. 529–533, 2015, DOI: 10.1038/nature14236.
  9. J. Kober, J. A. Bagnell, and J. Peters, “Reinforcement Learning in Robotics: A Survey”, The International Journal of Robotics Research, vol. 32, no. 11, pp. 1238–1274, 2013, DOI: 10.1177/0278364913495721.
  10. T. P. Lillicrap, J. J. Hunt, A. Pritzel, N. Heess, T. Erez, Y. Tassa, D. Silver, and D. Wierstra, “Continuous Control with Deep Reinforcement Learning”, preprint arXiv:1509.02971, 2016, DOI: 10.48550/arXiv.1509.02971.
  11. S. Fujimoto, H. van Hoof, and D. Meger, “Addressing Function Approximation Error in Actor-Critic Methods”, in Proceedings of the 35th International Conference on Machine Learning (ICML), vol. 80, pp. 1587–1596, 2018.
  12. T. Haarnoja, A. Zhou, P. Abbeel, and S. Levine, “Soft Actor-Critic: Off-Policy Maximum Entropy Deep Reinforcement Learning with a Stochastic Actor”, in Proceedings of the 35th International Conference on Machine Learning (ICML), vol. 80, pp. 1861–1870, 2018.
  13. J. Schulman, F. Wolski, P. Dhariwal, A. Radford, and O. Klimov, “Proximal Policy Optimization Algorithms”, arXiv preprint arXiv:1707.06347, 2017, DOI: 10.48550/arXiv.1707.06347.
  14. J. Schulman, P. Moritz, S. Levine, M. I. Jordan, and P. Abbeel, “High-Dimensional Continuous Control Using Generalized Advantage Estimation”, arXiv preprint arXiv:1506.02438, 2016, DOI: 10.48550/arXiv.1506.02438.
  15. S. Hochreiter and J. Schmidhuber, “Long Short-Term Memory”, Neural Computation, vol. 9, no. 8, pp. 1735–1780, 1997, DOI: 10.1162/neco.1997.9.8.1735.
  16. T. Wang, “CLE: An Integrated Framework of CNN, LSTM, and Enhanced A3C for Addressing Multi-Agent Pathfinding Challenges in Warehousing Systems”, IEEE Access, vol. 12, pp. 88904–88912, 2024, DOI: 10.1109/ACCESS.2024.3416111.
  17. D. Aghi, V. Mazzia, and M. Chiaberge, “Local Motion Planner for Autonomous Navigation in Vineyards with a RGB-D Camera-Based Algorithm and Deep Learning Synergy”, Machines, vol. 8, no. 2, pp. 27, 2020, DOI: 10.3390/machines8020027.
  18. P. Henderson, R. Islam, P. Bachman, J. Pineau, D. Precup, and D. Meger, “Deep Reinforcement Learning that Matters”, in Proceedings of the 32nd AAAI Conference on Artificial Intelligence, pp. 3207–3214, 2018, DOI: 10.1609/aaai.v32i1.11694.
  19. R. Agarwal, M. Schwarzer, P. S. Castro, A. Courville, and M. G. Bellemare, “Deep Reinforcement Learning at the Edge of the Statistical Precipice”, in Advances in Neural Information Processing Systems (NeurIPS), vol. 34, pp. 29304–29320, 2021.
  20. The MathWorks, Inc., Reinforcement Learning Toolbox. Software, Natick, MA, USA, 2024. https://www.mathworks.com/products/reinforcement-learning.html.
  21. “Vineyard Impassability RL”, GitHub repository. [Online]. Available: https://github.com/xHalaso/Vineyard_impassability_RL.
  22. A. Y. Ng, D. Harada, and S. Russell, “Policy Invariance under Reward Transformations: Theory and Application to Reward Shaping”, in Proceedings of the 16th International Conference on Machine Learning (ICML), pp. 278–287, 1999.
  23. A. Raffin, A. Hill, A. Gleave, A. Kanervisto, M. Ernestus, and N. Dormann, “Stable-Baselines3: Reliable Reinforcement Learning Implementations”, Journal of Machine Learning Research, vol. 22, no. 268, pp. 1–8, 2021.
  24. T. Haarnoja, A. Zhou, K. Hartikainen, G. Tucker, S. Ha, J. Tan, V. Kumar, H. Zhu, A. Gupta, P. Abbeel, and S. Levine, “Soft Actor-Critic Algorithms and Applications”, preprint arXiv:1812.05905, 2018, DOI: 10.48550/arXiv.1812.05905.
  25. M. Pleines, M. Pallasch, F. Zimmer, and M. Preuss, “Generalization, Mayhems and Limits in Recurrent Proximal Policy Optimization”, preprint arXiv:2205.11104, 2022, DOI: 10.48550/arXiv.2205.11104.
  26. M. Hausknecht and P. Stone, “Deep Recurrent Q-Learning for Partially Observable MDPs”, preprint arXiv:1507.06527, 2015, DOI: 10.48550/arXiv.1507.06527.
  27. L. Meng, R. Gorbet, and D. Kulić, “Memory-based Deep Reinforcement Learning for POMDPs”, in 2021 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS), pp. 5619–5626, 2021, DOI: 10.1109/IROS51168.2021.9636140.
DOI: https://doi.org/10.2478/jee-2026-0045 | Journal eISSN: 1339-309X (formerly 1335-3632) | Journal ISSN: 1335-3632
Language: English
Page range: 472 - 482
Submitted on: May 27, 2026
Published on: Aug 27, 2026
Published by: Slovak University of Technology in Bratislava
In partnership with: Paradigm Publishing Services

© 2026 Filip Zúbek, Oliver Halaš, Vendelín František Skokan, Ladislav Körösi, Ondrej Straka, Aleš Melichár, Martin Dekan, published by Slovak University of Technology in Bratislava
This work is licensed under the Creative Commons Attribution-NonCommercial-NoDerivatives 4.0 License.