This study explores the use of Unity3D as a versatile platform for developing, training, and evaluating intelligent agents through reinforcement learning. Leveraging the Unity ML-Agents Toolkit, a dynamic 3D environment was created to examine agent learning behavior using two advanced algorithms: Proximal Policy Optimization (PPO) and Soft Actor-Critic (SAC). The simulation environment consisted of navigable terrain bounded by red borders, with collectible blue balls serving as rewards and a purple cube representing the agent. A carefully designed reward system was implemented to encourage goal-directed behavior and penalize inefficiency, while time constraints introduced an additional challenge requiring both precision and speed. Through iterative training and refinement, the agent demonstrated increasingly complex behaviors, such as path optimization and efficient resource collection. Comparative analysis revealed that SAC exhibited rapid initial learning but suffered from performance instability due to excessive exploration, while PPO showed slower convergence but achieved more stable and consistent long-term results. The findings highlight Unity’s potential as a comprehensive simulation and experimentation framework, bridging the gap between real-time visualization and machine learning. Beyond game development, this approach can be extended to applications in robotics, industrial automation, and intelligent system design, offering an accessible yet powerful environment for studying adaptive, autonomous behaviors in virtual settings.
| Mualliflar | Rufat Mammadzada |
|---|---|
| Jurnal | Кимёвий технология. Назорат ва бошқарув |
| Nashr sanasi | 2025-11-01 |
| Jild | 2025 |
| Son | 5 |
| Betlar | 96-103 |
| Til | en |
| DOI | 10.59048/2181-1105.1722 |
DOI: 10.59048/2181-1105.1722 · Maqolaning asl sahifasi
Smart greenhouses offer a solution to sustainable food production under climate uncertainty, yet their management often depends on fixed rules or human intuition. This study proposes an intelligent decision-making…
The issue of synthesis of an adaptive control system for multi-channel and multi-mode objects is considered. Development that of effective adaptive control algorithms for nonlinear dynamic systems ensure quality…
The article studies issues related to the development of intelligent systems located in an environment capable of perceiving this environment and acting autonomously to achieve assigned tasks. To create an intelligent…
In this work, a mathematical model of industrial wastewater treatment technology was formed. First, wastewater treatment technologies were analyzed and the technology of ion-exchange resin treatment was selected. A…
Today, smart technologies are rapidly entering various areas of our lives. The role of intelligent technologies in the effective organization of control and management processes in the field of agriculture is…
The oxidative stability of vegetable oils has been introduced as a critical factor that affects on its quality, shelf life, and nutritional value. In this study, the Rancimat method was analyzed for investigation and…
This article analyzes emergency situations arising from natural disasters (fires, explosions, earthquakes, floods, landslides, etc.) and man-made accidents. The event in which the values of fire risks in the area…
This article presents a fairly reliable algorithm for generating an S-box using a trigonometric function in the parameters of the genetic algorithm. S-boxes are essential components of modern cryptography and are used…
This study explores the application of an Adaptive Neuro-Fuzzy Inference System (ANFIS) for controlling wastewater treatment processes using ion-exchange resins. It addresses the critical challenges of water scarcity…
In this article, the process of real-time prediction of the dynamics of fire development in non-stationary and gas-fired areas is carried out by developing a real-time prediction method based on artificial intelligence…