Advancing Production Systems with Online Reinforcement Learning: Real-time Monitoring, Control, and Optimization
Bakhtiyar Doskenov, Olanrewaju Okuyelu
Current Journal of Applied Science and Technology · pp. 1–22 · Published 27 Jan 2025
10.9734/cjast/2025/v44i24480Abstract
Modern production systems are increasingly complex and variable, requiring adaptive and intelligent solutions for real-time monitoring and control. Traditional methods, such as linear models and static optimization, often fall short in addressing the dynamic, high-dimensional demands of industrial environments. Online reinforcement learning (RL) offers a compelling alternative by enabling systems to continuously learn and optimize decision-making through real-time interactions with their environment. This review explores the current advancements in online RL, focusing on its applications in predictive maintenance, dynamic scheduling, and process optimization. Key methodologies, including Deep RL, policy-based approaches, and hybrid frameworks, are examined for their ability to enhance scalability, adaptability, and efficiency in Industry 4.0 ecosystems. While online RL holds great promise, challenges such as computational demands, algorithmic stability, and limited real-world validation remain significant barriers to its widespread adoption. The lack of standardized benchmarks further hinders the evaluation and comparability of RL solutions across different industrial contexts. The findings underscore the potential of RL to significantly reduce operational costs by optimizing resource utilization, minimizing downtime through predictive interventions, and streamlining production workflows. These improvements collectively enhance productivity and support the creation of more agile and efficient industrial processes. To address existing gaps, this paper synthesizes recent advancements, identifies unresolved challenges, and outlines critical research pathways, including the development of efficient algorithms, the integration of domain knowledge for improved stability, and the deployment of multi-agent RL systems for distributed manufacturing networks. By providing actionable insights, this review highlights the transformative potential of online RL in creating intelligent, autonomous, and resilient production systems that deliver tangible cost and efficiency benefits in real-world applications.
Cited by 11
Udit Mamodiya, Indra Kishor, Mohamed M. Awad · 2025 International Conference on Future Telecommunications and Artificial Intelligence (IC-FTAI) · 2025
Husnain Ali, Rizwan Safdar, Jinfeng Liu · Industrial & Engineering Chemistry Research · 2025
Beemkumar Nagappan, Ganesan Subbiah, Ravi Kumar Paliwal · International Journal of Chemical Reactor Engineering · 2025
Guolin Li, Linli Zhang, Debin Yin · Journal of Shanghai Jiaotong University (Science) · 2025
Showing 4 of 11 known citations — external sources report more than can currently be individually listed.
Related research
- A Review of Artificial Neural Networks for Chemical Process Optimization and Compound Property Prediction — shares topic coverage
- Optimization of Solvent Oil Extraction from Sandbox (Hura crepitans Linn.) Seeds — shares topic coverage
- Data-Driven Techniques and Data Analytics in Water Treatment Facilities: Innovative Safety Protocols and Optimization — shares topic coverage
- The Influence of Anti-scatter Grid Usage for Knee Computerized Radiography — shares topic coverage
- Research on Chemical Process Optimization Based on Artificial Neural Network Algorithm — shares topic coverage
Article metrics
Real usage data collected on this platform.
0
Page views
0
PDF downloads
0
Outbound clicks
11
Citations
Views by country
Approximate, from request IP at view time — not citizenship or institution. Countries with fewer than 5 views are grouped as "Other".
No views recorded yet.
Traffic sources
Referring site, by host.
No traffic recorded yet.
Views and downloads exclude known bots/crawlers. Citations combines this platform's own DOI-resolved index with each external source's own reported total — see Cited by above for individually listed citing works. Last refreshed 0 seconds ago.