Skip to content
Research Article Open access CC BY 4.0

Advancing Production Systems with Online Reinforcement Learning: Real-time Monitoring, Control, and Optimization

Bakhtiyar Doskenov, Olanrewaju Okuyelu

Current Journal of Applied Science and Technology · pp. 1–22 · Published 27 Jan 2025

10.9734/cjast/2025/v44i24480

Abstract

Modern production systems are increasingly complex and variable, requiring adaptive and intelligent solutions for real-time monitoring and control. Traditional methods, such as linear models and static optimization, often fall short in addressing the dynamic, high-dimensional demands of industrial environments. Online reinforcement learning (RL) offers a compelling alternative by enabling systems to continuously learn and optimize decision-making through real-time interactions with their environment. This review explores the current advancements in online RL, focusing on its applications in predictive maintenance, dynamic scheduling, and process optimization. Key methodologies, including Deep RL, policy-based approaches, and hybrid frameworks, are examined for their ability to enhance scalability, adaptability, and efficiency in Industry 4.0 ecosystems. While online RL holds great promise, challenges such as computational demands, algorithmic stability, and limited real-world validation remain significant barriers to its widespread adoption. The lack of standardized benchmarks further hinders the evaluation and comparability of RL solutions across different industrial contexts. The findings underscore the potential of RL to significantly reduce operational costs by optimizing resource utilization, minimizing downtime through predictive interventions, and streamlining production workflows. These improvements collectively enhance productivity and support the creation of more agile and efficient industrial processes. To address existing gaps, this paper synthesizes recent advancements, identifies unresolved challenges, and outlines critical research pathways, including the development of efficient algorithms, the integration of domain knowledge for improved stability, and the deployment of multi-agent RL systems for distributed manufacturing networks. By providing actionable insights, this review highlights the transformative potential of online RL in creating intelligent, autonomous, and resilient production systems that deliver tangible cost and efficiency benefits in real-world applications.

Online reinforcement learning real-time monitoring production systems dynamic scheduling predictive maintenance process optimization industry 4.0

Cited by 11

A Reinforcement Learning Approach for Adaptive Energy Management

Udit Mamodiya, Indra Kishor, Mohamed M. Awad · 2025 International Conference on Future Telecommunications and Artificial Intelligence (IC-FTAI) · 2025

Hybrid Fusion Paradigm in Advanced Process Monitoring: A Panoramic Review and Future Perspectives

Husnain Ali, Rizwan Safdar, Jinfeng Liu · Industrial & Engineering Chemistry Research · 2025

Data-driven optimization of biomass conversion pathways: integrating thermochemical processes

Beemkumar Nagappan, Ganesan Subbiah, Ravi Kumar Paliwal · International Journal of Chemical Reactor Engineering · 2025

Graph-Based Knowledge Reinforcement Learning for Flexible Job-Shop Scheduling

Guolin Li, Linli Zhang, Debin Yin · Journal of Shanghai Jiaotong University (Science) · 2025

Showing 4 of 11 known citations — external sources report more than can currently be individually listed.

Article metrics

Real usage data collected on this platform.

0

Page views

0

PDF downloads

0

Outbound clicks

11

Citations

Views by country

Approximate, from request IP at view time — not citizenship or institution. Countries with fewer than 5 views are grouped as "Other".

No views recorded yet.

Traffic sources

Referring site, by host.

No traffic recorded yet.

Views and downloads exclude known bots/crawlers. Citations combines this platform's own DOI-resolved index with each external source's own reported total — see Cited by above for individually listed citing works. Last refreshed 0 seconds ago.