Skip to content
Open access

Learning to Schedule Machines and Operators: A Human-Aware Deep Q-Learning Framework for Dynamic Flexible Job Shops

2026 · International Journal of Advanced Computer Science and Applications · 0 citations · 48 references

TL;DR

A human-centered Deep Reinforcement Learning framework, DD4LQN, for dynamic flexible job shop scheduling under operator learning and forgetting dynamics, which integrates a disturbance-aware scenario generator, bounded logistic learning-forgetting dynamics, and deterministic Dual Pair Ranking to ensure auditable decisions.

Abstract

Most learning-based schedulers for job shop problems assume static workforce performance. However, dynamic dual-resource job shops must navigate stochastic disturbances and time-varying task durations due to human factors. This study proposes a human-centered Deep Reinforcement Learning framework, DD4LQN, for dynamic flexible job shop scheduling under operator learning and forgetting dynamics. The environment integrates a disturbance-aware scenario generator, bounded logistic learning-forgetting dynamics, and deterministic Dual Pair Ranking to ensure auditable decisions. The training used a plan that included ENTRY and EXIT greedy evaluations and a Quote-then-Commit process. Tests over ten weeks showed that DD4LQN-EXIT did better than other scheduling methods. Empirical evaluations over a ten-week horizon demonstrate that DD4LQN-EXIT got the average schedule reward of 0.5843. It was better than Earliest Due Date (EDD) and Shortest Processing Time (SPT) by 17% and 14%, respectively. It also completed jobs and had fewer risky orders. Even though Shortest Processing Time had delays on average, it completed fewer jobs because it focused on short tasks. Additionally, an exploratory ablation across the same scenarios demonstrated that Dual Pair Rank and learning–forgetting dynamics come up with better results. Furthermore, representation diagnostics on the 227,717-parameter network confirm structural stability, with layer matrices retaining up to 98.2% of maximum effective rank without capacity collapse.

Read PDF

Similar papers

Open access Sep 2026

A Two-Layer Multi-Agent Deep Reinforcement Learning Framework for Flexible Job-Shop Scheduling with Multiple Batch-Processing Machines

An extended FJSP with multiple BPMs is formed and an end-to-end two-layer multi-agent deep reinforcement learning framework is proposed, supporting the framework as an effective scheduling approach for deterministic FJSP with BPMs and indicating cross-scale generalization across evaluated instances.

Ze-Peng Liu, Ai-Ming Wang · 0 citations
Book Open access Aug 2026

EDA Job Scheduling Using Reinforcement Learning with Adaptive Macro Actions

Adaptive Job Selection (AJS), a reinforcement learning-based agent that learns to schedule pending jobs for reduced job waiting and completion time on LSF clusters for EDA workloads, is presented, becoming the first open-source, deployable RL-based scheduler designed for production EDA environments.

Yiming Shao, Aijun An, Michael Spriggs et al. · 0 citations
#reinforcement learning Conference Open access Sep 2026

A Deep Reinforcement Learning (DRL) Based Transformer Method for Solving the Open Shop Scheduling Problem

In this study, we investigate the potential of a Deep Reinforcement Learning (DRL) based Transformer neural network architecture to solve large-scale open shop scheduling problems. The open shop scheduling problem (OSSP) involves sequencing n jobs across m machines, where each machine handles only one job at a time, an...

Faezeh Ardali, Gerald M. Knapp · 0 citations
Conference

Dynamic Graph-Based Reinforcement Learning for Efficient Scheduling in Large-Scale Flexible Job Shops

The Flexible Job Shop Scheduling Problem (FJSP) is an NP-hard optimization challenge with significant industrial applications, especially for large-scale instances. Traditional approaches, such as Priority Dispatching Rules (PDRs), often struggle with time-intensive design processes and suboptimal performance as proble...

Jeongwon Park, Feng Ju · 0 citations
2026

Feedback-Driven Population Self-Evolution Framework for Dispatching Rule Generation in Dynamic Job Shop via Knowledge Distillation

The dynamic job shop scheduling problem (DJSSP) is critical for optimizing production efficiency in intelligent manufacturing systems under dynamic constraints. Traditional approaches, including heuristic dispatching rules (HDRs) and evolutionary hyper-heuristics, often struggle to generalize across dynamic and unseen...

Jin Huang, Zhengqi Shi, Qi-Hao Liu et al. · 0 citations
Book Open access Sep 2026

ReLA: Representation Learning and Aggregation for Scalable Job Scheduling with Reinforcement Learning

ReLA is an RL scheduler built on structured representation learning and aggregation that learns intra-entity representations using self-attention and convolution, captures inter-entity operation–machine interactions using cross-attention, and aggregates multi-scale representations for parallel actor-based scoring of fe...

Zheng-Yi Kwan, Wei Zhang, Aik Beng Ng et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.