Skip to content
AI Atlas
PaperActive

Distributed Optimization of Modular Production Systems using Model-based Reinforcement Learning with Inverse Models

arxiv.org/abs/2609.11615

quality89

Updated 2 h ago · first seen 11 Sept 2026

paper_01M294FR4WXH5QT5JYXCC0286Y

Published
11 Sept 2026
T1 · 2 h ago
arXiv
2609.11615
T1 · 2 h ago
Category
cs.AI
T1 · 2 h ago

Abstract

This paper presents a novel approach for data-driven self-learning control of highly flexible, modular manufacturing systems. Specifically, we employ a novel framework for model-based reinforcement learning which introduces approximate inverse process models within the training of reinforcement policies. This approach disentangles the learning of actuation dynamics and the dynamics in state space, resulting in RL-based training solely within the task space. We propose a lightweight feedforward architecture for approximate inverse models and integrate them within the policy network of standard RL algorithms. We apply the approach to a laboratory modular production testbed with heterogeneous production modules. The results underline the efficiency improvements for modular manufacturing units in terms of both performance and training speed, particularly for off-policy algorithms.

Authors 4

Andreas Schwung, Steve Yuwono, Sofiene Lassoued, Dorothea Schwung

Specification

Official page

Source:arXiv (Atom API + RSS)T1observed 2 h agohigh

Arxiv announce type
cross

Source:arXiv (Atom API + RSS)T1observed 2 h agohigh

arXiv id
2609.11615

Source:arXiv (Atom API + RSS)T1observed 2 h agohigh

Categories
cs.AI, cs.LG, cs.SY, eess.SY

Source:arXiv (Atom API + RSS)T1observed 2 h agohigh

PDF

Source:arXiv (Atom API + RSS)T1observed 2 h agohigh

Primary category
cs.AI

Source:arXiv (Atom API + RSS)T1observed 2 h agohigh

Published
11 Sept 2026

Source:arXiv (Atom API + RSS)T1observed 2 h agohigh

Each value shows its source, tier and observation time. Conflicting claims are kept side by side and flagged — never averaged. How AI Atlas records facts →

Provenance

Attributed facts

9

Source tiers

T19

Freshest observation

2 h ago

Conflicts

None