Optimal Cooperative Control of Multi‐Modular Robot Manipulator Based on Adaptive Dynamic Programming With Experience Replay in Nonzero‐Sum Games

Bo Dong et al.

Optimal Control Applications and Methods2026https://doi.org/10.1002/oca.70096article
ABDC B
Weight
0.50

What the paper says

ABSTRACT This paper proposes an adaptive dynamic programming (ADP) method based on nonzero‐sum (NZS) game theory for the optimal cooperative control of a multi‐modular robot manipulator (MMRM). First, the system dynamics are established via the Newton‐Euler iterative algorithm, and load distribution is adopted to allocate driving forces to each module while maintaining force balance. The optimal cooperative control problem is then formulated as an NZS game, with each joint of the modular manipulators treated as a player. A radial basis function neural network (RBFNN)‐based state observer is constructed to estimate model unknowns. Moreover, a novel critic neural network weight‐adjustment rule that incorporates experience replay is proposed to relax the persistent excitation (PE) condition. Finally, Lyapunov theory proves the ultimate uniform boundedness (UUB) of the closed‐loop system errors, and experiments verify the method's effectiveness.

Open paper page →

Cite this paper

https://doi.org/https://doi.org/10.1002/oca.70096

Or copy a formatted citation

@article{bo2026,
  title        = {{Optimal Cooperative Control of Multi‐Modular Robot Manipulator Based on Adaptive Dynamic Programming With Experience Replay in Nonzero‐Sum Games}},
  author       = {Bo Dong et al.},
  journal      = {Optimal Control Applications and Methods},
  year         = {2026},
  doi          = {https://doi.org/https://doi.org/10.1002/oca.70096},
}

Paste directly into BibTeX, Zotero, or your reference manager.

Flag this paper

Optimal Cooperative Control of Multi‐Modular Robot Manipulator Based on Adaptive Dynamic Programming With Experience Replay in Nonzero‐Sum Games

Flags are reviewed by the Arbiter methodology team within 5 business days.


Evidence weight

0.50

Balanced mode · F 0.40 / M 0.15 / V 0.05 / R 0.40

F · citation impact0.50 × 0.4 = 0.20
M · momentum0.50 × 0.15 = 0.07
V · venue signal0.50 × 0.05 = 0.03
R · text relevance †0.50 × 0.4 = 0.20

† Text relevance is estimated at 0.50 on the detail page — for your query’s actual relevance score, open this paper from a search result.