Passer à la navigation principale Passer à la recherche Passer au contenu principal

A Lagrangian Framework for Safe Cooperative Reinforcement Learning

  • Texas AandM University
  • Rensselaer Polytechnic Institute

Résultats de recherche: Le chapitre dans un livre, un rapport, une anthologie ou une collectionContribution à une conférenceRevue par des pairs

Résumé

We consider the problem of safe cooperative multiagent reinforcement learning (MARL) within the framework of a constrained multiagent Markov decision process (MDP). Agents share a common value function and learn to coordinate their actions to maximize a joint objective while adhering to system-level constraints. These constraints can enforce safety, reliability, or additional regulatory requirements governing the evolution of the multiagent system. We propose a Lagrangian-based approach, where agents iteratively solve a relaxed Lagrangian MDP using a joint learning mechanism. During execution, agents independently follow their policies, accumulating constraint violations over an epoch, which are then used to update the Lagrange multipliers. We show that continuous execution of this primal-dual algorithm produces episodes which are feasible almost surely. Further, we prove that the sequence of policies generated by the algorithm yields a nonstationary approximately optimal solution for the safe cooperative MARL problem.

langue originaleAnglais
titre2025 IEEE 64th Conference on Decision and Control, CDC 2025
EditeurInstitute of Electrical and Electronics Engineers Inc.
Pages5112-5119
Nombre de pages8
ISBN (Electronique)9798331526276
Les DOIs
étatPublié - 1 janv. 2025
Evénement64th IEEE Conference on Decision and Control, CDC 2025 - Rio de Janeiro, Brésil
Durée: 9 déc. 202512 déc. 2025

Série de publications

NomProceedings of the IEEE Conference on Decision and Control
ISSN (imprimé)0743-1546
ISSN (Electronique)2576-2370

Une conférence

Une conférence64th IEEE Conference on Decision and Control, CDC 2025
Pays/TerritoireBrésil
La villeRio de Janeiro
période9/12/2512/12/25

Empreinte digitale

Examiner les sujets de recherche de « A Lagrangian Framework for Safe Cooperative Reinforcement Learning ». Ensemble, ils forment une empreinte digitale unique.

Contient cette citation