arxivcs.LGcs.AImath.OC2026-07-05
Regime-Conditional Stabilisation of LLM-Augmented Cooperative Multi-Agent Reinforcement Learning
Faid Keddouri, Sohaib Houhou, Aissa Boulmerka, Nadir Farhi
Large Language Models (LLMs) offer a natural interface for translating human objectives into reward signals for cooperative multi-agent reinforcement learning (MARL), yet the training-time dynamics of this integration remain poorly understood. We show that dynamically updating LL…