Skip to main content

Research timeline

Related research and updates

Public articles linked to the same research event.

arXiv

A foundation model as critic backbone speeds multi-agent RL convergence for random access networks by at least 55%

This work proposes a foundation-model-aided, fully decentralized multi-agent reinforcement learning framework in which a self-supervised forward-dynamics foundation model serves as a reward-agnostic critic backbone and devices exchange only scalar rewards via local consensus; it provides a finite-time convergence analysis and reports at least about 55% faster convergence than end-to-end training across fair-AoI, max-sum-rate, and fair-rate random access optimization tasks.