The Forager's Dilemma: Peer Enforcement Alone Is Not Enough for Stable Multi-Agent Alignment
- Introduced Dual-Enforcement RL, combining intrinsic constraints with social enforcement to solve free-rider vulnerabilities in multi-agent AI systems.
- Trained independent PPO agents in a custom gridworld, reducing deceptive signaling to 1% and improving sustainable resource gathering to 90% under adversarial stress tests.