Search

Tag: #marl5 results

Google AI 代理對抗不可預測對手學會合作

Google AI 代理對抗不可預測對手學會合作

Google 的 Paradigms of Intelligence 團隊發現,將 AI 代理對抗多樣且不可預測的對手訓練,即可產生合作行為,而無需硬編碼規則。透過分散式強化學習,LLM 代理利用情境學習即時適應混合動機情境。此可擴展方法適合企業多代理系統。

VentureBeatMediaMar 11#multi-agent#marl#cooperation
DrIGM Enables Robust Multi-Agent RL

DrIGM Enables Robust Multi-Agent RL

DrIGM introduces distributionally robust IGM for MARL, ensuring decentralized actions align under uncertainties via robust value factorization. Compatible with VDN/QMIX/QTRAN without reward shaping. Boosts OOD performance in SustainGym and StarCraft.

ArXiv AIResearchFeb 13#research#arxiv#drigm