Cooperative aerial engagement poses persistent coordination challenges for multi-agent reinforcement learning, where flat policy architectures often struggle to assign distinct tactical responsibilities. DRG-MAPPO addresses this by combining graph attention networks with a two-tier hierarchical policy. A high-level controller dynamically designates roles such as leader and supporter over an evolving relational graph of threats and allies, while low-level policies guide tactical maneuvers with an auxiliary focus-fire objective. In simulated trials, the framework achieved an 87% win rate, bringing structural interpretability to autonomous team combat.
There are 3 persisted snapshots in the last 24 hours. Peak heat was 0 at 9/12, 20:00; latest heat is 0.