DPBench: Large Language Models Struggle with Simultaneous Coordination
arXiv:2602.13255v1 Announce Type: new Abstract: Large language models are increasingly deployed in multi-agent systems, yet we lack benchmarks that test whether they can coordinate under …
Najmul Hasan, Prashanth BusiReddyGari
130 views