Benchmark tests show multi-agent AI teams struggle with group tasks
Models & ResearchSuperintelligence · 2h ago

Benchmark tests show multi-agent AI teams struggle with group tasks

A research study evaluating groups of up to 20 AI agents found that team success rates peaked at 52 percent on complex multi-step objectives. The findings indicate that simply increasing chat communication between agents is insufficient for solving coordination problems without structured division of labor.

Columbia UniversityUniversity of PennsylvaniaOpenAIGoogle
Read the original