1 story in this blend

Nous Research launched the Hermes Index, a public benchmark that tests AI models on agent performance while measuring the financial cost to finish each job. Models perform standardized automated tasks including research, email triage, and diagram generation. The initial leaderboard shows high end models leading in quality, while smaller budget models offer dramatically lower prices for basic operations.