
Frontier AI Benchmarks Outperform Human Certified Accountants on Month-End Close Tasks
Testing conducted by Mercor revealed that top AI systems achieved near perfect scores on complex accounting tasks where licensed professionals averaged below forty percent. The evaluation highlighted rapid model improvement over the past year, though researchers note human accountants perform broader workplace roles.
The Blend
Artificial intelligence models have reached a new milestone in routine financial work. According to research published by Mercor, leading AI systems scored nearly perfectly when tested on standard month end accounting close procedures. In contrast, certified human accountants who took the exact same evaluation scored an average of less than forty percent.
This gap highlights how rapidly automated tools are improving at structured, rule based tasks. Just a year ago, AI models struggled with complex multi step financial calculations. Now, software can process dense spreadsheets, reconcile discrepancies, and draft accurate reporting materials faster than entry level human workers. For businesses, this suggests that high volume desk work could soon be handed off to software agents to reduce errors and cut operational costs.
However, scoring well on a standardized test does not mean AI can replace an entire finance department overnight. Human accountants handle client relationships, strategic advice, and unexpected real world edge cases that fixed benchmarks cannot measure. The main open question is how quickly accounting firms will adapt their business models, shifting human staff from routine data reconciliation toward high level financial strategy.
Written independently by AI News Smoothie from the reporting listed below. Facts belong to the original publishers. Follow the links for their full coverage.