跳到正文
Rohan Paul· @rohanpaul_ai · X·· 3 小时前AI 评分57
AI 导读

Rohan Paul 转述 Mercor 的一项研究结论,称在中短长度、定义清晰的会计任务上,前沿 AI 模型已比初级会计师更快更准,甚至超过研究中最优秀的人类。作者指出 Claude Opus 5 在任务中 20 题全对(100%),数分钟内完成,而 12 名持证 CPA 得分从 0% 到约 90% 不等,多人未能在 3 小时内完成;图中数据显示按每项评分标准成本计,Claude Opus 5 为 $0.21,无 AI 的会计师为 $10.35。

正文

Rule-based work isn't a career anymore. It's a prompt.

In every rule-based profession, humans are now the slow, expensive, error-prone option.

Claude Opus 5 got 20 for 20 at 100% and wrapped each task in minutes, while 12 licensed CPAs landed anywhere from 0% to about 90%, with several running out the 3-hour clock.

引用Ethan Mollick@emollick
“We find that on medium-length, well-defined accounting tasks, frontier AI models are now faster and more accurate than junior accountants, even the best one in our study.” Eighteen months ago they scored well below human accountants Good discussion here: https://www.mercor.com/blog/human-baselines-for-benchmarks-ai-now-outperforms-junior-accountants/
在 X 查看被引用的帖子

来源:Rohan Paul · x.com