New Open Benchmark Creates Global Standard for Evaluating Enterprise AI. DevRev Tops the Leaderboard
DevRev, maker of the agentic AI product Computer, today announced Enterprise-Bench, an open, vendor-neutral benchmark for evaluating whether AI agents can operate in production enterprise environments ...
Moonshot AI's 2.8-trillion-parameter model Kimi K3 tops Anthropic's Claude Fable 5 and OpenAI's GPT 5.6 Sol on impressive ...
Results that may be inaccessible to you are currently showing.
Hide inaccessible results