Independent benchmark publication
Compare AI on practical Workday® tasks
Explore planned Finance, Reporting, Integrations, and Adaptive Planning task areas. Approved results stay tied to their dataset, product scope, evaluation track, and run conditions, so practitioners can compare like for like without blending unlike runs.
No approved benchmark results have been published. The category pages describe intended task areas, not completed questions or verified product instructions.
Initial task areas
Browse Workday benchmark categories
Each category page shows its planned topics and any approved runs for that area. A published result in one area does not mean the other areas have results or complete question coverage.
How to compare results
Compare models within one published run
Each run identifies its dataset and rubric versions, product scope, evaluation track, environment, and question count. Compare models within that documented scope; unlike runs are not blended into a site-wide score or ranking. Quality, completion, latency, and cost are separate measures. Missing values stay unknown, and failures remain visible in coverage.
Read the publication methodology