On August 7, 2026, Artificial Analysis announced an update to its text-to-image arena to broaden the range of use cases and capabilities tested, from UI/UX design to live-action film, from physics to text rendering. The updated evaluation measures not only which model is best overall, but also which model is best for specific jobs.
The new methodology is based on a taxonomy of 10 real-world use cases and 9 model capabilities, including Marketing & Advertising, Retail & E-commerce, Live-Action Film, Animation & Gaming, Architecture & Real Estate, Productivity & Knowledge Work, UI/UX Design, as well as capabilities such as reasoning, knowledge, text rendering, layout, physics, and human anatomy.
The overall quality benchmark now samples uniformly across use cases and capabilities. Each prompt is human-curated, sourced from anonymous crowdsourced data, and refreshed monthly. Prompts that no longer differentiate models or deviate from real-world usage are retired, and new prompts at the frontier of capabilities are added to track the cutting edge and prevent overfitting.