
Artificial Analysis launched Optima — a platform where users can create their own AI-model tests based on their data and workflows.
Optima lets users compare models not only by quality, but also by cost and task execution time. The description of the platform separately notes that for agent-based applications these metrics may be more informative than token price alone.
The source does not provide data on the number of users, results of specific tests, or comparison of individual models. The information is based on The Decoder's synopsis, not the full text of the original source.
editorial commentary
Why it matters
A likely consequence is that users will more often choose models based on the overall effectiveness for a specific task, rather than on a universal rating or token price. The next observable signal will be published comparisons on real user data. Substantial uncertainty — the source provides no information about Optima's methodology, scale of rollout, or testing results.