← Back to Model Beat
Research·6d ago·all news from August 16, 2026

Optima tackles AI benchmarking's biggest flaw by letting users test models against their own data

Artificial Analysis has released Optima, a platform that allows users to create custom AI benchmarks using their own specific data and operational workflows. By evaluating models based on internal task performance, cost, and time, the tool provides a more practical assessment for businesses than generic industry tests. This approach is particularly relevant for agent-based applications, where efficiency and outcome quality are often more critical to the user than base token pricing.

Covered by 1 source

Related stories

ResearchTwitch streamers can now opt out from training Amazon’s AIAug 12 · 8 sourcesResearchAirTag reveals how Amazon destroys rare books for AI trainingAug 17 · 5 sourcesResearchAnthropic set AI agents loose on the same task. They started a turf war.Aug 13 · 4 sourcesResearchGRPO Beyond English: A Large-Scale Study of GRPO in Non-English and Multilingual SettingsAug 17 · 3 sources