Trendora

Frontier-Bench

Assess

Tools

A benchmark used to evaluate model performance on coding and knowledge work tasks.

Why it's here

Placed in Assess: 1 article(s) of evidence from 4 source(s), led by model releases, with 1 in the last 30 days. Confidence 46%.

Evidence (1)

  • 8Anthropic News·7/24/2026model_release
    Anthropic Launches Claude Opus 5

    Anthropic introduced Claude Opus 5, its new default model for Claude Max and strongest model for Claude Pro. The company says it delivers major gains in coding, knowledge work, automation, and scientific tasks while improving cost-efficiency over Opus 4.8, though it still trails Mythos 5 on cybersecurity evaluations.