isitcooked.ai
← leaderboard

Claude Opus 4.8

NOT COOKED

Anthropic · anthropic/claude-opus-4.8 · verdict as of 2026-07-29

Daily test history (90 days, baseline band = trailing mean ± 2σ)

Structured Extraction

NOT COOKED

Math & Reasoning

NOT COOKED

Code Generation

NOT COOKED

Game Design

CALIBRATING

Instruction Following

NOT COOKED

Report Analysis

NOT COOKED

Summarization

NOT COOKED

Customer Service

NOT COOKED

Web Design

NOT COOKED

Creative Writing

CALIBRATING

Serving providers (last 30 days)

striped = multiple providers served this model that day (hover for detail)

Public benchmarks overall 86.7

MMLU-Pro

GPQA Diamond

93.6

SWE-bench Verified

LMArena Elo

1479

AIME 2025

retrieved 2026-07-03 from public sources — see methodology

Recent samples (latest run, one per test case)

No samples from the latest run.