JumalAIta IntelligenceIndependent evaluation

Benchmarks / Model

Claude Opus 4.7

Results for Claude Opus 4.7 by Anthropic across JumalAIta Intelligence benchmarks, with the configuration and run details of every evaluation.

Organization
Anthropic
Benchmarks
1

01 / Results

Benchmark results

Results for Claude Opus 4.7, one row per benchmark.
BenchmarkVersionRankEvaluationConfigurationStatus
Creative Reimplementation Benchmarkv1.03Expert-rankedSingle shot · HighFinal

02 / Run details

Creative Reimplementation Benchmark

Claude Opus 4.7

Rank 3

Anthropic · Single shot · High reasoning

Benchmark version
v1.0
Generation
Single shot
Reasoning effort
High
Harness
Not yet published
Evaluated on
Not yet published
Run time
Not yet published
Cost
Not yet published
Tokens
Not yet published
Evaluators
The original authors of the reference demo (members of Jumalauta)
Recording length
2 min 55 s
Recording published
2026-09-06

Evaluators' notes

  1. Rank 3

    Claude Opus 4.7

    According to the original authors, the Claude Opus 4.7 version looks noticeably rougher than the two newer entries.