JumalAIta IntelligenceIndependent evaluation

Benchmarks / Model

Claude Fable 5.1

Results for Claude Fable 5.1 by Anthropic across JumalAIta Intelligence benchmarks, with the configuration and run details of every evaluation.

Organization
Anthropic
Benchmarks
1

01 / Results

Benchmark results

Results for Claude Fable 5.1, one row per benchmark.
BenchmarkVersionRankEvaluationConfigurationStatus
Creative Reimplementation Benchmarkv1.01Expert-rankedSingle shot · HighFinal

02 / Run details

Creative Reimplementation Benchmark

Claude Fable 5.1

Rank 1

Anthropic · Single shot · High reasoning

Benchmark version
v1.0
Generation
Single shot
Reasoning effort
High
Harness
Not yet published
Evaluated on
Not yet published
Run time
Not yet published
Cost
Not yet published
Tokens
Not yet published
Evaluators
The original authors of the reference demo (members of Jumalauta)
Recording length
2 min 55 s

Evaluators' notes

  1. Rank 1

    Claude Fable 5.1

    According to the original authors, Claude Fable 5.1 understood the essence of the demo. Their example is the credits, where the model credited itself under an invented handle in the style of the group's own credits. The output is fully procedural, as its own credits state.

    gfx: taas yks tekoaely joka luulee osaavansa piirtaeae
    gfx: yet another AI that thinks it can drawRecording, about 1:28
    gfx: ei yhtaeaen kuvaa, kaikki proseduraalista
    gfx: not a single image, everything proceduralRecording, about 1:40