Benchmarks / Model
Claude Fable 5.1
Results for Claude Fable 5.1 by Anthropic across JumalAIta Intelligence benchmarks, with the configuration and run details of every evaluation.
01 / Results
Benchmark results
| Benchmark | Version | Rank | Evaluation | Configuration | Status |
|---|---|---|---|---|---|
| Creative Reimplementation Benchmark | v1.0 | 1 | Expert-ranked | Single shot · High | Final |
02 / Run details
Creative Reimplementation Benchmark
- Benchmark version
- v1.0
- Generation
- Single shot
- Reasoning effort
- High
- Harness
- Not yet published
- Evaluated on
- Not yet published
- Run time
- Not yet published
- Cost
- Not yet published
- Tokens
- Not yet published
- Evaluators
- The original authors of the reference demo (members of Jumalauta)
- Recording length
- 2 min 55 s
Evaluators' notes
- Rank 1
Claude Fable 5.1
According to the original authors, Claude Fable 5.1 understood the essence of the demo. Their example is the credits, where the model credited itself under an invented handle in the style of the group's own credits. The output is fully procedural, as its own credits state.
gfx: taas yks tekoaely joka luulee osaavansa piirtaeae
gfx: yet another AI that thinks it can drawRecording, about 1:28 gfx: ei yhtaeaen kuvaa, kaikki proseduraalista
gfx: not a single image, everything proceduralRecording, about 1:40
