Benchmarks / Model
GPT-6 Astra
Results for GPT-6 Astra by OpenAI across JumalAIta Intelligence benchmarks, with the configuration and run details of every evaluation.
01 / Results
Benchmark results
| Benchmark | Version | Rank | Evaluation | Configuration | Status |
|---|---|---|---|---|---|
| Creative Reimplementation Benchmark | v1.0 | 2 | Expert-ranked | Single shot · High | Final |
02 / Run details
Creative Reimplementation Benchmark
- Benchmark version
- v1.0
- Generation
- Single shot
- Reasoning effort
- High
- Harness
- Not yet published
- Evaluated on
- Not yet published
- Run time
- Not yet published
- Cost
- Not yet published
- Tokens
- Not yet published
- Evaluators
- The original authors of the reference demo (members of Jumalauta)
- Recording length
- 2 min 55 s
Evaluators' notes
- Rank 2
GPT-6 Astra
According to the original authors, the GPT-6 Astra version looks better at first glance: the model used image generation models for its assets, and the output contains four generated images. The authors weighted an understanding of the spirit of the work above visual polish.
