Specialist test. Evaluate this separately from the eight core tests. Explore the active suite →
Focused model compare
City Scroll Journey
GPT-6 Luna is the starting run, compared with Claude Opus 5.5, GPT-6 Astra, GPT-6 Sol.
Starting model
GPT-6 Luna
Anthropic
Claude Opus 5.5
OpenAI
GPT-6 Astra
OpenAI