Specialist test. Evaluate this separately from the eight core tests. Explore the active suite →
Focused model compare
City Scroll Journey
Qwen3.8-Max is the starting run, compared with Claude Opus 5.5, GPT-6 Astra, GPT-6 Sol.
Starting model
Qwen3.8-Max
Anthropic
Claude Opus 5.5
OpenAI
GPT-6 Astra
OpenAI