Every published VulcanBench measurement of GPT-6 Sol, newest first: combined score, Code quality, runtime, tokens and API-equivalent cost at every effort level, graded by deterministic hidden tests on real tasks. Measured through the Codex CLI on a ChatGPT subscription. GPT-6 Sol is a newer model than GPT-5.6 Sol, which has its own page; GPT-6 Luna, the other GPT-6 model on the board, has its own page too.
Public results answer "which model is stronger on this suite." Whether GPT-6 Sol belongs in
your routing table depends on your languages, your codebases, and your task mix, that gets measured, not
guessed. Grading rules are in the methodology.