It's benchmarked in the same class as other frontier coding harnesses — for example scoring below Claude Code (running Opus 5) and Meta's Muse Spark 1.2 on a Terminal-Bench 2.1 comparison Meta published in August 2026 — marking xAI's entry into the crowded agentic-coding-harness market.