The Short Version
Claude Opus 4.7 does show clear, quantified improvements over Opus 4.6 on multiple coding-specific benchmarks, including SWE-bench Verified (80.8%→87.6%), SWE-bench Pro (53.4%→64.3%), and CursorBench (58%→70%). These figures are consistently reported across Anthropic's official documentation, the AWS News Blog, and numerous third-party writeups. The primary caveat is that the benchmark data originates from Anthropic's own reporting and has not yet been independently replicated by a third-party benchmark aggregator.