Code generation: a model at 1/271st the price of the best scorer, at 97.8% of its quality.

claude-opus-5-fast vs gpt-4.1 for rewriting and editing

the measured answer

As of August 24, 2026, claude-opus-5-fast measures 0.950 on rewriting & editing vs 0.864 for gpt-4.1 (+0.086), at 15.0× the price ($29.93 vs $2.00 per 1K requests).

same suite, same items, same scoring — frontier v5, measured 2026-08-24 · method

modelvendormeasured quality$ / 1K requestsp95 latency
claude-opus-5-fastanthropic0.950$29.937591 ms
gpt-4.1openai0.864$2.002174 ms

Rewrite-and-edit work — tone changes, tightening, constraint-preserving edits — punishes models that "improve" text by discarding requirements. The measured suite scores whether stated constraints survive the edit, which is where cheap and expensive models genuinely separate.

These two are part of a larger measured frontier — what is the best model for rewriting and editing text shows every measured option for this workload, and other workloads rank these models differently: a model that wins here can lose on another kind of work, which is the whole argument for routing per workload rather than picking one model for everything.

the routed alternative

Potion routes each request to the cheapest option measured at your quality bar — including picks this public page does not name. Get an API key or read the docs.

claude-opus-5-fast vs gpt-4.1 for rewriting and editing — measured · Frontier Notes