Current verdict ·

The terminal default—and the two reasons to choose differently.

Codex Pro is the terminal-agent default now for workflow and current capability fit—not a durable capacity advantage. Choose Cursor Ultra for an IDE-first visible pool; choose Claude Code Max 20× for Claude-native workflow affinity.

  1. ● Terminal-agent default

    Codex Pro

    Workflow
    Terminal-first agent loops across the local CLI and Codex cloud.
    Capability
    GPT-5.6 Sol, Terra, and Luna replaced GPT-5.5 in the current Codex model line.
    Current headroom
    The five-hour gate is temporarily suspended; weekly limits remain. Durable headroom and the suspension duration are unknown.
    Bill behavior
    Optional ChatGPT credits permit paid continuation after included usage; this is not a clean subscription wall.

    Temporary / unknown: Temporary limit state: the five-hour gate is suspended, weekly limits still apply, and OpenAI has not stated durable headroom or how long the suspension will last. We infer no throughput from the temporary state.

  2. ↳ IDE-first / visible-pool exception

    Cursor Ultra

    Workflow
    IDE-first agenting when the editor and an inspectable monthly pool belong together.
    Capability
    Cursor’s agent can route across its supported model catalog inside the IDE workflow.
    Current headroom
    $200/month includes a visible $400 API-usage pool; Auto and Composer allowances remain separate.
    Bill behavior
    Premium/model routing drains the included pool; on-demand overage is visible and cappable.
    Primary receiptsCursor pricing and included pool ↗Cursor model and usage limits ↗Cursor overage controls ↗
  3. ↳ Claude-native workflow-affinity exception

    Claude Code Max 20×

    Workflow
    Claude-native terminal work when Claude Code’s interaction model is the reason to buy.
    Capability
    Choose it for Claude/Sonnet affinity rather than a claim of higher published throughput.
    Current headroom
    Max 20× is relative to Pro; exact rolling message/token capacity remains unpublished and weekly limits apply.
    Bill behavior
    Usage stops at the included limit by default; optional usage credits can continue at API rates when enabled.
    Primary receiptsAnthropic Max plan limits ↗Anthropic Claude Code with Pro or Max ↗Anthropic optional usage credits ↗

Decision order is workflow → capability → current headroom → bill behavior. No Codex throughput is inferred while the temporary limit state is in force. Permalink to the current verdict.

Current lived-reliability conclusions

Non-ranked, attribution-limited evidence from Jul 12–18, 2026. No popularity, prevalence, or incident-rate inference.

Claude Code

Strong model and workflow affinity remains visible, but current receipts bound completion trust and usage-control confidence.

Confidence: Moderate · corroborated reports. Evidence limit: Reports are self-selected and do not establish prevalence. The completion report explicitly describes the model as better and cleaner.

Codex

Capability and a concrete switching account are cautiously positive; quota accounting, capacity stops, and Windows app incidents limit completion trust.

Confidence: Moderate · corroborated reports. Evidence limit: No receipt measures prevalence. Windows incidents do not establish CLI or macOS reliability, and one switching account is not market momentum.

Cursor

A supported access-interruption warning exists; current evidence is insufficient for cross-billing or broad quality-decline claims.

Confidence: Limited · isolated or attribution-limited. Evidence limit: The billing attribution is unresolved, no broad quality sample is present, and neither receipt supports prevalence or causation.

How each spend route fails

Stay free / BYOK / local

Fit: Trials, occasional work, local/open-source harnesses, or when you already bring inference.

Bill ceiling: Tool UI can be $0; model/API or local compute may sit outside the subscription.

Failure mode: “Free” shifts cost or capability limits instead of guaranteeing agent headroom.

Cline provider docs ↗Claude Code costs ↗

Bundled subscription seat

Fit: Recurring agentic work where a quota, rolling window, or included pool is acceptable.

Bill ceiling: Plan price covers included use; optional Claude or ChatGPT credits and Cursor overage can continue billing only when enabled.

Failure mode: Work pauses at a remaining limit, or moves onto an explicitly enabled credit/overage meter.

Claude Code with Pro/Max ↗OpenAI Codex pricing ↗Cursor usage limits ↗Cursor pricing ↗

Direct API / BYOK billing

Fit: Spiky workloads, exact token-rate visibility, custom harnesses, or after a seat limit is exhausted.

Bill ceiling: Explicit token/API/credit spend, not an included-seat pool; bill scales with provider token or credit rates.

Failure mode: Every token or credit is spend; runaway agents can keep consuming unless externally capped.

Anthropic API pricing ↗Claude Code costs ↗Cursor models & pricing ↗OpenAI Codex pricing ↗

Audit the decision

This verdict separates workflow, capability, current headroom, and bill behavior. It does not convert Codex’s temporary five-hour-gate suspension into a throughput estimate. Read the Value Index normalization methodology.