Skip to content

feat(openai): GPT-6 models + per-model reasoning effort on the ChatGPT subscription - #954

Merged
ericleepi314 merged 2 commits into
mainfrom
feat/openai-gpt6-subscription
Sep 24, 2026
Merged

ericleepi314 merged 2 commits into
mainfrom
feat/openai-gpt6-subscription

Conversation

@ericleepi314

@ericleepi314 ericleepi314 commented Sep 24, 2026 •

Copy link
Copy Markdown
Collaborator

Summary

  • Subscription login tested live. The browser "Sign in with ChatGPT" flow ran end to end through clawcodex login in an isolated CLAWCODEX_CONFIG_DIR. Credentials were saved, the catalog was discovered, and a Read-tool agent turn succeeded on gpt-6-astra, gpt-6-sol and gpt-6-luna at --effort max.

  • GPT-6 now reaches subscription users. The Codex catalog hides any model whose minimal_client_version is above our pinned client_version. gpt-6-astra needs 0.153.0 and gpt-6-sol/luna need 0.155.0, but we were pinned at 0.149.1. The pin is now 0.155.1.

  • Subscription reasoning effort is clamped per model. Levels come from the login's cached catalog, with a static fallback. Live probes on 2026-09-24:

    • gpt-6 and gpt-5.6 accept up to max.
    • gpt-5.5 accepts up to xhigh.
    • ultra is advertised in the catalog but returns 400 on every model, so it is filtered out.
    • The old blanket rule sent high for every model, which capped GPT-6 two notches low.
  • Public API: max passes through for GPT-6 and still degrades to xhigh for everything else.

  • /model picker: the effort step uses the same per-model ladder and knows whether the session is on the subscription.

  • Model registration:

    • gpt-6-sol and gpt-6-luna added to the OpenAI model list; gpt-6-astra was already there.
    • Context rows set to 872K, the smaller of the 922K API input limit and the 872K subscription limit. OpenAI's overflow error does not trigger reactive compaction, so an over-estimate is a hard failure.
    • List pricing added, with the 272K long-context tier.
    • Verbosity support added.
    • Harness effort docs updated.
  • GPT-5.6 context window fixed too. All five gpt-5.6 rows drop from 1.05M to 872K for the same reason: the API caps input at 922K and the subscription at 872K.

Follow-ups (not in this PR)

  • OpenRouter has no openai/gpt-6-* rows.
  • The advisor_cost context-tier note is out of date.
  • Tests could be isolated from the real config dir suite-wide.

Test plan

  • Targeted suites pass: providers, picker, openai routing/subscription, pricing, effort, compact.
  • Full suite: 1 failure, test_system_prompt_full::...damped_wording.... It fails the same way on origin/main.
  • Live against the ChatGPT backend: catalog at several client_versions, an effort ladder per model, and -p agent turns with a tool call on all three GPT-6 models.
  • A fresh browser login saves gpt-6-astra as the default model, and a flag-less session runs on it.
  • Critic review: REQUEST_CHANGES on the context window, fixed, then APPROVE.

🤖 Generated with Claude Code

…T subscription

- Bump the Codex catalog client_version 0.149.1 -> 0.155.1. It gates the
  catalog: gpt-6-astra needs 0.153.0, gpt-6-sol/luna need 0.155.0, so
  subscription logins never saw GPT-6.
- Clamp subscription reasoning effort per model from the login's cached
  catalog (static fallback). Probed live 2026-09-24: gpt-6/gpt-5.6 take max,
  gpt-5.5 stops at xhigh, 'ultra' 400s everywhere. The old blanket
  xhigh/max -> high clamp capped GPT-6 two notches low.
- Public API: max passes through for GPT-6, still degrades to xhigh elsewhere.
- /model picker effort step uses the same per-model ladder, subscription-aware.
- Register gpt-6-sol/luna, context rows at 872K (smaller of the 922K API and
  872K subscription input limits), list pricing with the 272K long tier,
  verbosity support; harness effort docs updated.

Co-Authored-By: Claude Opus 5.5 (1M context) <noreply@anthropic.com>
@github-actions

github-actions Bot commented Sep 24, 2026 •

Copy link
Copy Markdown

Test Results

     5 files   1 021 suites   19m 33s ⏱️
16 054 tests 16 032 ✅ 22 💤 0 ❌
32 078 runs  32 007 ✅ 71 💤 0 ❌

Results for commit 4bf9762.

♻️ This comment has been updated with latest results.

The public API caps gpt-5.6 input at 922K and the ChatGPT subscription
catalog caps it at 872K (max_context_window). At 1.05M, auto-compact fired
near 998K, past both limits, into an overflow error reactive compaction
does not recognise. Same reasoning as the GPT-6 rows.

Co-Authored-By: Claude Opus 5.5 (1M context) <noreply@anthropic.com>
@ericleepi314
ericleepi314 merged commit ff06eb6 into main Sep 24, 2026
8 checks passed
@ericleepi314
ericleepi314 deleted the feat/openai-gpt6-subscription branch September 24, 2026 23:45
Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

1 participant