feat(openai): GPT-6 models + per-model reasoning effort on the ChatGPT subscription - #954
Merged
Merged
Conversation
…T subscription - Bump the Codex catalog client_version 0.149.1 -> 0.155.1. It gates the catalog: gpt-6-astra needs 0.153.0, gpt-6-sol/luna need 0.155.0, so subscription logins never saw GPT-6. - Clamp subscription reasoning effort per model from the login's cached catalog (static fallback). Probed live 2026-09-24: gpt-6/gpt-5.6 take max, gpt-5.5 stops at xhigh, 'ultra' 400s everywhere. The old blanket xhigh/max -> high clamp capped GPT-6 two notches low. - Public API: max passes through for GPT-6, still degrades to xhigh elsewhere. - /model picker effort step uses the same per-model ladder, subscription-aware. - Register gpt-6-sol/luna, context rows at 872K (smaller of the 922K API and 872K subscription input limits), list pricing with the 272K long tier, verbosity support; harness effort docs updated. Co-Authored-By: Claude Opus 5.5 (1M context) <noreply@anthropic.com>
Test Results 5 files 1 021 suites 19m 33s ⏱️ Results for commit 4bf9762. ♻️ This comment has been updated with latest results. |
The public API caps gpt-5.6 input at 922K and the ChatGPT subscription catalog caps it at 872K (max_context_window). At 1.05M, auto-compact fired near 998K, past both limits, into an overflow error reactive compaction does not recognise. Same reasoning as the GPT-6 rows. Co-Authored-By: Claude Opus 5.5 (1M context) <noreply@anthropic.com>
This file contains hidden or bidirectional Unicode text that may be interpreted or compiled differently than what appears below. To review, open the file in an editor that reveals hidden Unicode characters.
Learn more about bidirectional Unicode characters
Sign up for free
to join this conversation on GitHub.
Already have an account?
Sign in to comment
Add this suggestion to a batch that can be applied as a single commit.This suggestion is invalid because no changes were made to the code.Suggestions cannot be applied while the pull request is closed.Suggestions cannot be applied while viewing a subset of changes.Only one suggestion per line can be applied in a batch.Add this suggestion to a batch that can be applied as a single commit.Applying suggestions on deleted lines is not supported.You must change the existing code in this line in order to create a valid suggestion.Outdated suggestions cannot be applied.This suggestion has been applied or marked resolved.Suggestions cannot be applied from pending reviews.Suggestions cannot be applied on multi-line comments.Suggestions cannot be applied while the pull request is queued to merge.Suggestion cannot be applied right now. Please check back later.
Summary
Subscription login tested live. The browser "Sign in with ChatGPT" flow ran end to end through
clawcodex loginin an isolatedCLAWCODEX_CONFIG_DIR. Credentials were saved, the catalog was discovered, and a Read-tool agent turn succeeded on gpt-6-astra, gpt-6-sol and gpt-6-luna at--effort max.GPT-6 now reaches subscription users. The Codex catalog hides any model whose
minimal_client_versionis above our pinnedclient_version. gpt-6-astra needs 0.153.0 and gpt-6-sol/luna need 0.155.0, but we were pinned at 0.149.1. The pin is now 0.155.1.Subscription reasoning effort is clamped per model. Levels come from the login's cached catalog, with a static fallback. Live probes on 2026-09-24:
max.xhigh.ultrais advertised in the catalog but returns 400 on every model, so it is filtered out.highfor every model, which capped GPT-6 two notches low.Public API:
maxpasses through for GPT-6 and still degrades toxhighfor everything else./modelpicker: the effort step uses the same per-model ladder and knows whether the session is on the subscription.Model registration:
GPT-5.6 context window fixed too. All five gpt-5.6 rows drop from 1.05M to 872K for the same reason: the API caps input at 922K and the subscription at 872K.
Follow-ups (not in this PR)
openai/gpt-6-*rows.advisor_costcontext-tier note is out of date.Test plan
test_system_prompt_full::...damped_wording.... It fails the same way on origin/main.client_versions, an effort ladder per model, and-pagent turns with a tool call on all three GPT-6 models.gpt-6-astraas the default model, and a flag-less session runs on it.🤖 Generated with Claude Code