Repository navigation
Conversation
Keep glm-*[1m] in the listing catalog, but send the base modelCode upstream so Chat/Responses/Claude/async no longer hit 1214.
Coding Plan 1M context is opt-in via glm-*[1m]. Bare ids inherit the 200k catalog window; listing still keeps the [1m] aliases.
Owner
|
Hey I checked this commit and found a few issues: Blocking issues:
|
Owner
|
After careful review this PR will be closed without merging as it's considered NOT A BUG Breaking down the PR it's obvious that 3 issues are adressed, respectivly the 1M marker, disable thinking and proxy side reasoning effort map. Let's start with the effort map.
|
xz-dev
added a commit
to xz-dev/zcode-api
that referenced
this pull request
Aug 24, 2026
Fork-only overlay. Upstream closed PR TriDefender#30. Includes f76599e max_tokens catalog + request clamp so Z.AI 1210 does not fire on 200k/1M context.
This file contains hidden or bidirectional Unicode text that may be interpreted or compiled differently than what appears below. To review, open the file in an editor that reveals hidden Unicode characters.
Learn more about bidirectional Unicode characters
Sign up for free
to join this conversation on GitHub.
Already have an account?
Sign in to comment
Add this suggestion to a batch that can be applied as a single commit.This suggestion is invalid because no changes were made to the code.Suggestions cannot be applied while the pull request is closed.Suggestions cannot be applied while viewing a subset of changes.Only one suggestion per line can be applied in a batch.Add this suggestion to a batch that can be applied as a single commit.Applying suggestions on deleted lines is not supported.You must change the existing code in this line in order to create a valid suggestion.Outdated suggestions cannot be applied.This suggestion has been applied or marked resolved.Suggestions cannot be applied from pending reviews.Suggestions cannot be applied on multi-line comments.Suggestions cannot be applied while the pull request is queued to merge.Suggestion cannot be applied right now. Please check back later.
What
Aligns the proxy with current GLM Coding Plan contracts for reasoning effort and the official
[1m]listing aliases.none/minimal/low → low,medium/high → high,xhigh/max → max; GLM-5.2 can disable). Shared translator covers Chat, Responses, Claude Messages, and async. Catalog is the single source of truth: Codex?client_versionand Anthropicanthropic-versionadvertise the official efforts.[1m]request path — catalogs keepglm-5.2[1m]/glm-5.3[1m]. SharedtransformRequestBodystrips trailing[1m]before upstream so Chat/Responses/Claude/async stop hitting1214/modelCode不存在.glm-5.2/glm-5.3advertise 200k. Only[1m]aliases advertise 1M. Official Coding Plan 1M is opt-in via the suffix (Enable 1M Context, ZCode Connect Models).Why
noneeffort previously collapsed to Anthropicthinking:{type:"enabled"}, solowandmaxlooked identical upstream.[1m]aliases were forwarded asmodelCode, which the provider rejects (1214).glm-5.3call were 1M.Tests
Focused suites for the translator, body transformer, catalogs, and providers pass. Full
bun test: 550 pass / 1 fail. The failure is pre-existing (src/integration.test.tsstill loadsconfig.test.yaml, removed inf6aa147) and is unrelated to this branch.Notes
plan/asyncbehavior change.