You signed in with another tab or window. Reload to refresh your session.You signed out in another tab or window. Reload to refresh your session.You switched accounts on another tab or window. Reload to refresh your session.Dismiss alert
The Codex client model catalog exposes input-context metadata such as context_window, but does not expose the model registry output ceiling. Clients that need to reserve output headroom for proactive compaction therefore cannot derive a complete capacity contract from /v1/models.
This metadata must remain distinct from request translation: Codex currently rejects max_output_tokens when it is sent as a request parameter on this path. The catalog may advertise capacity without forwarding that field upstream.
Proposed behavior
Add max_output_tokens to Codex client model entries when the registry has a positive MaxCompletionTokens or OutputTokenLimit.
Apply the field consistently to template-backed and synthesized model entries.
Omit the field when capacity is unknown; do not invent a default.
Do not add or translate any request parameter.
Add focused tests for known, fallback, and unknown output limits.
Why this belongs upstream
The source data already exists in registry.ModelInfo; exposing it in the model catalog provides a provider-neutral, machine-readable capacity signal for any Codex-compatible client without adding client-specific routing or hard-coded model margins.
Privacy
This report is based solely on public code behavior and synthetic request shapes. It contains no credentials, account identifiers, local paths, logs, or conversation content.
This discussion was converted from issue #4607 on August 07, 2026 21:09.
Heading
Bold
Italic
Quote
Code
Link
Numbered list
Unordered list
Task list
Attach files
Mention
Reference
Menu
reacted with thumbs up emoji reacted with thumbs down emoji reacted with laugh emoji reacted with hooray emoji reacted with confused emoji reacted with heart emoji reacted with rocket emoji reacted with eyes emoji
Uh oh!
There was an error while loading. Please reload this page.
Problem
The Codex client model catalog exposes input-context metadata such as
context_window, but does not expose the model registry output ceiling. Clients that need to reserve output headroom for proactive compaction therefore cannot derive a complete capacity contract from/v1/models.This metadata must remain distinct from request translation: Codex currently rejects
max_output_tokenswhen it is sent as a request parameter on this path. The catalog may advertise capacity without forwarding that field upstream.Proposed behavior
max_output_tokensto Codex client model entries when the registry has a positiveMaxCompletionTokensorOutputTokenLimit.Why this belongs upstream
The source data already exists in
registry.ModelInfo; exposing it in the model catalog provides a provider-neutral, machine-readable capacity signal for any Codex-compatible client without adding client-specific routing or hard-coded model margins.Privacy
This report is based solely on public code behavior and synthetic request shapes. It contains no credentials, account identifiers, local paths, logs, or conversation content.
All reactions