Conversation
GPTZZZ (https://gptzzz.ai) is an OpenAI-compatible relay. Adds the provider entry plus the four chat models it currently serves, all via base_model on the existing OpenAI lab entries. Sources: - Endpoint and catalog: https://gptzzz.ai/docs/ and GET /v1/models (2026-09-16) - Effort values verified live against POST /v1/chat/completions: none, low, medium, high, xhigh and max are accepted; minimal is rejected upstream. - Pricing is the published 0.2x group rate applied to OpenAI list pricing, cross-checked against the cost/actual_cost fields from GET /v1/usage. Co-Authored-By: Claude Opus 5 <noreply@anthropic.com>
Action items
|
Review flagged that the gpt-6-astra effort list was copied from GPT-5.6 rather than verified on that model. Re-tested each model separately against the live endpoint. gpt-6-astra rejects none upstream: Unsupported value: 'none' is not supported with the 'gpt-6-astra' model. Supported values are: 'low', 'medium', 'high', 'xhigh', and 'max'. This now matches providers/openai/models/gpt-6-astra.toml and the other same-surface peers. gpt-5.6, gpt-5.6-sol and gpt-5.6-terra do accept none (HTTP 200), so their lists are unchanged; their header comments now say the values were verified per model. Co-Authored-By: Claude Opus 5 <noreply@anthropic.com>
|
Good catch, and correct. The I re-tested each model separately against the live endpoint. So
Effort values as retested on 2026-09-16:
|
|
No actionable findings. |
Re-verified all effort values against the live endpoint on 2026-09-28. gpt-6-astra rejected 'none' on 2026-09-16, which is what this branch recorded. The upstream has since added it: rejecting an unsupported value now reports the supported set as Unsupported value: 'minimal' is not supported with the 'gpt-6-astra' model. Supported values are: 'none', 'low', 'medium', 'high', 'xhigh'. Each value was re-tested per model rather than copied across models. 'minimal' remains rejected everywhere. Co-Authored-By: Claude Opus 5 <noreply@anthropic.com>
|
Correcting an earlier entry in this PR: When I first verified this on 2026-09-16 the host rejected it, and that is what the previous commit recorded: Re-verifying today (2026-09-28), the same probe reports a different supported set: So the upstream added the value at some point in the last twelve days. I re-tested every value against every model individually rather than copying a list across models, per the earlier review:
Flagging this because it is a property of the provider that can change without notice — the TOML comment now carries both the re-verification date and the exact rejection string it is based on, so the next person can tell when the record was last checked against reality. |
Action items
|
Prices = list price shown on https://gptzzz.ai/model-plaza x 0.05, read from /api/v1/model-plaza on 2026-09-29. List prices for gpt-5.6, gpt-5.6-sol, gpt-5.6-terra and gpt-6-astra were refreshed from the same source. Co-Authored-By: Claude Opus 5.5 <noreply@anthropic.com>
The host's own rejection message on 2026-09-28 lists 'none', 'low', 'medium', 'high', 'xhigh' for gpt-6-astra, so drop 'max'. Co-Authored-By: Claude Opus 5.5 <noreply@anthropic.com>
|
Two follow-ups:
|
Action items
|
gpt-5.6 / gpt-5.6-sol are billed on the host's own list price (5/30 per MTok), which is above OpenAI's; document it in the file headers. Remove gpt-6-astra from this PR; it will come back in a follow-up with per-value probes. Co-Authored-By: Claude Opus 5.5 <noreply@anthropic.com>
|
Correction to my previous comment: I wrote that OpenAI list prices had changed. That was wrong. OpenAI's list for Addressed the two [high] items:
|
… billing evidence Per-value probes on 2026-09-29: gpt-6-astra accepts low..max and rejects none, matching the OpenAI lab entry. gpt-5.6* accept none..max. /v1/usage cost matches the stated list prices and actual_cost/cost = 0.05. Co-Authored-By: Claude Opus 5.5 <noreply@anthropic.com>
|
Re-probed everything with a live key on 2026-09-29 and updated the description with the full table:
|
|
No actionable findings. |
|
Gentle ping — this has been open since 2026-09-16 and is currently the blocker for a downstream PR. The Cline maintainer @abeatrix asked on cline/cline#14182:
So this PR is the gating item for that integration. Current state here:
Happy to rebase, split it per-model, or adjust anything that would help it land. Thanks for maintaining this. |
GPTZZZ is an independent OpenAI-compatible relay serving OpenAI models (not affiliated with OpenAI). This adds the provider entry plus four chat models it currently serves, all through
base_modelon the existing OpenAI lab entries.Sources
Endpoint and catalog:
https://gptzzz.ai/v1, documented at https://gptzzz.ai/docs/. All four models are returned byGET /v1/models(2026-09-29).Reasoning options: probed one value at a time against
POST /v1/chat/completionson 2026-09-29 (prompt "What is 17*23?",max_completion_tokens: 256):gpt-6-astratherefore uses the lab set["low","medium","high","xhigh","max"]. The 400 bodies read e.g.Unsupported value: 'none' is not supported with the 'gpt-6-astra' model. Supported values are: 'low', 'medium', 'high', 'xhigh', and 'max'.Pricing: read from the host's public
/api/v1/model-plazaon 2026-09-29. The host bills its own published list price multiplied by a public group rate; these entries use the 0.05x group. Each model file states the list price it is derived from. Cross-check against the host'sGET /v1/usage: thecostfield reproduces exactly from those list prices (e.g.gpt-5.6at 5.00 / 30.00,gpt-6-astraat 10.00 / 50.00), andactual_cost / cost = 0.05on 2026-09-25, 09-28 and 09-29. Forgpt-5.6/gpt-5.6-solthe host's list (5.00 / 30.00) is above OpenAI's 4.00 / 20.00, so the values here are what a customer is actually charged. All values are USD per million tokens.Notes
gpt-5.6usesbase_model = "openai/gpt-5.6-sol"withname = "GPT-5.6": the host serves it as an alias ofgpt-5.6-sol(its 400 message forgpt-5.6namesgpt-5.6-sol).bun validatepasses.