Skip to content

chore: update model catalog from bot issues - #822

Merged
Erin McNulty (erin2722) merged 3 commits into
mainfrom
chore/autofix-bot-issues-2026-06-18
Jun 18, 2026
Merged

chore: update model catalog from bot issues#822
Erin McNulty (erin2722) merged 3 commits into
mainfrom
chore/autofix-bot-issues-2026-06-18

Conversation

@github-actions

Copy link
Copy Markdown
Contributor

Automated daily batch of model catalog updates from bot issues.

Included issues

Summary

Issue Provider Primary model Changed models Added models Updated models Verification sources
#805 bedrock xai.grok-4.3 xai.grok-4.3 xai.grok-4.3 None 1
2
3
4
#806 together Qwen/Qwen3.7-Plus Qwen/Qwen3.7-Plus Qwen/Qwen3.7-Plus None 1
2
3
#807 cerebras llama3.3-70b llama3.3-70b
llama3.1-8b
llama-4-scout-17b-16e-instruct
None llama3.3-70b
llama3.1-8b
llama-4-scout-17b-16e-instruct
1
2
#808 groq moonshotai/kimi-k2-instruct-0905 moonshotai/kimi-k2-instruct-0905
meta-llama/llama-4-maverick-17b-128e-instruct
gemma2-9b-it
None moonshotai/kimi-k2-instruct-0905
meta-llama/llama-4-maverick-17b-128e-instruct
gemma2-9b-it
1
2
#810 bedrock openai.gpt-5.5 openai.gpt-5.5
openai.gpt-5.4
openai.gpt-5.5
openai.gpt-5.4
None 1
2
3
4
#814 anthropic claude-mythos-5 claude-mythos-5 claude-mythos-5 None 1
2
3
4
#815 together deepseek-ai/DeepSeek-V4-Pro deepseek-ai/DeepSeek-V4-Pro
moonshotai/Kimi-K2.6
None deepseek-ai/DeepSeek-V4-Pro
moonshotai/Kimi-K2.6
1
2
3
#816 mistral mistral-large-2411 mistral-large-2411
pixtral-large-2411
open-mistral-nemo-2407
None mistral-large-2411
pixtral-large-2411
open-mistral-nemo-2407
1
#819 together zai-org/GLM-5.2 zai-org/GLM-5.2 zai-org/GLM-5.2 None 1
2
#820 mistral voxtral-small-2507 voxtral-small-2507 voxtral-small-2507 None 1
2

Verified metadata

#805: [BOT ISSUE] Bedrock: add missing xAI Grok 4.3 model (xai.grok-4.3)

Model Display name Parent Providers Format Flavor Token limits Pricing Lifecycle
xai.grok-4.3 Grok 4.3 bedrock openai chat input=1000000, output=1000000 in/out=1.25/2.5 per 1M; cache read=0.2 per 1M multimodal=true; reasoning=true

Verification notes

No LLM verification step ran; model metadata was already complete in the issue.

sync_models vs proposed update

sync_models cross-check found differences. Official provider verification was used for the applied values, and sync_models discrepancies are listed below for review.

Model Field Proposed update sync_models sync_models source models
xai.grok-4.3 catalog entry present missing None

#806: [BOT ISSUE] Together: add missing Qwen3.7-Plus model (Qwen/Qwen3.7-Plus)

Model Display name Parent Providers Format Flavor Token limits Pricing Lifecycle
Qwen/Qwen3.7-Plus Qwen3.7 Plus together openai chat input=1000000, output=65536 in/out=0.32/1.28 per 1M multimodal=true

Verification notes

No LLM verification step ran; model metadata was already complete in the issue.

sync_models vs proposed update

sync_models cross-check found differences. Official provider verification was used for the applied values, and sync_models discrepancies are listed below for review.

Model Field Proposed update sync_models sync_models source models
Qwen/Qwen3.7-Plus catalog entry present missing None

#807: [BOT ISSUE] Cerebras: remove deprecated models (llama3.3-70b, llama3.1-8b, llama-4-scout-17b-16e-instruct)

Model Display name Parent Providers Format Flavor Token limits Pricing Lifecycle
llama3.3-70b Llama 3.3 70B n/a openai chat input=n/a, output=not provided in/out=0.1/0.1 per 1M deprecated=true; date=2026-02-16
llama3.1-8b Llama 3.1 8B n/a openai chat input=128000, output=128000 in/out=0.1/0.1 per 1M deprecated=true; date=2026-05-27
llama-4-scout-17b-16e-instruct LLaMA 4 Scout 17B 16e Instruct n/a openai chat input=131072, output=8192 in/out=0.11/0.34 per 1M deprecated=true; date=2025-11-03

Verification notes

Verification Checklist

sync_models vs proposed update

sync_models cross-check found differences. Official provider verification was used for the applied values, and sync_models discrepancies are listed below for review.

Model Field Proposed update sync_models sync_models source models
llama3.3-70b catalog entry present missing None
llama3.1-8b deprecation_date 2026-05-27 n/a cerebras/llama3.1-8b
llama-4-scout-17b-16e-instruct catalog entry present missing None

#808: [BOT ISSUE] Groq: remove deprecated models (moonshotai/kimi-k2-instruct-0905, meta-llama/llama-4-maverick-17b-128e-instruct, gemma2-9b-it)

Model Display name Parent Providers Format Flavor Token limits Pricing Lifecycle
moonshotai/kimi-k2-instruct-0905 Kimi K2 0905 groq openai chat input=262144, output=16384 in/out=1/3 per 1M; cache read=0.5 per 1M deprecated=true; date=2026-04-15
meta-llama/llama-4-maverick-17b-128e-instruct Llama 4 Maverick (17Bx128E) groq openai chat input=131072, output=8192 in/out=0.2/0.6 per 1M deprecated=true; date=2026-03-09; multimodal=true
gemma2-9b-it Gemma 2 9B n/a openai chat input=n/a, output=not provided in/out=0.2/0.2 per 1M deprecated=true; date=2025-10-08

Verification notes

Verification Checklist

sync_models vs proposed update

sync_models cross-check found differences. Official provider verification was used for the applied values, and sync_models discrepancies are listed below for review.

Model Field Proposed update sync_models sync_models source models
moonshotai/kimi-k2-instruct-0905 deprecation_date 2026-04-15 n/a groq/moonshotai/kimi-k2-instruct-0905
meta-llama/llama-4-maverick-17b-128e-instruct deprecation_date 2026-03-09 n/a groq/meta-llama/llama-4-maverick-17b-128e-instruct
gemma2-9b-it catalog entry present missing None

#810: [BOT ISSUE] Missing AWS Bedrock models: openai.gpt-5.5, openai.gpt-5.4

Model Display name Parent Providers Format Flavor Token limits Pricing Lifecycle
openai.gpt-5.5 OpenAI GPT-5.5 (Bedrock) bedrock openai chat input=272000, output=not provided in/out=5.5/33 per 1M; cache read=0.55 per 1M multimodal=true; reasoning=true
openai.gpt-5.4 OpenAI GPT-5.4 (Bedrock) bedrock openai chat input=272000, output=not provided in/out=2.75/16.5 per 1M; cache read=0.275 per 1M multimodal=true; reasoning=true

Verification notes

No LLM verification step ran; model metadata was already complete in the issue.

sync_models vs proposed update

sync_models cross-check found differences. Official provider verification was used for the applied values, and sync_models discrepancies are listed below for review.

Model Field Proposed update sync_models sync_models source models
openai.gpt-5.5 catalog entry present missing None
openai.gpt-5.4 catalog entry present missing None

#814: [BOT ISSUE] Anthropic: add missing claude-mythos-5 model

Model Display name Parent Providers Format Flavor Token limits Pricing Lifecycle
claude-mythos-5 Claude Mythos 5 anthropic anthropic chat input=1000000, output=128000 in/out=10/50 per 1M; cache read=1 per 1M; cache write=12.5 per 1M multimodal=true; reasoning=true

Verification notes

No LLM verification step ran; model metadata was already complete in the issue.

sync_models vs proposed update

sync_models cross-check found differences. Official provider verification was used for the applied values, and sync_models discrepancies are listed below for review.

Model Field Proposed update sync_models sync_models source models
claude-mythos-5 catalog entry present missing None

#815: [BOT ISSUE] Together: stale pricing for deepseek-ai/DeepSeek-V4-Pro and missing pricing for moonshotai/Kimi-K2.6

Model Display name Parent Providers Format Flavor Token limits Pricing Lifecycle
deepseek-ai/DeepSeek-V4-Pro DeepSeek V4 Pro together openai chat input=512000, output=not provided in/out=1.74/3.48 per 1M; cache read=0.2 per 1M active
moonshotai/Kimi-K2.6 Kimi K2.6 baseten, together openai chat input=262144, output=not provided in/out=1.2/4.5 per 1M; cache read=0.2 per 1M multimodal=true

Verification notes

Verification Checklist

sync_models vs proposed update

sync_models cross-check found differences. Official provider verification was used for the applied values, and sync_models discrepancies are listed below for review.

Model Field Proposed update sync_models sync_models source models
deepseek-ai/DeepSeek-V4-Pro catalog entry present missing None
moonshotai/Kimi-K2.6 catalog entry present missing None

#816: [BOT ISSUE] Mistral: mark deprecated mistral-large-2411, pixtral-large-2411, open-mistral-nemo-2407

Model Display name Parent Providers Format Flavor Token limits Pricing Lifecycle
mistral-large-2411 mistral-large-latest mistral openai chat input=128000, output=128000 in/out=2/6 per 1M parent=mistral-large-latest; deprecated=true; date=2026-02-27
pixtral-large-2411 pixtral-large-latest mistral openai chat input=128000, output=128000 in/out=2/6 per 1M parent=pixtral-large-latest; deprecated=true; date=2026-02-27; multimodal=true
open-mistral-nemo-2407 mistral openai chat input=128000, output=128000 in/out=0.15/0.15 per 1M deprecated=true; date=2026-05-22

Verification notes

Verification Checklist

sync_models vs proposed update

sync_models cross-check found differences. Official provider verification was used for the applied values, and sync_models discrepancies are listed below for review.

Model Field Proposed update sync_models sync_models source models
mistral-large-2411 deprecation_date 2026-02-27 n/a mistral/mistral-large-2411
pixtral-large-2411 deprecation_date 2026-02-27 n/a mistral/pixtral-large-2411
open-mistral-nemo-2407 input_cost_per_mil_tokens 0.15 0.3 mistral/open-mistral-nemo-2407
open-mistral-nemo-2407 output_cost_per_mil_tokens 0.15 0.3 mistral/open-mistral-nemo-2407
open-mistral-nemo-2407 deprecation_date 2026-05-22 n/a mistral/open-mistral-nemo-2407

#819: [BOT ISSUE] Together: add missing zai-org/GLM-5.2 model

Model Display name Parent Providers Format Flavor Token limits Pricing Lifecycle
zai-org/GLM-5.2 together openai chat input=262144, output=not provided in/out=1.4/4.4 per 1M; cache read=0.26 per 1M active

Verification notes

No LLM verification step ran; model metadata was already complete in the issue.

sync_models vs proposed update

sync_models cross-check found differences. Official provider verification was used for the applied values, and sync_models discrepancies are listed below for review.

Model Field Proposed update sync_models sync_models source models
zai-org/GLM-5.2 catalog entry present missing None

#820: [BOT ISSUE] Mistral: add missing voxtral-small-2507 model

Model Display name Parent Providers Format Flavor Token limits Pricing Lifecycle
voxtral-small-2507 mistral openai chat input=32000, output=not provided in/out=0.1/0.4 per 1M multimodal=true

Verification notes

No LLM verification step ran; model metadata was already complete in the issue.

sync_models vs proposed update

sync_models cross-check found differences. Official provider verification was used for the applied values, and sync_models discrepancies are listed below for review.

Model Field Proposed update sync_models sync_models source models
voxtral-small-2507 catalog entry present missing None

@vercel

vercel Bot commented Jun 18, 2026

Copy link
Copy Markdown

The latest updates on your projects. Learn more about Vercel for GitHub.

Project Deployment Actions Updated (UTC)
ai-proxy Ready Ready Preview, Comment Jun 18, 2026 3:59pm

Request Review

@github-actions

Copy link
Copy Markdown
Contributor Author

Codex (@codex) review

@chatgpt-codex-connector chatgpt-codex-connector Bot left a comment

Copy link
Copy Markdown

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

💡 Codex Review

Here are some automated review suggestions for this pull request.

Reviewed commit: b0040c81da

ℹ️ About Codex in GitHub

Codex has been enabled to automatically review pull requests in this repo. Reviews are triggered when you

  • Open a pull request for review
  • Mark a draft as ready
  • Comment "Codex (@codex) review".

If Codex has suggestions, it will comment; otherwise it will react with 👍.

When you sign up for Codex through ChatGPT, Codex can also answer questions or update the PR, like "Codex (@codex) address that feedback".

Comment thread packages/proxy/schema/index.ts Outdated
Comment on lines +596 to +599
"openai.gpt-5.5": ["bedrock"],
"openai.gpt-5.4": ["bedrock"],
"Qwen/Qwen3.7-Plus": ["together"],
"xai.grok-4.3": ["bedrock"],

Copy link
Copy Markdown

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

P2 Badge Use a supported format for Bedrock-only models

These new Bedrock entries are backed by model_list.json specs with format: "openai", but this proxy dispatches by modelSpec.format; the OpenAI path immediately rejects Bedrock secrets with Bedrock does not support OpenAI format in fetchOpenAI. As a result, requests for the newly added Bedrock-only models such as openai.gpt-5.5, openai.gpt-5.4, and xai.grok-4.3 will fail locally before reaching AWS unless they use a Bedrock-supported format/path, such as converse, or the proxy adds an OpenAI-compatible Bedrock handler.

Useful? React with 👍 / 👎.

@github-actions
github-actions Bot force-pushed the chore/autofix-bot-issues-2026-06-18 branch from b0040c8 to b678b11 Compare June 18, 2026 13:49
@github-actions

Copy link
Copy Markdown
Contributor Author

Codex Review Response

Issue: Bedrock-only models using unsupported format: "openai"

Valid finding -- fixed.

The Codex review correctly identified that three newly added Bedrock-only models had format: "openai", which would cause the proxy to reject requests at runtime (Bedrock does not support OpenAI format error in proxy.ts:2131). All other Bedrock-only models in the catalog use format: "converse".

Changes made to packages/proxy/schema/model_list.json

Changed format from "openai" to "converse" for:

  • openai.gpt-5.5 (line 12408)
  • openai.gpt-5.4 (line 12422)
  • xai.grok-4.3 (line 13055)

No changes were needed in packages/proxy/schema/index.ts.

Erin McNulty (erin2722) and others added 2 commits June 18, 2026 13:57
The zai-org/GLM-* ids are shared across baseten and together, but the
audit attributes a newly-found shared id to a single provider (it filed
GLM-5.2 as together-only in #819) and the per-provider dedup drops the
other provider. Result: zai-org/GLM-5.1 was together-only despite baseten
serving it, and zai-org/GLM-5.2 was missing entirely.

Verified by direct provider calls (real 200 completions):
- zai-org/GLM-5.2: works on BOTH baseten and together
- zai-org/GLM-5.1: works on baseten (was missing from its providers)

Fixes:
- zai-org/GLM-5.1 available_providers -> ["baseten","together"] (model_list
  + index.ts; the normalize backstop only adds missing mappings, it does
  not widen an existing one, so index.ts was updated by hand).
- Add zai-org/GLM-5.2 ["baseten","together"] (pricing 1.4/4.4 corroborated
  by together #819 + the fireworks glm-5p2 entry; 262k context per #819).
- Also added the missing index.ts mapping for voxtral-small-2507 (another
  model this PR added without one) via normalize-local-models.

Co-Authored-By: Claude Opus 4.8 <noreply@anthropic.com>
xai.grok-4.3, openai.gpt-5.5, openai.gpt-5.4 are served on Amazon Bedrock's
"Mantle" engine via the OpenAI Chat Completions/Responses API on the
bedrock-mantle endpoint — NOT Converse (the AWS Grok 4.3 model card lists
Converse as unsupported). The gateway supports the /chat/completions path
on Bedrock, so set format "converse" -> "openai" for all three. Verified:
the gateway now routes xai.grok-4.3 to Bedrock (AWS_DEFAULT_CREDENTIALS,
format=openai); a clean 200 isn't reachable from the CI account because
grok-4.3 is us-west-2-only and the CI Bedrock creds are us-east-1 (env,
not a catalog defect).

Also fix grok-4.3 max_output_tokens 1000000 -> 131072 (1M is the context
window; the model card's max_completion_tokens default is 131072).

Co-Authored-By: Claude Opus 4.8 <noreply@anthropic.com>
@erin2722
Erin McNulty (erin2722) merged commit ac94547 into main Jun 18, 2026
6 of 8 checks passed
Erin McNulty (erin2722) added a commit that referenced this pull request Jun 22, 2026
…rmat to converse

Both model-sync workflows have a 'Respond to Codex review with Claude Code'
step that auto-applies Codex's suggestions to the catalog. When the index.ts
Bedrock mappings for openai.gpt-5.5 / openai.gpt-5.4 / xai.grok-4.3 are present,
Codex posts a P1 ('openai format on a Bedrock model fails') because the TS
proxy's fetchOpenAI rejects Bedrock secrets, and the auto-apply step 'fixes' it
by flipping format openai->converse. That is wrong: these are Bedrock Mantle
models served only via the OpenAI-compatible bedrock-mantle endpoint, which does
not support the Converse API or InvokeModel, so converse breaks invocation at
AWS. This recurred across #834/#836/#840/#843 (all reverting the validated #822
openai fix). The LiteLLM field sync never touches format (no checkAndUpdateFormat),
so this guardrail on the codex-response prompt is the correct lever.

Adds a 'Bedrock Mantle exception' instruction to the codex-response prompt in
both sync-models.yaml and fix-missing-model-bot-issues.yaml: keep format=openai
for bedrock-only openai-format models and explain in the summary that this is an
intentional Mantle case to be handled proxy-side, not a catalog change.

Co-Authored-By: Claude Opus 4.8 <noreply@anthropic.com>
Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment