Skip to content

chore: update model catalog from bot issues - #1026

Merged
Erin McNulty (erin2722) merged 5 commits into
mainfrom
chore/autofix-bot-issues-2026-07-27
Jul 27, 2026
Merged

chore: update model catalog from bot issues#1026
Erin McNulty (erin2722) merged 5 commits into
mainfrom
chore/autofix-bot-issues-2026-07-27

Conversation

@github-actions

Copy link
Copy Markdown
Contributor

Automated daily batch of model catalog updates from bot issues.

Included issues

Summary

Issue Provider Primary model Changed models Added models Updated models Verification sources
#975 databricks databricks-gpt-5-5-pro databricks-gpt-5-4-mini databricks-gpt-5-4-mini None 1
2
#980 vertex publishers/meta/models/llama-4-maverick-17b-128e-instruct-maas publishers/meta/models/llama-4-maverick-17b-128e-instruct-maas
publishers/meta/models/llama-4-scout-17b-16e-instruct-maas
publishers/meta/models/llama-4-maverick-17b-128e-instruct-maas
publishers/meta/models/llama-4-scout-17b-16e-instruct-maas
None 1
2
3
4
#990 xai grok-4.5 grok-4.5
grok-4.5-latest
None grok-4.5
grok-4.5-latest
1
2
#991 openai gpt-realtime-2.1 gpt-realtime-2.1
gpt-realtime-2.1-mini
gpt-realtime-2.1
gpt-realtime-2.1-mini
None 1
2
3
4
#994 vertex publishers/mistralai/models/mistral-medium-3 publishers/mistralai/models/mistral-medium-3
publishers/mistralai/models/mistral-small-2503
publishers/mistralai/models/codestral-2
publishers/mistralai/models/mistral-medium-3
publishers/mistralai/models/mistral-small-2503
publishers/mistralai/models/codestral-2
None 1
2
3
4
5
#996 vertex publishers/openai/models/gpt-oss-120b-maas publishers/openai/models/gpt-oss-120b-maas
publishers/openai/models/gpt-oss-20b-maas
publishers/openai/models/gpt-oss-120b-maas
publishers/openai/models/gpt-oss-20b-maas
None 1
2
3
4
#999 together thinkingmachines/inkling thinkingmachines/inkling None thinkingmachines/inkling 1
2
#1001 together openai/gpt-oss-20b openai/gpt-oss-20b None openai/gpt-oss-20b 1
2
3
#1011 databricks databricks-gemini-3-6-flash databricks-gemini-3-6-flash databricks-gemini-3-6-flash None 1
2
#1025 fireworks accounts/fireworks/models/kimi-k3 accounts/fireworks/models/kimi-k3 accounts/fireworks/models/kimi-k3 None 1
2
3

Verified metadata

#975: [BOT ISSUE] Databricks: add missing databricks-gpt-5-5-pro, databricks-gpt-5-4-mini, databricks-gemini-3-5-flash

Model Display name Parent Providers Format Flavor Token limits Pricing Lifecycle
databricks-gpt-5-4-mini GPT-5.4 mini databricks openai chat input=272000, output=128000 n/a multimodal=true; reasoning=true

Verification notes

No LLM verification step ran; model metadata was already complete in the issue.

sync_models vs proposed update

sync_models cross-check found differences. Official provider verification was used for the applied values, and sync_models discrepancies are listed below for review.

Model Field Proposed update sync_models sync_models source models
databricks-gpt-5-4-mini catalog entry present missing None

#980: [BOT ISSUE] Vertex: add missing Meta Llama 4 Maverick and Scout MaaS entries

Model Display name Parent Providers Format Flavor Token limits Pricing Lifecycle
publishers/meta/models/llama-4-maverick-17b-128e-instruct-maas Llama 4 Maverick (17Bx128E) Instruct vertex openai chat input=524288, output=8192 in/out=0.5/1.5 per 1M multimodal=true
publishers/meta/models/llama-4-scout-17b-16e-instruct-maas Llama 4 Scout (17Bx16E) Instruct vertex openai chat input=1310720, output=8192 in/out=0.2/0.6 per 1M multimodal=true

Verification notes

Verification

sync_models vs proposed update

sync_models cross-check found differences. Official provider verification was used for the applied values, and sync_models discrepancies are listed below for review.

Model Field Proposed update sync_models sync_models source models
publishers/meta/models/llama-4-maverick-17b-128e-instruct-maas catalog entry present missing None
publishers/meta/models/llama-4-scout-17b-16e-instruct-maas catalog entry present missing None

#990: [BOT ISSUE] xAI: fix stale cached input pricing for grok-4.5 and grok-4.5-latest ($0.50 → $0.30)

Model Display name Parent Providers Format Flavor Token limits Pricing Lifecycle
grok-4.5 Grok 4.5 xAI openai chat input=500000, output=500000 in/out=2/6 per 1M; cache read=0.3 per 1M multimodal=true; reasoning=true
grok-4.5-latest Grok 4.5 (Latest) grok-4.5 xAI openai chat input=500000, output=500000 in/out=2/6 per 1M; cache read=0.3 per 1M parent=grok-4.5; multimodal=true; reasoning=true

Verification notes

No LLM verification step ran; model metadata was already complete in the issue.

sync_models vs proposed update

sync_models cross-check found differences. Official provider verification was used for the applied values, and sync_models discrepancies are listed below for review.

Model Field Proposed update sync_models sync_models source models
grok-4.5 input_cache_read_cost_per_mil_tokens 0.3 0.5 xai/grok-4.5
grok-4.5-latest input_cache_read_cost_per_mil_tokens 0.3 0.5 xai/grok-4.5-latest

#991: [BOT ISSUE] OpenAI: add missing gpt-realtime-2.1 and gpt-realtime-2.1-mini

Model Display name Parent Providers Format Flavor Token limits Pricing Lifecycle
gpt-realtime-2.1 GPT Realtime 2.1 openai, azure openai chat input=128000, output=32000 in/out=4/24 per 1M; cache read=0.4 per 1M reasoning=true
gpt-realtime-2.1-mini GPT Realtime 2.1 Mini openai, azure openai chat input=128000, output=32000 in/out=0.6/2.4 per 1M; cache read=0.06 per 1M reasoning=true

Verification notes

Verification

sync_models vs proposed update

sync_models cross-check found differences. Official provider verification was used for the applied values, and sync_models discrepancies are listed below for review.

Model Field Proposed update sync_models sync_models source models
gpt-realtime-2.1-mini max_output_tokens 32000 4096 gpt-realtime-2.1-mini

#994: [BOT ISSUE] Vertex: add missing Mistral MaaS entries (mistral-medium-3, mistral-small-2503, codestral-2)

Model Display name Parent Providers Format Flavor Token limits Pricing Lifecycle
publishers/mistralai/models/mistral-medium-3 Mistral Medium 3 (Vertex) vertex openai chat input=131072, output=8191 in/out=0.4/2 per 1M active
publishers/mistralai/models/mistral-small-2503 Mistral Small 3.1 (Vertex) vertex openai chat input=128000, output=not provided in/out=0.1/0.3 per 1M active
publishers/mistralai/models/codestral-2 Codestral 2 (Vertex) vertex openai chat input=256000, output=not provided in/out=0.3/0.9 per 1M active

Verification notes

Verification

sync_models vs proposed update

sync_models cross-check found differences. Official provider verification was used for the applied values, and sync_models discrepancies are listed below for review.

Model Field Proposed update sync_models sync_models source models
publishers/mistralai/models/mistral-medium-3 catalog entry present missing None
publishers/mistralai/models/mistral-small-2503 catalog entry present missing None
publishers/mistralai/models/codestral-2 catalog entry present missing None

#996: [BOT ISSUE] Vertex: add missing OpenAI GPT-OSS MaaS entries (gpt-oss-120b-maas, gpt-oss-20b-maas)

Model Display name Parent Providers Format Flavor Token limits Pricing Lifecycle
publishers/openai/models/gpt-oss-120b-maas GPT-OSS 120B (Vertex) vertex openai chat input=131072, output=32768 n/a reasoning=true
publishers/openai/models/gpt-oss-20b-maas GPT-OSS 20B (Vertex) vertex openai chat input=131072, output=32768 n/a reasoning=true

Verification notes

Verification

sync_models vs proposed update

sync_models cross-check found differences. Official provider verification was used for the applied values, and sync_models discrepancies are listed below for review.

Model Field Proposed update sync_models sync_models source models
publishers/openai/models/gpt-oss-120b-maas catalog entry present missing None
publishers/openai/models/gpt-oss-20b-maas catalog entry present missing None

#999: [BOT ISSUE] Together: add together to available_providers for thinkingmachines/inkling

Model Display name Parent Providers Format Flavor Token limits Pricing Lifecycle
thinkingmachines/inkling Inkling baseten, together openai chat input=1048576, output=not provided in/out=1/4.05 per 1M; cache read=0.17 per 1M multimodal=true; reasoning=true

Verification notes

Verification

sync_models vs proposed update

sync_models cross-check found differences. Official provider verification was used for the applied values, and sync_models discrepancies are listed below for review.

Model Field Proposed update sync_models sync_models source models
thinkingmachines/inkling catalog entry present missing None

#1001: [BOT ISSUE] Together: stale pricing for openai/gpt-oss-20b ($0.075/$0.30 → $0.05/$0.20)

Model Display name Parent Providers Format Flavor Token limits Pricing Lifecycle
openai/gpt-oss-20b OpenAI GPT-OSS (20B) groq, together openai chat input=131072, output=65536 in/out=0.05/0.2 per 1M; cache read=0.0375 per 1M reasoning=true

Verification notes

No LLM verification step ran; model metadata was already complete in the issue.

sync_models vs proposed update

sync_models cross-check found differences. Official provider verification was used for the applied values, and sync_models discrepancies are listed below for review.

Model Field Proposed update sync_models sync_models source models
openai/gpt-oss-20b max_output_tokens 65536 32768 groq/openai/gpt-oss-20b, together_ai/openai/gpt-oss-20b
openai/gpt-oss-20b input_cost_per_mil_tokens 0.05 0.075 groq/openai/gpt-oss-20b, together_ai/openai/gpt-oss-20b
openai/gpt-oss-20b output_cost_per_mil_tokens 0.2 0.3 groq/openai/gpt-oss-20b, together_ai/openai/gpt-oss-20b

#1011: [BOT ISSUE] Databricks: add missing databricks-gemini-3-6-flash, databricks-glm-5-2, and databricks-inkling

Model Display name Parent Providers Format Flavor Token limits Pricing Lifecycle
databricks-gemini-3-6-flash Gemini 3.6 Flash databricks openai chat input=1048576, output=65536 n/a multimodal=true; reasoning=true

Verification notes

No LLM verification step ran; model metadata was already complete in the issue.

sync_models vs proposed update

sync_models cross-check found differences. Official provider verification was used for the applied values, and sync_models discrepancies are listed below for review.

Model Field Proposed update sync_models sync_models source models
databricks-gemini-3-6-flash catalog entry present missing None

#1025: [BOT ISSUE] Fireworks: add missing accounts/fireworks/models/kimi-k3

Model Display name Parent Providers Format Flavor Token limits Pricing Lifecycle
accounts/fireworks/models/kimi-k3 Kimi K3 fireworks openai chat input=1040000, output=not provided in/out=3/15 per 1M; cache read=0.3 per 1M multimodal=true; reasoning=true

Verification notes

Verification

sync_models vs proposed update

sync_models cross-check found differences. Official provider verification was used for the applied values, and sync_models discrepancies are listed below for review.

Model Field Proposed update sync_models sync_models source models
accounts/fireworks/models/kimi-k3 catalog entry present missing None

@vercel

vercel Bot commented Jul 27, 2026

Copy link
Copy Markdown

The latest updates on your projects. Learn more about Vercel for GitHub.

Project Deployment Actions Updated (UTC)
ai-proxy Ready Ready Preview, Comment Jul 27, 2026 7:31pm

Request Review

@github-actions

Copy link
Copy Markdown
Contributor Author

Codex (@codex) review

@chatgpt-codex-connector chatgpt-codex-connector Bot left a comment

Copy link
Copy Markdown

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

💡 Codex Review

Here are some automated review suggestions for this pull request.

Reviewed commit: 733a7dc49f

ℹ️ About Codex in GitHub

Codex has been enabled to automatically review pull requests in this repo. Reviews are triggered when you

  • Open a pull request for review
  • Mark a draft as ready
  • Comment "Codex (@codex) review".

If Codex has suggestions, it will comment; otherwise it will react with 👍.

When you sign up for Codex through ChatGPT, Codex can also answer questions or update the PR, like "Codex (@codex) address that feedback".

"databricks-meta-llama-3-1-8b-instruct": ["databricks"],
"openai/gpt-oss-120b": ["groq", "together", "baseten"],
"openai/gpt-oss-20b": ["groq"], // NOTE: We use groq pricing for this and Together pricing for the 120B model
"openai/gpt-oss-20b": ["groq", "together"],

Copy link
Copy Markdown

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

P1 Badge Keep the shared model pinned to Groq pricing

When this model is selected through Groq, the catalog now applies Together's $0.05/$0.20 rates even though Groq's rates remain $0.075/$0.30. model_list.json stores only one price per model, and this change both replaces that price and widens the direct endpoint mapping to Together; the previous mapping deliberately pinned this ID to Groq for pricing. Consequently, Groq usage for this model is underestimated by one third. Keep the direct mapping and catalog rates pinned to Groq, or introduce provider-specific pricing before routing the same ID through both providers.

Useful? React with 👍 / 👎.

Erin McNulty (erin2722) added a commit that referenced this pull request Jul 27, 2026
The daily bot-issues resolver (fix_bot_issue.ts) added models straight from a
ticket without consulting the sync-exclusion list, so models we removed kept
coming back every day (e.g. #1026 re-added realtime, Vertex MaaS, and the
ENDPOINT_NOT_FOUND Databricks endpoints). Wire isModelExcludedFromSync into
resolveIssueCommand: excluded ids are dropped from the actionable set and the
issue is closed as unsupported when nothing else remains.

Extend MANUAL_SYNC_EXCLUDED_MODELS with the entries trimmed from the catalog PR:
- databricks-gpt-5-4-mini, databricks-gemini-3-6-flash (ENDPOINT_NOT_FOUND, not
  deployed on the account)
- the 9 Vertex MaaS ids (real models, but added deliberately via a dedicated PR
  with the gateway OpenAPI routing change, not by the daily automation)

Verified: resolve-issue on an excluded model now returns action "unsupported"
and leaves the catalog untouched. Tests: 44 pass.

Co-Authored-By: Claude Opus 4.8 <noreply@anthropic.com>
Erin McNulty (erin2722) and others added 3 commits July 27, 2026 18:57
Gemini Live uses a bidirectional API and is rejected by generateContent
("not supported for generateContent"), so it is not a chat/completions model —
same rationale as the gpt-realtime-2.1 entries already excluded. Add it to
MANUAL_SYNC_EXCLUDED_MODELS so the LiteLLM sync stops re-adding it.

Co-Authored-By: Claude Opus 4.8 <noreply@anthropic.com>
The daily bot-issues resolver (fix_bot_issue.ts) added models straight from a
ticket without consulting the sync-exclusion list, so models we removed kept
coming back every day (e.g. #1026 re-added realtime, Vertex MaaS, and the
ENDPOINT_NOT_FOUND Databricks endpoints). Wire isModelExcludedFromSync into
resolveIssueCommand: excluded ids are dropped from the actionable set and the
issue is closed as unsupported when nothing else remains.

Extend MANUAL_SYNC_EXCLUDED_MODELS with the entries trimmed from the catalog PR:
- databricks-gpt-5-4-mini, databricks-gemini-3-6-flash (ENDPOINT_NOT_FOUND, not
  deployed on the account)
- the 9 Vertex MaaS ids (real models, but added deliberately via a dedicated PR
  with the gateway OpenAPI routing change, not by the daily automation)

Verified: resolve-issue on an excluded model now returns action "unsupported"
and leaves the catalog untouched. Tests: 44 pass.

Co-Authored-By: Claude Opus 4.8 <noreply@anthropic.com>
…d Databricks)

Now that this branch carries the permanent sync-exclusion list, drop the 11 entries
it otherwise both adds and excludes: gpt-realtime-2.1(+mini), the 7 Vertex MaaS ids
present, and databricks-gpt-5-4-mini / databricks-gemini-3-6-flash (ENDPOINT_NOT_FOUND).
Leaves the day's legitimate additions (kimi-k3, mythos, non-ENF databricks, bedrock).

Co-Authored-By: Claude Opus 4.8 <noreply@anthropic.com>
The daily bot dropped `openrouter` from 8 models' provider lists (gpt-4o, o1,
o3-mini, o4-mini, gpt-4-turbo, gpt-4.1-nano, mistral-small-2603, grok-4.5).
Keep openrouter for now — the correct fix is a slug-keyed openrouter entry
(e.g. `x-ai/grok-4.5`), coming via a `sync-openrouter` automation, rather than
dropping openrouter support. Restores available_providers (and grok-4.5's
index.ts entry); leaves the batch's other field changes intact.

Co-Authored-By: Claude Opus 4.8 <noreply@anthropic.com>
@erin2722
Erin McNulty (erin2722) merged commit 46aa35a into main Jul 27, 2026
5 of 6 checks passed
Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Projects

None yet

2 participants