Skip to content

chore: update model catalog from bot issues#1000

Open
github-actions[bot] wants to merge 1 commit into
mainfrom
chore/autofix-bot-issues-2026-07-19
Open

chore: update model catalog from bot issues#1000
github-actions[bot] wants to merge 1 commit into
mainfrom
chore/autofix-bot-issues-2026-07-19

Conversation

@github-actions

Copy link
Copy Markdown
Contributor

Automated daily batch of model catalog updates from bot issues.

Included issues

Summary

Issue Provider Primary model Changed models Added models Updated models Verification sources
#990 xai grok-4.5 grok-4.5
grok-4.5-latest
None grok-4.5
grok-4.5-latest
1
2
#991 openai gpt-realtime-2.1 gpt-realtime-2.1
gpt-realtime-2.1-mini
gpt-realtime-2.1
gpt-realtime-2.1-mini
None 1
2
3
4
5
#992 openai gpt-4o gpt-4o
o1
o3-mini
None gpt-4o
o1
o3-mini
1
2
#993 openai o4-mini o4-mini
gpt-4-turbo
gpt-4.1-nano
None o4-mini
gpt-4-turbo
gpt-4.1-nano
1
#994 vertex publishers/mistralai/models/mistral-medium-3 publishers/mistralai/models/mistral-medium-3
publishers/mistralai/models/mistral-small-2503
publishers/mistralai/models/codestral-2
publishers/mistralai/models/mistral-medium-3
publishers/mistralai/models/mistral-small-2503
publishers/mistralai/models/codestral-2
None 1
2
3
4
5
6
#996 vertex publishers/openai/models/gpt-oss-120b-maas publishers/openai/models/gpt-oss-120b-maas
publishers/openai/models/gpt-oss-20b-maas
publishers/openai/models/gpt-oss-120b-maas
publishers/openai/models/gpt-oss-20b-maas
None 1
2
3
4
5
#997 vertex publishers/xai/models/grok-4.3 publishers/xai/models/grok-4.3
publishers/xai/models/grok-4.20-non-reasoning
publishers/xai/models/grok-4.3
publishers/xai/models/grok-4.20-non-reasoning
None 1
2
3
#999 together thinkingmachines/inkling thinkingmachines/inkling None thinkingmachines/inkling 1
2

Verified metadata

#990: [BOT ISSUE] xAI: fix stale cached input pricing for grok-4.5 and grok-4.5-latest ($0.50 → $0.30)

Model Display name Parent Providers Format Flavor Token limits Pricing Lifecycle
grok-4.5 Grok 4.5 xAI openai chat input=500000, output=500000 in/out=2/6 per 1M; cache read=0.3 per 1M multimodal=true; reasoning=true
grok-4.5-latest Grok 4.5 (Latest) grok-4.5 xAI openai chat input=500000, output=500000 in/out=2/6 per 1M; cache read=0.3 per 1M parent=grok-4.5; multimodal=true; reasoning=true

Verification notes

No LLM verification step ran; model metadata was already complete in the issue.

sync_models vs proposed update

sync_models cross-check found differences. Official provider verification was used for the applied values, and sync_models discrepancies are listed below for review.

Model Field Proposed update sync_models sync_models source models
grok-4.5 input_cache_read_cost_per_mil_tokens 0.3 0.5 xai/grok-4.5
grok-4.5-latest input_cache_read_cost_per_mil_tokens 0.3 0.5 xai/grok-4.5-latest

#991: [BOT ISSUE] OpenAI: add missing gpt-realtime-2.1 and gpt-realtime-2.1-mini

Model Display name Parent Providers Format Flavor Token limits Pricing Lifecycle
gpt-realtime-2.1 GPT Realtime 2.1 openai, azure openai chat input=128000, output=32000 in/out=4/24 per 1M; cache read=0.4 per 1M active
gpt-realtime-2.1-mini GPT Realtime 2.1 mini openai, azure openai chat input=128000, output=32000 in/out=0.6/2.4 per 1M; cache read=0.06 per 1M active

Verification notes

Verification

sync_models vs proposed update

sync_models cross-check found differences. Official provider verification was used for the applied values, and sync_models discrepancies are listed below for review.

Model Field Proposed update sync_models sync_models source models
gpt-realtime-2.1-mini max_output_tokens 32000 4096 gpt-realtime-2.1-mini

#992: [BOT ISSUE] OpenAI: add deprecation_date for gpt-4o, o1, o3-mini (shutdown 2026-10-23)

Model Display name Parent Providers Format Flavor Token limits Pricing Lifecycle
gpt-4o GPT-4o openai, azure openai chat input=128000, output=16384 in/out=2.5/10 per 1M; cache read=1.25 per 1M date=2026-10-23; multimodal=true
o1 o1 openai, azure openai chat input=200000, output=100000 in/out=15/60 per 1M; cache read=7.5 per 1M date=2026-10-23; multimodal=true; reasoning=true
o3-mini o3 mini openai, azure openai chat input=200000, output=100000 in/out=1.1/4.4 per 1M; cache read=0.55 per 1M date=2026-10-23; multimodal=true; reasoning=true

Verification notes

Verification

sync_models vs proposed update

sync_models cross-check found differences. Official provider verification was used for the applied values, and sync_models discrepancies are listed below for review.

Model Field Proposed update sync_models sync_models source models
gpt-4o deprecation_date 2026-10-23 n/a gpt-4o
o1 deprecation_date 2026-10-23 n/a o1
o3-mini deprecation_date 2026-10-23 n/a o3-mini

#993: [BOT ISSUE] OpenAI: add deprecation_date for o4-mini, gpt-4-turbo, gpt-4.1-nano (shutdown 2026-10-23)

Model Display name Parent Providers Format Flavor Token limits Pricing Lifecycle
o4-mini o4-mini openai, azure openai chat input=200000, output=100000 in/out=1.1/4.4 per 1M; cache read=0.275 per 1M date=2026-10-23; multimodal=true; reasoning=true
gpt-4-turbo GPT-4 Turbo openai, azure openai chat input=128000, output=4096 in/out=10/30 per 1M date=2026-10-23; multimodal=true
gpt-4.1-nano GPT-4.1 nano openai, azure openai chat input=1047576, output=32768 in/out=0.1/0.4 per 1M; cache read=0.025 per 1M date=2026-10-23; multimodal=true

Verification notes

Verification

sync_models vs proposed update

sync_models cross-check found differences. Official provider verification was used for the applied values, and sync_models discrepancies are listed below for review.

Model Field Proposed update sync_models sync_models source models
o4-mini deprecation_date 2026-10-23 n/a o4-mini
gpt-4-turbo deprecation_date 2026-10-23 n/a gpt-4-turbo
gpt-4.1-nano deprecation_date 2026-10-23 n/a gpt-4.1-nano

#994: [BOT ISSUE] Vertex: add missing Mistral MaaS entries (mistral-medium-3, mistral-small-2503, codestral-2)

Model Display name Parent Providers Format Flavor Token limits Pricing Lifecycle
publishers/mistralai/models/mistral-medium-3 Mistral Medium 3 (Vertex) vertex openai chat input=131072, output=not provided in/out=0.4/2 per 1M active
publishers/mistralai/models/mistral-small-2503 Mistral Small 3.1 (Vertex) vertex openai chat input=131072, output=not provided in/out=0.1/0.3 per 1M active
publishers/mistralai/models/codestral-2 Codestral 2 (Vertex) vertex openai chat input=262144, output=not provided in/out=0.3/0.9 per 1M active

Verification notes

Verification

sync_models vs proposed update

sync_models cross-check found differences. Official provider verification was used for the applied values, and sync_models discrepancies are listed below for review.

Model Field Proposed update sync_models sync_models source models
publishers/mistralai/models/mistral-medium-3 catalog entry present missing None
publishers/mistralai/models/mistral-small-2503 catalog entry present missing None
publishers/mistralai/models/codestral-2 catalog entry present missing None

#996: [BOT ISSUE] Vertex: add missing OpenAI GPT-OSS MaaS entries (gpt-oss-120b-maas, gpt-oss-20b-maas)

Model Display name Parent Providers Format Flavor Token limits Pricing Lifecycle
publishers/openai/models/gpt-oss-120b-maas GPT-OSS 120B (Vertex) vertex openai chat input=131072, output=32768 in/out=0.15/0.6 per 1M reasoning=true
publishers/openai/models/gpt-oss-20b-maas GPT-OSS 20B (Vertex) vertex openai chat input=131072, output=32768 in/out=0.075/0.3 per 1M reasoning=true

Verification notes

Verification

sync_models vs proposed update

sync_models cross-check found differences. Official provider verification was used for the applied values, and sync_models discrepancies are listed below for review.

Model Field Proposed update sync_models sync_models source models
publishers/openai/models/gpt-oss-120b-maas catalog entry present missing None
publishers/openai/models/gpt-oss-20b-maas catalog entry present missing None

#997: [BOT ISSUE] Vertex: add missing xAI Grok MaaS entries (grok-4.3, grok-4.20-non-reasoning)

Model Display name Parent Providers Format Flavor Token limits Pricing Lifecycle
publishers/xai/models/grok-4.3 Grok 4.3 (Vertex) vertex openai chat input=1000000, output=1000000 in/out=1.25/2.5 per 1M; cache read=0.2 per 1M multimodal=true; reasoning=true
publishers/xai/models/grok-4.20-non-reasoning Grok 4.20 Non-Reasoning (Vertex) vertex openai chat input=1000000, output=not provided in/out=1.25/2.5 per 1M multimodal=true

Verification notes

Verification

sync_models vs proposed update

sync_models cross-check found differences. Official provider verification was used for the applied values, and sync_models discrepancies are listed below for review.

Model Field Proposed update sync_models sync_models source models
publishers/xai/models/grok-4.3 catalog entry present missing None
publishers/xai/models/grok-4.20-non-reasoning catalog entry present missing None

#999: [BOT ISSUE] Together: add together to available_providers for thinkingmachines/inkling

Model Display name Parent Providers Format Flavor Token limits Pricing Lifecycle
thinkingmachines/inkling Inkling baseten, together openai chat input=1048576, output=not provided in/out=1/4.05 per 1M; cache read=0.17 per 1M multimodal=true; reasoning=true

Verification notes

Verification

sync_models vs proposed update

sync_models cross-check found differences. Official provider verification was used for the applied values, and sync_models discrepancies are listed below for review.

Model Field Proposed update sync_models sync_models source models
thinkingmachines/inkling catalog entry present missing None

@github-actions

Copy link
Copy Markdown
Contributor Author

Codex (@codex) review

@vercel

vercel Bot commented Jul 19, 2026

Copy link
Copy Markdown

The latest updates on your projects. Learn more about Vercel for GitHub.

Project Deployment Actions Updated (UTC)
ai-proxy Ready Ready Preview, Comment Jul 19, 2026 10:52am

Request Review

@chatgpt-codex-connector chatgpt-codex-connector Bot left a comment

Copy link
Copy Markdown

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

💡 Codex Review

Here are some automated review suggestions for this pull request.

Reviewed commit: 30b332db65

ℹ️ About Codex in GitHub

Codex has been enabled to automatically review pull requests in this repo. Reviews are triggered when you

  • Open a pull request for review
  • Mark a draft as ready
  • Comment "Codex (@codex) review".

If Codex has suggestions, it will comment; otherwise it will react with 👍.

When you sign up for Codex through ChatGPT, Codex can also answer questions or update the PR, like "Codex (@codex) address that feedback".

Comment on lines +1089 to +1092
"publishers/openai/models/gpt-oss-120b-maas": ["vertex"],
"publishers/openai/models/gpt-oss-20b-maas": ["vertex"],
"publishers/xai/models/grok-4.3": ["vertex"],
"publishers/xai/models/grok-4.20-non-reasoning": ["vertex"],

Copy link
Copy Markdown

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

P1 Badge Route OpenAI/xAI Vertex MaaS through OpenAPI

When these new OpenAI/xAI Vertex models are selected, packages/proxy/src/proxy.ts:2202-2225 only sends publishers/meta and publishers/qwen through the OpenAI-compatible /endpoints/openapi/chat/completions path; these entries therefore fall into the rawPredict branch. Google/xAI docs for Grok use the OpenAI-compatible endpoint and xai/... model IDs, and Google's gpt-oss MaaS examples likewise use the OpenAI-compatible API with openai/... IDs, so these newly advertised Vertex options will return not-found/unsupported until the routing and model-id rewrite are extended. Sources: https://docs.cloud.google.com/gemini-enterprise-agent-platform/models/partner-models/grok/capabilities/function-calling, https://discuss.google.dev/t/now-ga-openais-gpt-oss-qwen3-models-on-vertex-ai-as-open-model-apis/253945

Useful? React with 👍 / 👎.

"available_providers": [
"baseten"
"baseten",
"together"

Copy link
Copy Markdown

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

P2 Badge Use Together's exact Inkling model id

For Together requests the proxy forwards bodyData.model unchanged to the OpenAI-compatible endpoint, and a repo-wide search only finds this lowercase thinkingmachines/inkling catalog key. Together's serverless catalog lists the API model string as thinkingmachines/Inkling, so enabling Together on the lowercase Baseten id advertises a model name that Together users are likely to have rejected; add a correct-case Together entry or a provider-specific translation instead. Source: https://docs.together.ai/docs/serverless/models

Useful? React with 👍 / 👎.

Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment