Skip to content

chore: update model catalog from bot issues - #1081

Open
github-actions[bot] wants to merge 1 commit into
mainfrom
chore/autofix-bot-issues-2026-08-08
Open

chore: update model catalog from bot issues#1081
github-actions[bot] wants to merge 1 commit into
mainfrom
chore/autofix-bot-issues-2026-08-08

Conversation

@github-actions

@github-actions github-actions Bot commented Aug 8, 2026

Copy link
Copy Markdown
Contributor

Automated daily batch of model catalog updates from bot issues.

Included issues

Summary

Issue Provider Primary model Changed models Added models Updated models Verification sources
#1072 perplexity perplexity/kimi-k2.7-code perplexity/kimi-k2.7-code perplexity/kimi-k2.7-code None 1
2
#1073 groq qwen/qwen3.6-27b qwen/qwen3.6-27b None qwen/qwen3.6-27b 1
2
#1074 mistral codestral-2508 codestral-2508 None codestral-2508 1
2
#1075 databricks databricks-kimi-k3 databricks-kimi-k3 databricks-kimi-k3 None 1
2
#1076 bedrock anthropic.claude-3-5-sonnet-20241022-v2:0 anthropic.claude-3-5-sonnet-20241022-v2:0
anthropic.claude-3-5-sonnet-20240620-v1:0
None anthropic.claude-3-5-sonnet-20241022-v2:0
anthropic.claude-3-5-sonnet-20240620-v1:0
1
#1077 bedrock google.gemma-4-31b google.gemma-4-31b
google.gemma-4-26b-a4b
google.gemma-4-e2b
None google.gemma-4-31b
google.gemma-4-26b-a4b
google.gemma-4-e2b
1
2
#1079 vertex publishers/google/models/gemini-2.5-flash publishers/google/models/gemini-2.5-flash
publishers/google/models/gemini-2.5-pro
None publishers/google/models/gemini-2.5-flash
publishers/google/models/gemini-2.5-pro
1
2
#1080 bedrock meta.llama4-scout-17b-instruct-v1:0 meta.llama4-scout-17b-instruct-v1:0
meta.llama4-maverick-17b-instruct-v1:0
None meta.llama4-scout-17b-instruct-v1:0
meta.llama4-maverick-17b-instruct-v1:0
1
2

Verified metadata

#1072: [BOT ISSUE] Perplexity: add missing perplexity/kimi-k2.7-code

Model Display name Parent Providers Format Flavor Token limits Pricing Lifecycle
perplexity/kimi-k2.7-code Perplexity: Kimi K2.7 Code perplexity openai chat input=262144, output=not provided in/out=0.95/4 per 1M; cache read=0.19 per 1M active

Verification notes

No LLM verification step ran; model metadata was already complete in the issue.

sync_models vs proposed update

sync_models cross-check found differences. Official provider verification was used for the applied values, and sync_models discrepancies are listed below for review.

Model Field Proposed update sync_models sync_models source models
perplexity/kimi-k2.7-code catalog entry present missing None

#1073: [BOT ISSUE] Groq: fix qwen/qwen3.6-27b max_output_tokens (32768 → 16384)

Model Display name Parent Providers Format Flavor Token limits Pricing Lifecycle
qwen/qwen3.6-27b Qwen 3.6 27B groq, openrouter openai chat input=131072, output=16384 in/out=0.6/3 per 1M active

Verification notes

No LLM verification step ran; model metadata was already complete in the issue.

sync_models vs proposed update

sync_models cross-check found differences. Official provider verification was used for the applied values, and sync_models discrepancies are listed below for review.

Model Field Proposed update sync_models sync_models source models
qwen/qwen3.6-27b catalog entry present missing None

#1074: [BOT ISSUE] Mistral: fix codestral-2508 max_input_tokens regression (256k → 128k)

Model Display name Parent Providers Format Flavor Token limits Pricing Lifecycle
codestral-2508 Codestral 2508 codestral-latest mistral, openrouter openai chat input=128000, output=not provided in/out=0.3/0.9 per 1M parent=codestral-latest

Verification notes

No LLM verification step ran; model metadata was already complete in the issue.

sync_models vs proposed update

sync_models cross-check found differences. Official provider verification was used for the applied values, and sync_models discrepancies are listed below for review.

Model Field Proposed update sync_models sync_models source models
codestral-2508 max_input_tokens 128000 256000 mistral/codestral-2508
codestral-2508 max_output_tokens n/a 256000 mistral/codestral-2508

#1075: [BOT ISSUE] Databricks: add missing databricks-kimi-k3 and databricks-gpt-5-4-mini

Model Display name Parent Providers Format Flavor Token limits Pricing Lifecycle
databricks-kimi-k3 Kimi K3 databricks openai chat input=1048576, output=not provided n/a multimodal=true; reasoning=true

Verification notes

Verification

sync_models vs proposed update

sync_models cross-check found differences. Official provider verification was used for the applied values, and sync_models discrepancies are listed below for review.

Model Field Proposed update sync_models sync_models source models
databricks-kimi-k3 catalog entry present missing None

#1076: [BOT ISSUE] Bedrock: fix Claude 3.5 Sonnet pricing (Extended Access $6/$30 since Dec 2025)

Model Display name Parent Providers Format Flavor Token limits Pricing Lifecycle
anthropic.claude-3-5-sonnet-20241022-v2:0 Claude Sonnet 3.5 v2 bedrock anthropic chat input=1000000, output=8192 in/out=6/30 per 1M; cache read=0.6 per 1M; cache write=7.5 per 1M multimodal=true
anthropic.claude-3-5-sonnet-20240620-v1:0 Claude Sonnet 3.5 bedrock anthropic chat input=1000000, output=4096 in/out=6/30 per 1M; cache read=0.3 per 1M; cache write=3.75 per 1M multimodal=true

Verification notes

Verification

sync_models vs proposed update

sync_models cross-check found differences. Official provider verification was used for the applied values, and sync_models discrepancies are listed below for review.

Model Field Proposed update sync_models sync_models source models
anthropic.claude-3-5-sonnet-20241022-v2:0 input_cost_per_mil_tokens 6 3 anthropic.claude-3-5-sonnet-20241022-v2:0
anthropic.claude-3-5-sonnet-20241022-v2:0 output_cost_per_mil_tokens 30 15 anthropic.claude-3-5-sonnet-20241022-v2:0
anthropic.claude-3-5-sonnet-20241022-v2:0 input_cache_read_cost_per_mil_tokens 0.6 0.3 anthropic.claude-3-5-sonnet-20241022-v2:0
anthropic.claude-3-5-sonnet-20241022-v2:0 input_cache_write_cost_per_mil_tokens 7.5 3.75 anthropic.claude-3-5-sonnet-20241022-v2:0
anthropic.claude-3-5-sonnet-20240620-v1:0 input_cost_per_mil_tokens 6 3 anthropic.claude-3-5-sonnet-20240620-v1:0
anthropic.claude-3-5-sonnet-20240620-v1:0 output_cost_per_mil_tokens 30 15 anthropic.claude-3-5-sonnet-20240620-v1:0

#1077: [BOT ISSUE] Bedrock: add missing pricing for Gemma 4 models (31b, 26b-a4b, e2b)

Model Display name Parent Providers Format Flavor Token limits Pricing Lifecycle
google.gemma-4-31b Gemma 4 31B bedrock openai chat input=256000, output=not provided in/out=0.14/0.4 per 1M multimodal=true; reasoning=true
google.gemma-4-26b-a4b Gemma 4 26B A4B bedrock openai chat input=256000, output=not provided in/out=0.13/0.4 per 1M multimodal=true; reasoning=true
google.gemma-4-e2b Gemma 4 E2B bedrock openai chat input=128000, output=not provided in/out=0.04/0.08 per 1M multimodal=true; reasoning=true

Verification notes

Verification

sync_models vs proposed update

sync_models cross-check found differences. Official provider verification was used for the applied values, and sync_models discrepancies are listed below for review.

Model Field Proposed update sync_models sync_models source models
google.gemma-4-31b catalog entry present missing None
google.gemma-4-26b-a4b catalog entry present missing None
google.gemma-4-e2b catalog entry present missing None

#1079: [BOT ISSUE] Vertex: add missing cache read pricing for Gemini 2.5 Flash and Pro

Model Display name Parent Providers Format Flavor Token limits Pricing Lifecycle
publishers/google/models/gemini-2.5-flash Gemini 2.5 Flash vertex google chat input=1048576, output=65535 in/out=0.3/2.5 per 1M; cache read=0.03 per 1M multimodal=true; reasoning=true
publishers/google/models/gemini-2.5-pro Gemini 2.5 Pro vertex google chat input=1048576, output=65535 in/out=1.25/10 per 1M; cache read=0.13 per 1M multimodal=true; reasoning=true

Verification notes

Verification

sync_models vs proposed update

sync_models cross-check found differences. Official provider verification was used for the applied values, and sync_models discrepancies are listed below for review.

Model Field Proposed update sync_models sync_models source models
publishers/google/models/gemini-2.5-flash catalog entry present missing None
publishers/google/models/gemini-2.5-pro catalog entry present missing None

#1080: [BOT ISSUE] Bedrock: add missing available_providers to Llama 4 Scout and Maverick base entries

Model Display name Parent Providers Format Flavor Token limits Pricing Lifecycle
meta.llama4-scout-17b-instruct-v1:0 Llama 4 Scout bedrock converse chat input=10000000, output=8000 n/a multimodal=true
meta.llama4-maverick-17b-instruct-v1:0 Llama 4 Maverick bedrock converse chat input=1000000, output=8000 n/a multimodal=true

Verification notes

Verification

sync_models vs proposed update

sync_models cross-check found differences. Official provider verification was used for the applied values, and sync_models discrepancies are listed below for review.

Model Field Proposed update sync_models sync_models source models
meta.llama4-scout-17b-instruct-v1:0 max_input_tokens 10000000 128000 meta.llama4-scout-17b-instruct-v1:0
meta.llama4-scout-17b-instruct-v1:0 max_output_tokens 8000 4096 meta.llama4-scout-17b-instruct-v1:0
meta.llama4-scout-17b-instruct-v1:0 input_cost_per_mil_tokens n/a 0.17 meta.llama4-scout-17b-instruct-v1:0
meta.llama4-scout-17b-instruct-v1:0 output_cost_per_mil_tokens n/a 0.66 meta.llama4-scout-17b-instruct-v1:0
meta.llama4-maverick-17b-instruct-v1:0 max_input_tokens 1000000 128000 meta.llama4-maverick-17b-instruct-v1:0
meta.llama4-maverick-17b-instruct-v1:0 max_output_tokens 8000 4096 meta.llama4-maverick-17b-instruct-v1:0
meta.llama4-maverick-17b-instruct-v1:0 input_cost_per_mil_tokens n/a 0.24 meta.llama4-maverick-17b-instruct-v1:0
meta.llama4-maverick-17b-instruct-v1:0 output_cost_per_mil_tokens n/a 0.97 meta.llama4-maverick-17b-instruct-v1:0

@vercel

vercel Bot commented Aug 8, 2026

Copy link
Copy Markdown

The latest updates on your projects. Learn more about Vercel for GitHub.

Project Deployment Actions Updated (UTC)
ai-proxy Ready Ready Preview Aug 8, 2026 10:25am

Request Review

@github-actions

github-actions Bot commented Aug 8, 2026

Copy link
Copy Markdown
Contributor Author

Codex (@codex) review

@chatgpt-codex-connector chatgpt-codex-connector Bot left a comment

Copy link
Copy Markdown

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

💡 Codex Review

Here are some automated review suggestions for this pull request.

Reviewed commit: 440a746668

ℹ️ About Codex in GitHub

Codex has been enabled to automatically review pull requests in this repo. Reviews are triggered when you

  • Open a pull request for review
  • Mark a draft as ready
  • Comment "Codex (@codex) review".

If Codex has suggestions, it will comment; otherwise it will react with 👍.

When you sign up for Codex through ChatGPT, Codex can also answer questions or update the PR, like "Codex (@codex) address that feedback".

"gpt-5.4-nano": ["openai", "azure", "openrouter"],
"Qwen/Qwen3.7-Plus": ["together"],
"qwen/qwen3.6-27b": ["groq"],
"qwen/qwen3.6-27b": ["groq", "openrouter"],

Copy link
Copy Markdown

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

P2 Badge Keep OpenRouter out of Qwen's direct mapping

This widens the direct endpoint mapping for a model that still has Groq as its native provider. The catalog sync code explicitly keeps OpenRouter aggregator unions out of direct index.ts endpoint types unless OpenRouter is the only provider (packages/proxy/scripts/sync_models.ts:1360-1366); otherwise getDirectModelEndpointTypes("qwen/qwen3.6-27b") starts offering OpenRouter routing/credential fallback instead of leaving this native model direct-routed to Groq. Keep openrouter in available_providers if desired, but don't add it to AvailableEndpointTypes here.

Useful? React with 👍 / 👎.

"databricks"
]
},
"databricks-kimi-k3": {

Copy link
Copy Markdown

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

P2 Badge Add the missing Databricks GPT-5.4 Mini entry

The batch text says #1075 adds both databricks-kimi-k3 and databricks-gpt-5-4-mini, but only Kimi is introduced here; I checked the repo with rg 'databricks-gpt-5-4-mini' and there is still no model_list.json or AvailableEndpointTypes entry. In the scenario where this release is expected to expose the new Databricks GPT-5.4 Mini model, requests or selection for that ID will continue to fail as unknown even though the issue is closed.

Useful? React with 👍 / 👎.

Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment