Skip to content

feat(aws-bedrock/eu.anthropic.claude-sonnet-5-5): add new models [bot] - #4097

Open
models-bot[bot] wants to merge 2 commits into
mainfrom
bot/add-aws-bedrock-eu.anthropic.claude-sonnet-5-5-20261001-023455
Open

models-bot[bot] wants to merge 2 commits into
mainfrom
bot/add-aws-bedrock-eu.anthropic.claude-sonnet-5-5-20261001-023455

Conversation

@models-bot

@models-bot models-bot Bot commented Oct 1, 2026 •

Copy link
Copy Markdown
Contributor

Auto-generated by model-addition-agent for aws-bedrock/eu.anthropic.claude-sonnet-5-5.


Note

Low Risk
Adds static model metadata only; no runtime or auth logic changes.

Overview
Adds a new provider catalog file for eu.anthropic.claude-sonnet-5-5 on AWS Bedrock so routing, pricing, and capability metadata can target the EU inference profile separately from existing eu.anthropic.claude-sonnet-5 and global.anthropic.claude-sonnet-5-5 entries.

The definition marks the model active and serverless, with per-token costs (including prompt-cache fields) for eight EU regions, chat mode, text/image input, thinking enabled, a reasoning_effort parameter (default high), and standard Claude tool/caching features. Limits are set to a 1M context window and 128k max output tokens.

Reviewed by Cursor Bugbot for commit 6f4af14. Bugbot is set up for automated code reviews on this repo. Configure here.

@cursor cursor Bot left a comment

Copy link
Copy Markdown

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

Cursor Bugbot has reviewed your changes using default effort and found 2 potential issues.

Fix All in Cursor

❌ Bugbot Autofix is OFF. To automatically fix reported issues with cloud agents, enable autofix in the Cursor dashboard.

Reviewed by Cursor Bugbot for commit 6f4af14. Configure here.

- high
- xhigh
- max
type: string

Copy link
Copy Markdown

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

Inherited max_tokens cap too low

High Severity

This fuller config sets max_output_tokens to 128000 but does not override inherited max_tokens maxValue. Provider default.yaml caps max_tokens at 2048, so the gateway clamps or rejects outputs far below the model's 128000 limit.

Fix in Cursor Fix in Web

Triggered by learned rule: Override max_tokens maxValue to match limits

Reviewed by Cursor Bugbot for commit 6f4af14. Configure here.

cache_read_input_token_cost: 2.2e-7
input_cost_per_token: 0.0000022
output_cost_per_token: 0.000011
region: eu-west-3

Copy link
Copy Markdown

Choose a reason for hiding this comment

The reason will be displayed to describe this comment to others. Learn more.

Fuller costs omit batch rates

Medium Severity

This fuller Claude cost table has token and cache rates but omits input_cost_per_token_batches and output_cost_per_token_batches. Batch inference estimates miss the Claude Batch 50% discount and price jobs at on-demand rates.

Fix in Cursor Fix in Web

Triggered by learned rule: AWS Bedrock per-id metadata is not a sibling copy

Reviewed by Cursor Bugbot for commit 6f4af14. Configure here.

Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

0 participants