Skip to content

chore(sync): update Cloudflare Workers AI model catalog - #7640

Merged
opencode-agent[bot] merged 1 commit into
devfrom
automation/sync-models-cloudflare-workers-ai
Sep 21, 2026
Merged

opencode-agent[bot] merged 1 commit into
devfrom
automation/sync-models-cloudflare-workers-ai

Conversation

@opencode-agent

Copy link
Copy Markdown
Contributor

Updates model TOMLs for the cloudflare-workers-ai sync target.

Provider Status Created Updated Deleted
Cloudflare Workers AI changed 0 3 0
Cloudflare Workers AI changed files
  • updated: /home/runner/work/models.dev/models.dev/providers/cloudflare-workers-ai/models/@cf/zai-org/glm-5.3.toml
  • updated: /home/runner/work/models.dev/models.dev/providers/cloudflare-workers-ai/models/@cf/zai-org/glm-5.3-flash.toml
  • updated: /home/runner/work/models.dev/models.dev/providers/cloudflare-workers-ai/models/@cf/deepseek-ai/deepseek-v4-flash-0731.toml

This PR was created automatically by the model sync workflow.

@opencode-agent opencode-agent Bot added automation Automated model catalog sync model-sync Automated model catalog sync provider:cloudflare-workers-ai Automated model catalog sync labels Sep 21, 2026
@opencode-agent
opencode-agent Bot enabled auto-merge (squash) September 21, 2026 16:28
@opencode-agent
opencode-agent Bot merged commit 739a4e2 into dev Sep 21, 2026
2 checks passed
@opencode-agent
opencode-agent Bot deleted the automation/sync-models-cloudflare-workers-ai branch September 21, 2026 16:29
rekram1-node added a commit that referenced this pull request Sep 28, 2026
Some listings advertise a larger window than the serving backend accepts, so
a client that fits its output to the listed window still sends requests the
provider rejects.

Workers AI: glm-5.3, glm-5.3-flash and deepseek-v4-flash-0731 are listed at
1,310,720 tokens by /ai/models/search and the model pages, but the backend
rejects any request over 1,048,576 ("Requested token count exceeds the model's
maximum context length of 1048576 tokens"). The deepseek-v4-flash-0731 model
page already says 1,048,576. #7454 capped these, then the automated sync in
#7640 restored 1,310,720 because it re-derives context from search on every
run. Pin the served window in the sync (it only ever lowers a window) and
correct the three files.

AI Gateway: anthropic/claude-sonnet-4.5 was curated to 1,000,000 since the
generator port, but the gateway rejects prompts over 200,000 tokens
("prompt is too long: 231024 tokens > 200000 maximum") and the model page
says 200,000. Drop the curated override so the catalog value applies.
Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Labels

automation Automated model catalog sync model-sync Automated model catalog sync provider:cloudflare-workers-ai Automated model catalog sync

Projects

None yet

Development

Successfully merging this pull request may close these issues.

0 participants