Skip to content

[vscode-lm] Use measured input budgets for Opus 5.5 and Astra - #21

Merged
simurg79 merged 2 commits into
mainfrom
fix/vscode-lm-opus-55-astra-budgets
Sep 24, 2026
Merged

simurg79 merged 2 commits into
mainfrom
fix/vscode-lm-opus-55-astra-budgets

Conversation

@simurg79

Copy link
Copy Markdown
Owner

What is the problem?

Opus 5.5 is missing from the current table, and Astra's advertised and accepted input budgets are stale. Using advertised limits as usable input budgets can reject requests; borrowing the earlier 197,897-token placeholder understates Opus 5.5's verified budget by about 3.4x.

How does this PR solve the problem?

Adds Opus 5.5 and updates GPT-6 Astra ("Astra 6") using session measurements from a live VS Code 1.137.0 extension host on 2026-09-23, with 62 models enumerated. Both measurements below use the copilot vendor.

Model Advertised input Accepted input Rejected input Size trials
claude-opus-5.5 871,793 677,108 695,778 13
gpt-6-astra 921,793 271,789 271,790 20

Opus's configured 677,108 is a verified lower bound, not an exact ceiling. Repeated "Model declined to respond" errors prevented convergence. The earlier 935,793 advertised figure was incorrect; 871,793 is the observed advertisement. Astra's adjacent-token boundary applies to the tested prompt format, not every possible request shape. Image input and tool calling worked for both models.

Unverified/default fields: prompt caching is unverifiable through the VS Code LM API: there is no cache-hit evidence or separately exposed total-context window. The supportsPromptCache value of false is a schema-required default, not a measured finding. The contextWindow field stores the advertised input value, not a measured total-context capacity. Pricing was not measured: Astra's optional prices are omitted; Opus's zero prices are placeholders, not evidence of free usage. Neither copilotcli variant could be measured because token counting was unavailable and streams were empty.

Tests pin the measured rows and defaults while preserving main's other model updates. Only the provider table and its tests are changed. No changeset is added, consistent with recent model-table changes in this fork.

How did you test the PR?

cd packages/types && npx vitest run src/__tests__/vscode-llm.spec.ts

11 passed, 0 failed, 0 skipped in one file; exit 0 (Vitest 3.2.4). Commit-hook lint and push-hook type checks each passed 10/10 tasks. Hooks warned that Node 24.15.0 differs from the requested 20.19.2. Live-probe results above come from the implementation session; they were not rerun during PR preparation.

Agent notes

Base: ac5411f (simurg79/Roo-Code main).
Scope excludes roo-vault and unrelated scratch files.

@simurg79
simurg79 merged commit 4bf5874 into main Sep 24, 2026
7 checks passed
Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

1 participant