fix: correct maxTokens test expectations to match model definition

The test expectations incorrectly expected maxTokens of 16384, but the
claudeCodeModels definition specifies maxTokens of 32768 for all models.
The handler code 'model.info.maxTokens ?? 16384' correctly evaluates to
32768 since the model definition provides the value.

Updated 3 test expectations to use the correct value of 32768.
This commit is contained in:
Hannes Rudolph 2025-12-13 21:29:55 -07:00
parent 537ef5424a
commit 224f21e4d7

View file

@ -110,13 +110,13 @@ describe("ClaudeCodeHandler", () => {
// Verify createStreamingMessage was called with correct parameters
// Default model has reasoning effort of "medium" so thinking should be enabled
// With interleaved thinking, maxTokens comes from model definition, not reasoning config
// With interleaved thinking, maxTokens comes from model definition (32768 for claude-sonnet-4-5)
expect(mockCreateStreamingMessage).toHaveBeenCalledWith({
accessToken: "test-access-token",
model: "claude-sonnet-4-5",
systemPrompt,
messages,
maxTokens: 16384, // model's maxTokens (interleaved thinking uses context window)
maxTokens: 32768, // model's maxTokens from claudeCodeModels definition
thinking: {
type: "enabled",
budget_tokens: 32000, // medium reasoning budget_tokens
@ -159,7 +159,7 @@ describe("ClaudeCodeHandler", () => {
model: "claude-sonnet-4-5",
systemPrompt,
messages,
maxTokens: 16384, // model default maxTokens
maxTokens: 32768, // model maxTokens from claudeCodeModels definition
thinking: { type: "disabled" },
tools: undefined,
toolChoice: undefined,
@ -194,13 +194,13 @@ describe("ClaudeCodeHandler", () => {
await iterator.next()
// Verify createStreamingMessage was called with high thinking config
// With interleaved thinking, maxTokens comes from model definition, not reasoning config
// With interleaved thinking, maxTokens comes from model definition (32768 for claude-sonnet-4-5)
expect(mockCreateStreamingMessage).toHaveBeenCalledWith({
accessToken: "test-access-token",
model: "claude-sonnet-4-5",
systemPrompt,
messages,
maxTokens: 16384, // model's maxTokens (interleaved thinking uses context window)
maxTokens: 32768, // model's maxTokens from claudeCodeModels definition
thinking: {
type: "enabled",
budget_tokens: 64000, // high reasoning budget_tokens