open-webui/backend
Classic298 e88e1f8119
fix: prompt cache misses after a model calls tools one after another (#31593)
When a model called one tool, got its result and then called a second tool with no text in between, the next request merged both calls into an earlier assistant message that the provider had already seen. That changed the conversation's beginning, so the provider's prompt cache stopped matching from there for the rest of the chat. Each tool call and its result are now sent as their own messages, so the start of the conversation stays identical from one request to the next.

Fixes #31588
2026-10-01 07:34:59 +04:00
..
data refac: mv backend files to /open_webui dir 2024-09-04 16:54:48 +02:00
open_webui fix: prompt cache misses after a model calls tools one after another (#31593) 2026-10-01 07:34:59 +04:00
.dockerignore fix: litellm config issue 2024-02-24 22:35:11 -08:00
.gitignore refac 2024-09-06 04:59:20 +02:00
dev.sh perf: allow disabling websocket per-message-deflate (#28613) 2026-08-24 18:46:07 -04:00
requirements-slim.txt chore: bump pycrdt to 0.14.8 (#31636) 2026-09-30 19:00:23 +04:00
requirements.txt chore: bump pycrdt to 0.14.8 (#31636) 2026-09-30 19:00:23 +04:00
start.sh refac 2026-09-06 16:48:30 -04:00
start_windows.bat chore: drop nltk, unused at the pinned versions (#29725) 2026-09-06 16:39:05 -04:00