mirror of
https://github.com/BerriAI/litellm.git
synced 2026-10-09 03:18:44 +00:00
get_fireworks_session_id fell back to litellm_trace_id when no session id was given. That id is generated per request (uuid4 when absent), so x-session-affinity carried a different value every time and Fireworks prompt caching never hit; cached_tokens stayed 0 across identical prompts. The None path the original change described was effectively unreachable because of it. Drop the fallback so affinity comes only from an id the caller actually supplied: litellm_session_id, session_id, or metadata.session_id. Callers who were relying on a trace id for affinity can pass litellm_session_id instead, which is stable across the requests they want grouped. Co-authored-by: mubashir1osmani <mubashir.osmani777@gmail.com> |
||
|---|---|---|
| .. | ||
| chat | ||
| completion | ||
| rerank | ||
| test_fireworks_ai_common_utils.py | ||
| test_fireworks_ai_cost_calculator.py | ||
| test_fireworks_ai_kimi_model_metadata.py | ||