litellm/tests/test_litellm/caching
Kent 1260f09679 fix(cache): resolve semantic-cache embedding model via Router model-group lookup
resolve_embedding_router gated on exact model-name membership, so wildcard/pattern
(bedrock/*) and visible model_group_alias embedding models fell back to a direct
litellm.embedding() call and lost deployment-level auth (e.g. Bedrock
aws_role_name), the same class of failure as #28244. Resolve via
Router.get_model_list(model_name=...), which unifies exact + visible alias +
provider-prefixed wildcard/pattern, and drop the now-redundant llm_model_list arg.

get_model_list is a listing resolver, not a faithful 'would the Router route this'
predicate: hidden model_group_alias entries, a bare model name matched only against
a provider-prefixed wildcard, a raw litellm_params.model / model_info.id deployment
address, and team-public names without a team_id still resolve empty and keep
falling back to a direct embedding call. These are niche embedding-model configs
and degrade to current behavior; documented as known gaps.
2026-06-29 18:35:49 +08:00
..
test_azure_blob_cache.py style: run black formatter on files from main merge 2026-04-17 13:02:59 -07:00
test_caching.py feat(caching): add valkey-semantic cache backend and fix semantic cache scope keys (#30675) 2026-06-19 17:09:17 -07:00
test_caching_handler.py fix(caching): restore stored prompt_tokens on embedding cache hits instead of recomputing (#30046) 2026-06-10 15:49:20 +05:30
test_check_and_fix_namespace_none_guard.py chore: litellm oss staging160626 (#30527) 2026-06-16 18:23:13 -07:00
test_dual_cache.py feat(rate-limiter): allow opting out of v3 TPM reservation and Redis circuit breaker (#30211) 2026-06-11 10:34:26 -07:00
test_embedding_router.py fix(cache): resolve semantic-cache embedding model via Router model-group lookup 2026-06-29 18:35:49 +08:00
test_gcs_cache.py chore: litellm oss staging (#30745) 2026-06-18 13:55:35 -07:00
test_in_memory_cache.py fix: prune expired in-memory cache heap entries (#25664) 2026-04-14 23:37:49 +05:30
test_llm_caching_handler.py fix: don't close HTTP/SDK clients on LLMClientCache eviction (#22925) 2026-03-05 12:00:38 -08:00
test_llm_client_cache_e2e.py fix: don't close HTTP/SDK clients on LLMClientCache eviction (#22925) 2026-03-05 12:00:38 -08:00
test_qdrant_semantic_cache.py fix(cache): resolve semantic-cache embedding model via Router model-group lookup 2026-06-29 18:35:49 +08:00
test_redis_cache.py fix(redis): loop-scope async Lua script registration (#31501) 2026-06-27 12:20:40 -07:00
test_redis_cluster_cache.py fix(caching): check REDIS_CLUSTER_NODES env var in Cache and Router class selection (#22790) 2026-03-06 17:31:30 -08:00
test_redis_connection_pool.py style: run black formatter on files from main merge 2026-04-17 13:02:59 -07:00
test_redis_semantic_cache.py fix(cache): resolve semantic-cache embedding model via Router model-group lookup 2026-06-29 18:35:49 +08:00
test_s3_cache.py style: run black formatter on files from main merge 2026-04-17 13:02:59 -07:00
test_valkey_semantic_cache.py feat(caching): add valkey-semantic cache backend and fix semantic cache scope keys (#30675) 2026-06-19 17:09:17 -07:00