litellm/tests/test_litellm/proxy/anthropic_endpoints
Krrish Dholakia fcd285facd fix(claude-code): retry transient fetch failures, don't abort sync on one
A sync fans out one request per skill folder; on a repo with dozens of
skills, a single request occasionally hitting a transient connection
blip was expected, not exceptional. Previously any one failure aborted
the entire asyncio.gather with an unhandled exception, so admin retries
on a big repo were coin flips even after the concurrency fix.

Two changes:
- _http_get retries transient httpx errors (not real HTTP status codes)
  with a short backoff before giving up.
- A single skill that still fails after retries is skipped rather than
  fatal, and counted via a new persisted skipped_count column so an
  admin can actually see whether every skill loaded, instead of an
  opaque all-or-nothing success/error.

Verified live: 10/10 registrations succeeded for both a manifest-based
repo (alirezarezvani/claude-skills, 83 skills) and the dir-scan fallback
path that was actually flaky before (garrytan/gstack, 53 skills).
2026-07-10 21:35:58 -07:00
..
__init__.py Litellm fix GitHub action testing (#11163) 2025-05-26 14:41:42 -07:00
test_claude_code_marketplace.py test: isolate proxy master_key/prisma_client module globals between tests 2026-04-23 15:31:16 -07:00
test_claude_code_marketplace_sources_endpoints.py fix(claude-code): retry transient fetch failures, don't abort sync on one 2026-07-10 21:35:58 -07:00
test_claude_code_marketplace_sync.py fix(claude-code): retry transient fetch failures, don't abort sync on one 2026-07-10 21:35:58 -07:00
test_claude_code_skill_authz.py feat(claude-code): import multiple skill marketplaces with per-skill access grants 2026-07-10 19:14:23 -07:00
test_endpoints.py feat: litellm oss staging (#31935) 2026-07-03 09:27:31 +05:30