litellm/litellm-rust/crates/gateway-inference
2026-09-04 14:57:48 -07:00
..
benchmarks/realtime refactor(rust): replace ai gateway with inference crate 2026-09-04 14:57:48 -07:00
src refactor(rust): replace ai gateway with inference crate 2026-09-04 14:57:48 -07:00
tests refactor(rust): replace ai gateway with inference crate 2026-09-04 14:57:48 -07:00
AGENTS.md refactor(rust): replace ai gateway with inference crate 2026-09-04 14:57:48 -07:00
ARCHITECTURE.md refactor(rust): replace ai gateway with inference crate 2026-09-04 14:57:48 -07:00
Cargo.toml refactor(rust): replace ai gateway with inference crate 2026-09-04 14:57:48 -07:00
README.md refactor(rust): replace ai gateway with inference crate 2026-09-04 14:57:48 -07:00

LiteLLM Gateway Inference

litellm-gateway-inference is the reusable, framework-independent inference domain used by the Rust gateway server and Python bridge

It owns inference integrations, transport-neutral orchestration, and legacy OCR, audio transcription, and realtime I/O that have not yet moved into litellm-core. It has no Axum or Tower dependency

The executable server, HTTP/WebSocket routes, auth extractors, application state, config startup, Docker image, and deployment files live in ../gateway-server