mirror of
https://github.com/BerriAI/litellm.git
synced 2026-10-06 02:48:13 +00:00
The Rust token counter now mirrors _count_input_tokens key precedence (messages, prompt, input, query/documents) so every LLM route that goes through budget reservation gets the GIL-free count, not only /v1/messages and /v1/chat/completions. Objects are serialised like json.dumps before tokenizing; floats and unknown shapes still decline to Python. The body model is optional so route-selected models can be matched by the caller Co-Authored-By: Devin AI <158243242+devin-ai-integration[bot]@users.noreply.github.com> |
||
|---|---|---|
| .. | ||
| native_route_wheel_test.py | ||
| test_bindings.py | ||
| test_chat_completions.py | ||
| test_configuration.py | ||
| test_runtime.py | ||
| test_token_counter.py | ||
| test_verify_linux_native_wheel.py | ||