litellm/tests/llm_translation/test_text_completion.py
devin-ai-integration[bot] fa2c8984ba
test: move the unit half of 126 mixed legacy files into tests/unit (#45090)
* test: move the unit half of 126 mixed legacy files into tests/unit

* test: restore litellm globals that moved tests set

* test: finalize migration test cleanup

* test: restore original bodies of moved legacy tests

The move into tests/unit had rewritten 612 test bodies, and some of the rewrites dropped assertions. Each moved test now carries its original body from the legacy file, with only the imports, helpers, fake provider credentials and monkeypatched env it needs to run under tests/unit

test_timeout_streaming goes back to tests/local_testing because it needs the fake OpenAI endpoint server. The image payload fixture moves with its only user, and two tests that leaked global state (a registered model cost entry and queued logging tasks) are now isolated

* test: drop module imports shadowed by restored local imports

* test: assert on LiteLLM output in no-assertion moved tests and isolate leaks

Twenty no-assertion candidates get one assertion on the value LiteLLM returns, with the original lines unchanged. Four tests go back to their legacy files because they only check types or imports, write into the working directory, or cannot assert without a body change

Two moved tests leaked globals into later tests in the same worker, so monkeypatch fixtures now restore the retry-after header parser and the end user cost tracking flags

* test: drain queued logging tasks before the Phoenix span test

The moved Phoenix test counted spans from logging tasks that earlier tests had queued, so the drain fixture moves to tests/unit/conftest.py and both it and the Datadog batch test use it. test_factory_function goes back to its legacy file because its returned wrapper calls the real Assistants API and cannot be asserted on without a body change

---------

Co-authored-by: yuneng <yuneng@berri.ai>
2026-10-07 14:07:43 -07:00

37 lines
1 KiB
Python

import litellm
import pytest
@pytest.mark.asyncio
@pytest.mark.parametrize("sync_mode", [True, False])
async def test_text_completion_include_usage(sync_mode):
"""Test text completion with include_usage"""
last_chunk = None
if sync_mode:
response = await litellm.atext_completion(
model="gpt-3.5-turbo",
prompt="Hello, world!",
stream=True,
stream_options={"include_usage": True},
)
async for chunk in response:
print(chunk)
last_chunk = chunk
else:
response = litellm.text_completion(
model="gpt-3.5-turbo",
prompt="Hello, world!",
stream=True,
stream_options={"include_usage": True},
)
for chunk in response:
print(chunk)
last_chunk = chunk
assert last_chunk is not None
assert last_chunk.usage is not None
assert last_chunk.usage.prompt_tokens > 0
assert last_chunk.usage.completion_tokens > 0
assert last_chunk.usage.total_tokens > 0