litellm/tests
Krish Dholakia 4be0ec8e35
GA Multi-instance rate limiting v2 Requirements + New - specify token rate limit type - output / input / total (#11646)
* feat(parallel_request_limiter_v3.py): allows admin to enforce token rate limit based on just output tokens

Useful when trying to rate limit for primarily self hosted model use-cases

* test(test_parallel_request_limiter_v3.py): add unit test for token rate limit type

* feat(parallel_request_limiter_v3.py): return remaining token limits in header

* feat: return rate limit headers in response

* feat(parallel_request_limiter_v3.py): working rate limit response headers

* feat(parallel_request_limiter_v3.py): fix rate limit tracking for tpm when rpm also set

* feat(parallel_request_limiter_v3.py): show headers for key/user/team

* feat(parallel_request_limiter_v3.py): decrement max parallel request limiter on failure event

* feat(parallel_request_limiter_v3.py): add in-memory cache implementation of parallel request rate limiter

allows rate limiter to work even without redis cache setup

Work for GA of parallel request limiter v3

* refactor(proxy/hooks/__init__.py): replace with new parallel request handler

* test: update testing

* fix: fix ruff check

* fix: revert ga of multi instance rate limiting - needs more work to pass testing
2025-06-11 22:05:13 -07:00
..
basic_proxy_startup_tests
batches_tests Litellm managed file updates combined (#11040) 2025-05-22 17:20:41 -07:00
code_coverage_tests Xai, VertexAI, Google AI Studio - live web search support in OpenAI format (#11251) 2025-05-31 14:26:16 -07:00
documentation_tests
enterprise Support returning virtual key in custom auth + Handle provider-specific optional params for embedding calls (#11346) 2025-06-03 07:24:13 -07:00
guardrails_tests [Feat] Add Lasso Guardrail to LiteLLM (#11565) 2025-06-09 18:47:26 -07:00
image_gen_tests [Fix]: Add cost tracking for image edits endpoint [OpenAI, Azure] (#11186) 2025-05-27 17:52:15 -07:00
litellm_utils_tests Update enduser spend and budget reset date based on budget duration (#8460) 2025-06-08 08:39:14 -07:00
litellm-proxy-extras Prisma Migrate - support setting custom migration dir (#10336) 2025-04-26 12:05:06 -07:00
llm_responses_api_testing [Feat] New LLM API Endpoint - Add List input items for Responses API (#11602) 2025-06-10 15:47:16 -07:00
llm_translation [Feat] Perf fix - ensure deepgram provider uses async httpx calls (#11641) 2025-06-11 18:32:01 -07:00
load_tests test: test_embedding_performance 2025-05-14 21:31:07 -07:00
local_testing [Bug Fix] Add audio/ogg mapping for Audio MIME types (#11635) 2025-06-11 14:19:53 -07:00
logging_callback_tests fix(prometheus.py): update tests 2025-06-06 09:12:54 -07:00
mcp_tests [Feat] MCP - Add support for streamablehttp_client MCP Servers (#11628) 2025-06-11 17:09:46 -07:00
multi_instance_e2e_tests
old_proxy_tests/tests test: update tests to new deployment model (#10142) 2025-04-18 14:22:12 -07:00
openai_endpoints_tests # expect an error when getting the response again since 2025-04-25 09:42:35 -07:00
otel_tests [Feat] Prometheus - Track route on proxy_* metrics (#10992) 2025-05-20 22:55:55 -07:00
pass_through_tests test: add more debug logs 2025-04-29 15:40:18 -07:00
pass_through_unit_tests [Feat] Allow using litellm.completion with /v1/messages API Spec (use gpt-4, gemini etc with claude code) (#11502) 2025-06-06 20:35:53 -07:00
proxy_admin_ui_tests test(test_sso_sign_in.py): update test 2025-06-03 21:46:34 -07:00
proxy_security_tests
proxy_unit_tests fix(internal_user_endpoints.py): support user with + in email on us… (#11601) 2025-06-10 22:13:10 -07:00
router_unit_tests Support env var vertex credentials for passthrough + ignore space id on watsonx deployment (throws Json validation errors) (#11527) 2025-06-07 20:31:05 -07:00
scim_tests [Feat SSO] Add LiteLLM SCIM Integration for Team and User management (#10072) 2025-04-16 19:21:47 -07:00
spend_tracking_tests
store_model_in_db_tests feat: Allow Adding MCP Servers Through LiteLLM UI (#11208) 2025-05-28 16:29:27 -07:00
test_litellm GA Multi-instance rate limiting v2 Requirements + New - specify token rate limit type - output / input / total (#11646) 2025-06-11 22:05:13 -07:00
windows_tests [Bug Fix] UnicodeDecodeError: 'charmap' on Windows during litellm import (#10542) 2025-05-03 21:31:05 -07:00
__init__.py Litellm fix GitHub action testing (#11163) 2025-05-26 14:41:42 -07:00
gettysburg.wav
large_text.py
openai_batch_completions.jsonl
README.MD
test_budget_management.py Update enduser spend and budget reset date based on budget duration (#8460) 2025-06-08 08:39:14 -07:00
test_callbacks_on_proxy.py
test_config.py
test_debug_warning.py
test_end_users.py
test_entrypoint.py
test_fallbacks.py Ollama Chat - parse tool calls on streaming (#11171) 2025-05-27 16:14:49 -07:00
test_health.py
test_keys.py test: temporarily skip test due to change testing model change - need to update test for new model 2025-05-09 09:02:08 -07:00
test_logging.conf
test_models.py Ollama Chat - parse tool calls on streaming (#11171) 2025-05-27 16:14:49 -07:00
test_openai_endpoints.py
test_organizations.py UI - fix adding vertex models with reusable credentials + fix pagination on keys table + fix showing org budgets on table (#10528) 2025-05-03 08:16:53 -07:00
test_passthrough_endpoints.py
test_ratelimit.py
test_spend_logs.py
test_team_logging.py
test_team_members.py
test_team.py build: publish new litellm-proxy-extras file 2025-05-27 17:44:23 -07:00
test_users.py test: fix imports 2025-05-26 22:06:53 -07:00

In total litellm runs 1000+ tests

[02/20/2025] Update:

To make it easier to contribute and map what behavior is tested,

we've started mapping the litellm directory in tests/litellm

This folder can only run mock tests.