litellm/tests
Jugal D. Bhatt 7c49197f29
Add Hosted VLLM rerank provider integration (#12738)
* Vllm rerank (#12737)

* Add Hosted VLLM rerank provider integration

This commit implements the Hosted VLLM rerank provider integration for LiteLLM. The integration includes:
Adding Hosted VLLM as a supported rerank provider in the main rerank function
Implementing the HostedVLLMRerank handler class for making API requests
Creating a transformation class to convert Hosted VLLM responses to LiteLLM's standardized format
The integration supports both synchronous and asynchronous rerank operations. API credentials can be provided directly or through environment variables (HOSTED_VLLM_API_KEY and HOSTED_VLLM_API_BASE).
Notable features:
Proper error handling for missing credentials
Standard response transformation
Support for common rerank parameters (top_n, return_documents, etc.)
Proper token usage tracking
This expands LiteLLM's rerank provider ecosystem to include Hosted VLLM alongside existing providers like Cohere, Together AI, Azure AI, and Bedrock.

* refactor(rerank): use base_llm_http_handler for hosted_vllm rerank

- Replace custom HostedVLLMRerank handler with base_llm_http_handler
- Implement proper HostedVLLMRerankConfig inheriting from BaseRerankConfig
- Follow Cohere-compatible implementation pattern
- Clean up unnecessary comments

* Fix lint errors in hosted_vllm rerank transformer: remove unused imports

* Fix linting errors in rerank transformation modules

* fix: resolve type errors in Hosted VLLM rerank module

---------

Co-authored-by: Philip D'Souza <philip.dsouza@macro4.com>
Co-authored-by: Philip D'Souza <philip.a.dsouza@gmail.com>

* added a few tests

---------

Co-authored-by: Philip D'Souza <philip.dsouza@macro4.com>
Co-authored-by: Philip D'Souza <philip.a.dsouza@gmail.com>
2025-07-18 10:55:50 -07:00
..
basic_proxy_startup_tests
batches_tests openai.ConflictError 2025-07-12 17:07:21 -07:00
code_coverage_tests [Bug fix] [Bug]: Verbose log is enabled by default (#12596) 2025-07-14 20:06:16 -07:00
documentation_tests
enterprise (#11794) use upsert for managed object table rather than create to avoid UniqueViolationError (#11795) 2025-07-15 20:20:01 -07:00
guardrails_tests [Feat] Bedrock Guardrails - Allow disabling exception on 'BLOCKED' action (#12693) 2025-07-17 08:46:33 -07:00
image_gen_tests feat: add input_fidelity parameter for OpenAI image generation (#12662) 2025-07-16 16:56:05 -07:00
litellm_utils_tests [Bug Fix] Always include tool calls in output of trim_messages (#11517) 2025-07-17 16:01:59 -07:00
litellm-proxy-extras
llm_responses_api_testing test_mcp_tools_with_responses_api 2025-07-03 14:53:30 -07:00
llm_translation fix: don't fail request if unmapped item in responses list 2025-07-16 09:25:26 -07:00
load_tests
local_testing fix cohere InternalServerError error mapping 2025-07-16 16:13:34 -07:00
logging_callback_tests [Refactor] Vector Stores - Use class VectorStorePreCallHook for all Vector Store Integrations (#12715) 2025-07-17 16:31:58 -07:00
mcp_tests [MCP Gateway] List tools from access list for keys (#12657) 2025-07-16 14:28:53 -07:00
multi_instance_e2e_tests
old_proxy_tests/tests
openai_endpoints_tests Batches - support batch retrieve with target model Query Param + Anthropic - completion bridge, yield content_block_stop chunk (#12228) 2025-07-01 22:13:48 -07:00
otel_tests test_bedrock_guardrail_triggered 2025-07-09 17:05:06 -07:00
pass_through_tests test_anthropic_basic_completion_with_headers ci/cd fix 2025-07-12 11:08:03 -07:00
pass_through_unit_tests test fix - test_anthropic_messages_passthrough.py 2025-06-30 21:56:31 -07:00
proxy_admin_ui_tests Fix e2e test (#12549) 2025-07-12 10:42:57 -07:00
proxy_security_tests
proxy_unit_tests [Feat] Allow reading custom logger python scripts from s3 (#12623) 2025-07-16 15:07:01 -07:00
router_unit_tests test_mock_router_testing_params_str_to_bool_conversion 2025-06-27 18:07:33 -07:00
scim_tests
spend_tracking_tests
store_model_in_db_tests test_create_mcp_server_invalid_alias 2025-07-16 15:58:34 -07:00
test_litellm Add Hosted VLLM rerank provider integration (#12738) 2025-07-18 10:55:50 -07:00
unified_google_tests [Bug Fix] Using gemini-cli with Vertex Anthropic Models (#12246) 2025-07-03 12:29:27 -07:00
vector_store_tests [Refactor] Use Existing config structure for bedrock vector stores (#12672) 2025-07-17 11:55:11 -07:00
windows_tests
__init__.py [Feat] Add github co-pilot as a new LLM API provider (#12325) 2025-07-04 13:12:16 -07:00
gettysburg.wav
large_text.py
openai_batch_completions.jsonl
README.MD
test_budget_management.py Update enduser spend and budget reset date based on budget duration (#8460) 2025-06-08 08:39:14 -07:00
test_callbacks_on_proxy.py
test_config.py
test_debug_warning.py
test_end_users.py
test_entrypoint.py
test_fallbacks.py Ollama Chat - parse tool calls on streaming (#11171) 2025-05-27 16:14:49 -07:00
test_health.py
test_keys.py
test_logging.conf
test_models.py Add model access groups on UI (#11719) 2025-06-13 21:20:25 -07:00
test_openai_endpoints.py
test_organizations.py
test_passthrough_endpoints.py
test_ratelimit.py
test_resource_cleanup.py Fix: Properly close aiohttp client sessions to prevent resource leaks (#12251) 2025-07-09 09:25:17 -07:00
test_spend_logs.py
test_team_logging.py
test_team_members.py
test_team.py build: publish new litellm-proxy-extras file 2025-05-27 17:44:23 -07:00
test_users.py test: fix imports 2025-05-26 22:06:53 -07:00

In total litellm runs 1000+ tests

[02/20/2025] Update:

To make it easier to contribute and map what behavior is tested,

we've started mapping the litellm directory in tests/litellm

This folder can only run mock tests.