* Add `litellm-proxy` CLI (#10478) * First cut at a Python client module for proxy * Add UnauthorizedError + add_model method * Add delete_model method * Add example model_id to delete_model docstring * Make delete_model raise NotFoundError * Add get_model * Add get_all_model_info * Rename models.list_models to models.list * Rename models.get_all_model_info to models.info * Move ModelsManagementClient.get_all_model_group_info to ModelGroupsManagementClient.info * Rename get_model to get * Rename add_model to new * Rename delete_model to delete * In client classes, rename base_url attribute to _base_url and api_key attribute to _api_key * Add ModelsManagementClient.updae method * Add client.chat.completions (ChatClient) * ruff format litellm/proxy/client * ruff format tests/litellm/proxy/client/*.py * Add latest changes * Rename KeysManagementClient.create to KeysManagementClient.generate * Add new parameters to KeysManagementClient.generate * Add CredentialsManagementClient * Remove api_key parameter from KeysManagementClient.generate * Fix lint errors * Add litellm/proxy/client/README.md * README.md: Remove api_key param to client.keys.generate * Fix mypy errors * First cut at litellm-proxy cli * Add test for `litellm-proxy models list` * Nicer get_models_info * get_models_info: --columns option * Use format_timestamp in list_models * ruff format litellm/proxy/client * Simpler JSON printing with rich.print_json * Move models-related commands to separate file From `cli.py` to `groups/models.py` * Improve directory structure * Cleanup cli/groups/models.py - esp. usage of rich * Refactoring * Refactor mocking in cli/test_main.py * Dedup models commands tests * Update poetry.lock * Fix mypy errors * ruff format litellm/proxy/client/cli * ruff format tests/litellm/proxy/client/*.py * Fix timezone issue in test_models_list_table_format * Add cli/README.md * Small README.md tweaks * README.md enhancements * Add credentials commands * Add chat commands * Add http commands * ruff format litellm/proxy/client/cli * Fix lint errors in credentials and http commands * json => json_lib * test-key => sk-test-key * Mock HTTP responses so http command tests pass * Fix mypy error in credentials.py * bump: version 1.67.5 → 1.67.6 * build: update litellm version * cli/main.py: show_envvar=True * Increase test job timeout to 8 minutes because it looks like maybe the job is getting canceled because it takes too long with the additional tests? This probably could be reverted once #10484 is merged, since that speeds up pytest runs greatly. * Add keys functionality to library/CLI * Add info about keys commands to litellm/proxy/client/cli/README.md * Move Model Information section in CLI README * Make Model Information a level 4 heading * Move rich to extras as suggested by @ishaan-jaff --------- Co-authored-by: Krrish Dholakia <krrishdholakia@gmail.com> * pin rich=13.7.1 --------- Co-authored-by: Marc Abramowitz <abramowi@adobe.com> Co-authored-by: Krrish Dholakia <krrishdholakia@gmail.com>
56 lines
2.2 KiB
Plaintext
56 lines
2.2 KiB
Plaintext
# LITELLM PROXY DEPENDENCIES #
|
|
anyio==4.5.0 # openai + http req.
|
|
httpx==0.27.0 # Pin Httpx dependency
|
|
openai==1.68.2 # openai req.
|
|
fastapi==0.115.5 # server dep
|
|
backoff==2.2.1 # server dep
|
|
pyyaml==6.0.2 # server dep
|
|
uvicorn==0.29.0 # server dep
|
|
gunicorn==23.0.0 # server dep
|
|
uvloop==0.21.0 # uvicorn dep, gives us much better performance under load
|
|
boto3==1.34.34 # aws bedrock/sagemaker calls
|
|
redis==5.2.1 # redis caching
|
|
prisma==0.11.0 # for db
|
|
mangum==0.17.0 # for aws lambda functions
|
|
pynacl==1.5.0 # for encrypting keys
|
|
google-cloud-aiplatform==1.47.0 # for vertex ai calls
|
|
anthropic[vertex]==0.21.3
|
|
mcp==1.5.0 # for MCP server
|
|
google-generativeai==0.5.0 # for vertex ai calls
|
|
async_generator==1.10.0 # for async ollama calls
|
|
langfuse==2.45.0 # for langfuse self-hosted logging
|
|
prometheus_client==0.20.0 # for /metrics endpoint on proxy
|
|
ddtrace==2.19.0 # for advanced DD tracing / profiling
|
|
orjson==3.10.12 # fast /embedding responses
|
|
apscheduler==3.10.4 # for resetting budget in background
|
|
fastapi-sso==0.16.0 # admin UI, SSO
|
|
pyjwt[crypto]==2.9.0
|
|
python-multipart==0.0.18 # admin UI
|
|
Pillow==11.0.0
|
|
azure-ai-contentsafety==1.0.0 # for azure content safety
|
|
azure-identity==1.16.1 # for azure content safety
|
|
azure-storage-file-datalake==12.20.0 # for azure buck storage logging
|
|
opentelemetry-api==1.25.0
|
|
opentelemetry-sdk==1.25.0
|
|
opentelemetry-exporter-otlp==1.25.0
|
|
sentry_sdk==2.21.0 # for sentry error handling
|
|
detect-secrets==1.5.0 # Enterprise - secret detection / masking in LLM requests
|
|
cryptography==43.0.1
|
|
tzdata==2025.1 # IANA time zone database
|
|
litellm-proxy-extras==0.1.15 # for proxy extras - e.g. prisma migrations
|
|
### LITELLM PACKAGE DEPENDENCIES
|
|
python-dotenv==1.0.0 # for env
|
|
tiktoken==0.8.0 # for calculating usage
|
|
importlib-metadata==6.8.0 # for random utils
|
|
tokenizers==0.20.2 # for calculating usage
|
|
click==8.1.7 # for proxy cli
|
|
rich==13.7.1 # for litellm proxy cli
|
|
jinja2==3.1.6 # for prompt templates
|
|
aiohttp==3.10.2 # for network calls
|
|
aioboto3==12.3.0 # for async sagemaker calls
|
|
tenacity==8.2.3 # for retrying requests, when litellm.num_retries set
|
|
pydantic==2.10.2 # proxy + openai req.
|
|
jsonschema==4.22.0 # validating json schema
|
|
websockets==13.1.0 # for realtime API
|
|
####
|