litellm/litellm
2024-04-23 12:01:13 +02:00
..
deprecated_litellm_server
integrations fix(lowest_tpm_rpm_v2.py): use a combined tpm+rpm query in async get cache, to reduce redis client calls in high traffic 2024-04-20 16:13:11 -07:00
llms feat - watsonx refractoring, removed dependency, and added support for embedding calls 2024-04-23 12:01:13 +02:00
proxy ui - new build 2024-04-20 19:32:40 -07:00
router_strategy fix(lowest_tpm_rpm_v2.py): use a combined tpm+rpm query in async get cache, to reduce redis client calls in high traffic 2024-04-20 16:13:11 -07:00
tests (ci/cd) fix test_master_key_hashing 2024-04-20 14:50:34 -07:00
types feat(prometheus_services.py): emit proxy latency for successful llm api requests 2024-04-18 16:04:35 -07:00
__init__.py feat - watsonx refractoring, removed dependency, and added support for embedding calls 2024-04-23 12:01:13 +02:00
_logging.py fix(parallel_request_limiter.py): handle metadata being none 2024-03-14 10:02:41 -07:00
_redis.py fix(_redis.py): support redis ssl as a kwarg REDIS_SSL 2024-04-20 10:19:44 -07:00
_service_logger.py fix(test_lowest_tpm_rpm_routing_v2.py): unit testing for usage-based-routing-v2 2024-04-18 21:38:00 -07:00
_version.py
budget_manager.py
caching.py Merge branch 'main' into litellm_ssl_caching_fix 2024-04-19 17:20:27 -07:00
cost.json
exceptions.py fix(main.py): map list input to ollama prompt input format 2024-02-16 11:54:12 -08:00
main.py feat - watsonx refractoring, removed dependency, and added support for embedding calls 2024-04-23 12:01:13 +02:00
model_prices_and_context_window_backup.json feat - add llama3 on groq 2024-04-19 14:34:38 -07:00
requirements.txt
router.py fix(router.py): async simple-shuffle support 2024-04-20 15:01:12 -07:00
timeout.py
utils.py feat - watsonx refractoring, removed dependency, and added support for embedding calls 2024-04-23 12:01:13 +02:00