litellm/tests/test_litellm/llms
Krish Dholakia ee6e76e1f9
Bedrock Passthrough cost tracking (/invoke + /converse routes - streaming + non-streaming) (#12123)
* refactor(passthrough_endpoints-success-handler): refactor llm passthrough logging logic

isolate the llm translation work to enable cost tracking on sdk

* feat: initial implementation of passthrough SDK cost calculation

enables bedrock passthrough cost tracking to work

* feat(cost_calculator.py): working cost calculation for bedrock passthrough

* feat(litellm_logging.py): consider allm_passthrough in cost tracking

allows async calls (e.g. via proxy) to work

* feat(bedrock/passthrough): working event stream decoding for bedrock passthrough calls + logging instrumentation for passthrough sdk calls (log on stream completion)

Enables bedrock streaming cost calculation

* feat(litellm_logging.py): support streaming passthrough cost tracking

* feat(passthrough/main.py): working async streaming cost calculation

Closes https://github.com/BerriAI/litellm/issues/11359

* feat(proxy_server.py): fix passthrough routing when llm router enabled

* feat: further fixes

* feat(bedrock/): working bedrock passthrough cost tracking (non-streaming)

* feat(litellm_logging.py): working usage tracking for bedrock passthrough calls

ensures tokens are logged

* feat(bedrock/passthrough): add converse passthrough cost tracking support

* feat(base_llm/passthrough): remove redundant function

* refactor(litellm_logging.py): refactor function to be below 50 LOC

* test: update test

* test: remove redundant test
2025-06-27 20:01:12 -07:00
..
anthropic [Bug Fix] Anthropic - Token Usage Null Handling in calculate_usage (#12068) 2025-06-27 10:00:23 -07:00
azure Bedrock Passthrough cost tracking (/invoke + /converse routes - streaming + non-streaming) (#12123) 2025-06-27 20:01:12 -07:00
azure_ai/chat Litellm dev 06 18 2025 p1 (#11872) 2025-06-18 21:24:36 -07:00
bedrock fix aws bedrock claude tool call index (#11842) 2025-06-20 23:21:08 -07:00
chat
cohere/chat
custom_httpx [Fix] Allow using HTTP_ Proxy settings with trust_env (#12066) 2025-06-26 08:37:22 -07:00
databricks
datarobot Add support for DataRobot as a provider in LiteLLM (#10385) 2025-06-02 08:37:21 -07:00
deepgram [Feat] Add Eleven Labs - Speech To Text Support on LiteLLM (#12119) 2025-06-27 17:50:49 -07:00
featherless_ai/chat test: fixes 2025-05-31 12:42:56 -07:00
fireworks_ai/chat
gemini feat: Add audio parameter support to gemini tts models (#11287) 2025-05-31 16:20:19 -07:00
hosted_vllm/chat
huggingface refactor: cleanup huggingface rerank transformation 2025-06-06 10:30:44 -07:00
litellm_proxy/chat LiteLLM SDK <-> Proxy improvement (don't transform message client-side) + Bedrock - handle qs:.. in base64 file data + Tag Management - support adding public model names (#11908) 2025-06-19 22:34:18 -07:00
llamafile/chat
lm_studio
meta_llama [Feat] Enable Tool Calling for meta_llama (#11895) 2025-06-19 13:44:22 -07:00
mistral [Fix] Magistral small system prompt diverges too much from the official recommendation (#12007) 2025-06-24 13:45:58 -07:00
nebius test: fixes 2025-05-31 12:42:56 -07:00
novita/chat
nscale/chat
ollama Update mistral 'supports_response_schema' field + Fix ollama embedding (#12024) 2025-06-25 07:20:13 -07:00
openai fix: make response api support Azure Authentication method (#11941) 2025-06-23 08:43:20 -07:00
openrouter/chat
perplexity feat: implement Perplexity citation tokens and search queries cost calculation (#11938) 2025-06-23 14:15:25 -07:00
sagemaker
vertex_ai Litellm dev 06 23 2025 p1 (#11989) 2025-06-23 22:33:06 -07:00
watsonx Fixing watsonx error: 'model_id' or 'model' cannot be specified in the request body for models in a deployment space (#11854) 2025-06-23 10:14:10 -07:00
test_volcengine.py Anthropic /v1/messages - Custom LLM Server support (#12016) 2025-06-24 22:00:44 -07:00