* refactor(passthrough_endpoints-success-handler): refactor llm passthrough logging logic isolate the llm translation work to enable cost tracking on sdk * feat: initial implementation of passthrough SDK cost calculation enables bedrock passthrough cost tracking to work * feat(cost_calculator.py): working cost calculation for bedrock passthrough * feat(litellm_logging.py): consider allm_passthrough in cost tracking allows async calls (e.g. via proxy) to work * feat(bedrock/passthrough): working event stream decoding for bedrock passthrough calls + logging instrumentation for passthrough sdk calls (log on stream completion) Enables bedrock streaming cost calculation * feat(litellm_logging.py): support streaming passthrough cost tracking * feat(passthrough/main.py): working async streaming cost calculation Closes https://github.com/BerriAI/litellm/issues/11359 * feat(proxy_server.py): fix passthrough routing when llm router enabled * feat: further fixes * feat(bedrock/): working bedrock passthrough cost tracking (non-streaming) * feat(litellm_logging.py): working usage tracking for bedrock passthrough calls ensures tokens are logged * feat(bedrock/passthrough): add converse passthrough cost tracking support * feat(base_llm/passthrough): remove redundant function * refactor(litellm_logging.py): refactor function to be below 50 LOC * test: update test * test: remove redundant test |
||
|---|---|---|
| .. | ||
| anthropic | ||
| azure | ||
| azure_ai/chat | ||
| bedrock | ||
| chat | ||
| cohere/chat | ||
| custom_httpx | ||
| databricks | ||
| datarobot | ||
| deepgram | ||
| featherless_ai/chat | ||
| fireworks_ai/chat | ||
| gemini | ||
| hosted_vllm/chat | ||
| huggingface | ||
| litellm_proxy/chat | ||
| llamafile/chat | ||
| lm_studio | ||
| meta_llama | ||
| mistral | ||
| nebius | ||
| novita/chat | ||
| nscale/chat | ||
| ollama | ||
| openai | ||
| openrouter/chat | ||
| perplexity | ||
| sagemaker | ||
| vertex_ai | ||
| watsonx | ||
| test_volcengine.py | ||