Commit Graph

30881 Commits

Author SHA1 Message Date
Sameer Kankute
5b5a2d12a8 remove test 2025-08-21 00:32:16 +05:30
Sameer Kankute
1d1d637ac0 fix automated checks 2025-08-21 00:09:17 +05:30
Sameer Kankute
37b84d1582 Add rerank endpoint support for deepinfra 2025-08-20 23:36:05 +05:30
Ishaan Jaff
12bae8fdda test_partner_models_httpx_streaming 2025-08-20 09:44:47 -07:00
Ishaan Jaff
e69a895884 test_apply_default_settings 2025-08-20 08:53:38 -07:00
Ishaan Jaff
45a9af8063 security fix 2025-08-20 08:40:49 -07:00
Ishaan Jaff
e229a6a796 security fix 2025-08-20 08:35:59 -07:00
Qingchuan Hao
f2a6be390b feat(utils.py): accept 'api_version' as param for validate_environment 2025-08-20 14:29:58 +00:00
hgyun.lee
036b358636 Synchronize cache behavior between acompletion and completion 2025-08-20 18:27:36 +09:00
Fang Gong
d2b943f391 fix application inference profile for pass-through endpoints for bedrock 2025-08-20 01:06:29 -07:00
Fang Gong
fe51647e3c fix application inference profile for pass-through endpoints for bedrock 2025-08-20 01:03:49 -07:00
Fang Gong
caad9b3ca7 fix application inference profile for pass-through endpoints for bedrock 2025-08-20 00:40:43 -07:00
Tomu Hirata
5879f6e930 comment 2025-08-20 15:46:53 +09:00
Tomu Hirata
c1dbd1c2c1 remove comment
Signed-off-by: Tomu Hirata <tomu.hirata@gmail.com>
2025-08-20 15:42:26 +09:00
Tomu Hirata
d64b579131 Include predicted output in tracing
Signed-off-by: Tomu Hirata <tomu.hirata@gmail.com>
2025-08-20 15:39:32 +09:00
Tim Elfrink
972d7f7133 Merge branch 'main' of https://github.com/BerriAI/litellm into feat/github-copilot-thinking-reasoning-support 2025-08-20 07:58:11 +02:00
Krish Dholakia
4df07a5060
Merge pull request #13622 from BerriAI/fix-passthrough-deletion-failure
Fix query passthrough deletion
2025-08-19 22:34:07 -07:00
Krish Dholakia
bfe814f981
Merge pull request #13637 from Tasmay-Tibrewal/main
Added Qwen3, Deepseek R1 0528 Throughput, GLM 4.5 and GPT-OSS models for Together AI
2025-08-19 22:33:17 -07:00
Krish Dholakia
749bafc550
Merge pull request #13660 from BerriAI/litellm_azure_gpt-t
[LLM Translation] Adjust max_input_tokens for azure/gpt-5-chat models in JSON configuration
2025-08-19 22:32:56 -07:00
Krish Dholakia
04feaf33c0
Merge pull request #13748 from hakasecurity/migrate-to-aim-new-firewall-api
Migrate to aim new firewall api
2025-08-19 22:32:22 -07:00
Krish Dholakia
be30bc68ae
Merge pull request #13759 from kankute-sameer/litellm_feat_correct_cost_calculations
Add long context support for claude-4-sonnet
2025-08-19 22:30:25 -07:00
Krrish Dholakia
7d09375d52 fix: fix gpt-5-chat mappings 2025-08-19 22:21:00 -07:00
Ishaan Jaff
1832c09d6b
[Feat] - UI Allow using Key/Team Based Logging for Langfuse OTEL (#13791)
* add langfuse_otel

* ui allow setting langfuse OTEL

* fixes - using langfuse OTEL

* add description for logging integrations

* add description
2025-08-19 19:06:26 -07:00
Ishaan Jaff
a328ad56e3
[Bug Fix] Fixes for using Auto Router with LiteLLM Docker Image (#13788)
* fix install auto router.sh

* fixes for Docker IMG
2025-08-19 18:36:30 -07:00
Ishaan Jaff
56ac778316
[Bug Fix] Bedrock KB - Using LiteLLM Managed Credentials for Query (#13787)
* fix: add get_credentials_for_vector_store

* test_search_uses_registry_credentials

* test_bedrock_search_with_credentials_managed_registry
2025-08-19 15:39:36 -07:00
mubashir1osmani
ef9e50458d removed faq 2025-08-19 17:01:31 -04:00
tanjiro
b0e3469202
Models page row UI restructure (#13771)
* remove unused model dashboard

* move model_dashboard to templates folder

* moving  columns to molecules directory

* group model names and provider icon

* increase table column name font size. cleanup extra title and description inside the tab

* fix column width and truncate string

* moved credentials column

* combined created by and created at

* combine in and out costs

* remove edit button

* remove 2nd extra tooltip
2025-08-19 14:00:58 -07:00
Philip Kiely
b8167cf304 rip out some old stuff 2025-08-19 13:55:33 -07:00
mubashir1osmani
2edcc51e57 updated claude-code docs and deployment faq 2025-08-19 16:52:46 -04:00
Philip Kiely
6fd1705183 lint 2025-08-19 13:36:41 -07:00
Sameer Kankute
48622a4ee7 add cache above 200k keys in INTENDED_SCHEMA 2025-08-20 01:47:34 +05:30
Philip Kiely
7c3d522435 Update Baseten LiteLLM integration 2025-08-19 12:21:05 -07:00
Sameer Kankute
fa81b6c639
Update test_cost_calculator.py 2025-08-19 23:14:22 +05:30
Ishaan Jaff
195ea6515e
[Feat] Datadog LLM Observability - Add support for tracing guardrail input/output (#13767)
* add guardrail information on DD LLM Obs

* test_guardrail_information_in_metadata
2025-08-19 10:26:25 -07:00
Sameer Kankute
b5f0a7b49b
Merge branch 'main' into litellm_feat_correct_cost_calculations 2025-08-19 18:55:36 +05:30
Sameer Kankute
d39b2e8888 Add test for long context cost calculation 2025-08-19 16:56:13 +05:30
Sameer Kankute
4f38a63152 Add long context support for claude-4-sonnet 2025-08-19 16:40:12 +05:30
drorbaron
7fedcf1ea9 migrate stream ws 2025-08-19 12:36:29 +03:00
drorbaron
6b78ade918 migrate to use new aim FW API 2025-08-19 12:20:19 +03:00
openhands
7530f13e05 Fix IRSA role assumption logic for AWS Bedrock
- Fixed condition to properly detect IRSA environments using AWS_WEB_IDENTITY_TOKEN_FILE
- Skip role assumption when already running as target role in IRSA environment
- Prevents unnecessary AWS API calls that cause 'root account cannot assume role' errors
- Resolves failing test test_auth_with_aws_role_same_role_irsa

The original issue was that aws_access_key_id and aws_secret_access_key were being
populated from environment variables even when passed as None, causing the IRSA
detection condition to fail. The fix checks for IRSA-specific environment variables
instead of relying on the absence of explicit credentials.

Fixes #13417
2025-08-19 08:27:24 +00:00
openhands
93651c9da7 Revert "Revert "fix: role chaining and session name with webauthentication for aws be…" (#13230)"
This reverts commit 342fd2d8b6.
2025-08-19 08:19:18 +00:00
Tim Elfrink
b5fa2ee73f Merge remote-tracking branch 'origin/main' into feat/github-copilot-thinking-reasoning-support 2025-08-19 10:11:59 +02:00
Krish Dholakia
de59691c4b Enable update/delete org members on UI (#8560)
* feat(organization_endpoints.py): expose new `/organization/delete` endpoint. Cascade org deletion to member, teams and keys

Ensures any org deletion is handled correctly

* test(test_organizations.py): add simple test to ensure org deletion works

* feat(organization_endpoints.py): expose /organization/update endpoint, and define response models for org delete + update

* fix(organizations.tsx): support org delete on UI + move org/delete endpoint to use DELETE

* feat(organization_endpoints.py): support `/organization/member_update` endpoint

Allow admin to update member's role within org

* feat(organization_endpoints.py): support deleting member from org

* test(test_organizations.py): add e2e test to ensure org member flow works

* fix(organization_endpoints.py): fix code qa check

* fix(schema.prisma): don't introduce ondelete:cascade - breaking change

* docs(organization_endpoints.py): document missing params
2025-08-19 10:53:24 +03:00
Tim Elfrink
9b0fda7b14 fix: resolve case sensitivity and test failures for extended thinking support
- Fix supports_reasoning() call to use lowercase model names for proper lookup
- Remove custom_llm_provider parameter as model registry entries are provider-agnostic
- Update tests to use full model names with date stamps (required for supports_reasoning)
- Add test coverage for models without extended thinking support
2025-08-19 08:40:10 +02:00
Tim Elfrink
9f82b89051 fix: remove redundant github_copilot check in get_supported_openai_params
The provider_config_manager already handles github_copilot provider
through LlmProviders.GITHUB_COPILOT mapping, making the explicit
check unnecessary.
2025-08-19 08:21:18 +02:00
Tim Elfrink
8b66b50c31 fix: restrict thinking/reasoning_effort parameters to models with extended thinking support
Only models in the 4 family and 3-7 family support extended thinking features.
Previously all models would incorrectly receive these parameters.

Now uses supports_reasoning() to check model registry for actual capability.
2025-08-19 08:15:50 +02:00
Krrish Dholakia
137a98a5af bump: version 1.75.8 → 1.75.9 2025-08-18 23:11:08 -07:00
Krish Dholakia
435995ba5c
Merge pull request #13617 from moandersson/fix/migratejob-resources
Add possibility to configure resources for migrations-job in Helm chart
2025-08-18 23:03:35 -07:00
Krish Dholakia
88e52c55d0
Merge pull request #13675 from colesmcintosh/fix/groq-streaming-encoding
Fix Groq streaming ASCII encoding issue
2025-08-18 23:00:00 -07:00
Krish Dholakia
9dadd279a4
Merge pull request #13741 from BerriAI/litellm_dev_08_18_2025_p1
Refactor - forward model group headers - reuse same logic as global header forwarding
2025-08-18 22:58:39 -07:00