Commit Graph

31693 Commits

Author SHA1 Message Date
Sameer Kankute
1ec89b8a04 Feat: add inference_geo based pricing 2026-02-06 13:58:47 +05:30
Sameer Kankute
0934a4ab68 Correct litellm/litellm/llms/anthropic/chat/transformation.py 2026-02-06 13:08:43 +05:30
Sameer Kankute
358a081f63 Add compaction support for vertex ai 2026-02-06 12:52:28 +05:30
Sameer Kankute
1396813d74 The compact beta feature is not currently supported on the Converse and ConverseStream APIs 2026-02-06 12:52:28 +05:30
Sameer Kankute
d0444f402c Add test for compaction in anthropic 2026-02-06 12:52:28 +05:30
Sameer Kankute
7a473f2954
Apply suggestion from @greptile-apps[bot]
Co-authored-by: greptile-apps[bot] <165735046+greptile-apps[bot]@users.noreply.github.com>
2026-02-06 12:07:13 +05:30
Sameer Kankute
ea518a7684
Apply suggestion from @greptile-apps[bot]
Co-authored-by: greptile-apps[bot] <165735046+greptile-apps[bot]@users.noreply.github.com>
2026-02-06 12:06:50 +05:30
Sameer Kankute
887a977ab4 Add doc on how to enable compaction via chat completion 2026-02-06 12:05:36 +05:30
Sameer Kankute
24dda99bd7 Handle compaction block in the input request 2026-02-06 11:55:12 +05:30
Sameer Kankute
c03ba8394e Add compaction block in provider spcific fields streaming+ non streaming 2026-02-06 11:54:47 +05:30
Sameer Kankute
039b37fac1 Add compaction type block in the output 2026-02-06 11:31:35 +05:30
Sameer Kankute
1bcd407af6 Add adaptive thiking for bedrock converse 2026-02-06 09:40:16 +05:30
Sameer Kankute
f15dd691b4 Fix anthropic.claude-opus-4-6-v1 for bedrock 2026-02-06 09:28:50 +05:30
Sameer Kankute
186fd2e64e Add adaptive thinking support for anthropic opus 4.6 2026-02-06 09:24:43 +05:30
Sameer Kankute
3923ef2857
Merge pull request #20491 from BerriAI/litellm_gemini_files_gcs
fix: make sure gcs_bucket_name passes
2026-02-06 08:26:14 +05:30
yuneng-jiang
3bececfa7a
Merge pull request #20539 from BerriAI/litellm_usage_failed_req
[Refactor] UI - Usage Page: Spend By Provider
2026-02-05 16:39:55 -08:00
yuneng-jiang
f64f949715 Spend by provider refactor 2026-02-05 16:22:07 -08:00
Shivam Rawat
93cf1ef517
Merge pull request #20535 from BerriAI/litellm_move_anthropic_test_script
[Chore] Move anthropic input/output test to right folder
2026-02-05 16:12:08 -08:00
yuneng-jiang
523a36ed53
Merge pull request #20530 from BerriAI/litellm_ui_team_soft_budget
[Feature] Add soft_budget to Team Table + Create/Update Endpoints
2026-02-05 15:52:17 -08:00
shivam
3dc70c3398 moved the test anthropic file 2026-02-05 15:44:24 -08:00
Ishaan Jaffer
bff5f8d7b0 doc fix 2026-02-05 15:31:41 -08:00
yuneng-jiang
45c6382667 cz bump + builds 2026-02-05 14:47:11 -08:00
yuneng-jiang
8b0d0b2610 bump: version 0.4.30 → 0.4.31 2026-02-05 14:46:38 -08:00
yuneng-jiang
a17efa1c8e Add soft_budget to team table and create update endpoints 2026-02-05 14:43:48 -08:00
Ishaan Jaff
887a907e42
[Fix] Guardrails API - Ensure OpenAI Moderations Guard works with OpenAI Embeddings (#20523)
* init OpenAIEmbeddingsHandler

* init apply_guardrail

* use apply guardrails for OpenAI moderations

* test_embeddings_handler_string_input

* test_openai_moderation_guardrail_apply_guardrail

* fix typing

* test_openai_moderation_responses_api_input_field

* test fixes
2026-02-05 14:40:15 -08:00
shin-bot-litellm
0649720f79
Fix test isolation for test_log_langfuse_v2_handles_null_usage_values (#20475)
Create fresh mock objects within the test instead of reusing mocks
from setUp that have side_effect configured. The setUp's side_effect
on mock_langfuse_client.trace can interfere with return_value settings
when tests try to reset and reconfigure mocks.

Using dedicated mock objects for this test avoids state pollution
from setUp's side_effect configuration and makes the test more
deterministic in parallel execution environments.
2026-02-05 12:57:41 -08:00
shin-bot-litellm
1d92968e17
Fix test isolation for test_watsonx_gpt_oss_prompt_transformation (#20474)
Set cached tokenizer config directly and mock both sync and async
tokenizer functions to avoid race conditions when running with
parallel test execution (-n 16).

The issue was that parallel tests could populate the
litellm.known_tokenizer_config cache between clearing it and
when the code checked it. This caused the sync code path to be
used instead of the async path, bypassing the mocked async functions.

Fix:
1. Set cache directly instead of clearing it
2. Also mock sync versions _get_tokenizer_config and _get_chat_template_file

This ensures the test is deterministic regardless of test execution order.
2026-02-05 12:57:20 -08:00
yuneng-jiang
b5956cb020
Merge pull request #20513 from BerriAI/litellm_ui_test_cov_01
[Infra] UI - Adding Unit Tests for Coverage
2026-02-05 12:19:03 -08:00
yuneng-jiang
b41876b5f5
Merge pull request #20452 from BerriAI/litellm_non_root_pkglk
[Fix] Non Root Dockerfile: Keep package-lock.json
2026-02-05 12:18:29 -08:00
yuneng-jiang
a04a2de6a1
Merge pull request #20472 from nina-hu/fix/timezone-daily-spend-filtering
fix(ui): adjust daily spend date filtering for user timezone
2026-02-05 12:12:44 -08:00
yuneng-jiang
21d6025c01 adding test for converage 2026-02-05 12:02:12 -08:00
Ishaan Jaffer
ea35ab6d4e day 0 blog post 2026-02-05 11:47:36 -08:00
Peter Dave Hello
dbd4bb5cf4
Add Claude Opus 4.6 (#20508)
Add Claude Opus 4.6 entries for Anthropic, Bedrock Converse, and Vertex AI.

Align pricing and capability metadata with Anthropic docs, including
long-context rates, above-200k prompt-caching rates, prefill removal,
and tool-use system prompt token counts.

Register the Bedrock Converse model ID in constants and add targeted
tests to validate model map values and converse registration.
2026-02-05 11:33:44 -08:00
Ishaan Jaff
eae5fa195e
[Feat] add claude-opus-4-6 to model cost map (#20506)
* add claude-opus-4-6 to model cost map

* fix azure_ai

* new model
2026-02-05 10:59:56 -08:00
Ishaan Jaff
d2bd029fa4
[Fix] 404 Not Found on /api/event_logging/batch endpoint (#20504)
* fix typing

* add event_logging_batch

* TestEventLoggingBatchEndpoint
2026-02-05 10:58:08 -08:00
Krish Dholakia
f65eec8fa8
docs: Update CLI arguments documentation with all available options (#20437)
- Reorganized documentation into logical sections (Server Configuration,
  Server Backend Options, SSL/TLS Configuration, Model Configuration,
  Model Parameters, Database Configuration, Debugging, Testing & Health Checks)
- Added missing CLI arguments:
  - --iam_token_db_auth: RDS IAM token authentication for database connections
  - --run_gunicorn: Start proxy with gunicorn backend
  - --run_hypercorn: Start proxy with hypercorn backend (HTTP/2 support)
  - --ssl_keyfile_path, --ssl_certfile_path, --ciphers: SSL/TLS configuration
  - --keepalive_timeout: Uvicorn keepalive timeout setting
  - --max_requests_before_restart: Worker recycling for memory management
  - --use_prisma_db_push: Alternative database schema update method
  - --test_async, --num_requests: Async endpoint testing
  - --version/-v: Print version
  - --max_budget: Set budget limits for API calls
  - --headers, --add_key, --save: Model configuration options
  - --use_queue: Celery workers for async endpoints
  - --local: Local debugging flag
- Updated default value for --num_workers to reflect actual behavior
- Added environment variable documentation for applicable options

Co-authored-by: Cursor Agent <cursoragent@cursor.com>
2026-02-05 08:46:37 -08:00
Sameer Kankute
9376fcf771
Apply suggestion from @greptile-apps[bot]
Co-authored-by: greptile-apps[bot] <165735046+greptile-apps[bot]@users.noreply.github.com>
2026-02-05 19:53:01 +05:30
Sameer Kankute
eabe3b2e9c fix: mypy issue 2026-02-05 19:50:16 +05:30
Sameer Kankute
b74afec2be
Merge pull request #20398 from BerriAI/litellm_oss_staging_02_04_2026
Litellm oss staging 02 04 2026
2026-02-05 19:42:36 +05:30
Sameer Kankute
26d25de577 fix: make sure gcs_bucket_name passes 2026-02-05 19:42:17 +05:30
Sameer Kankute
df590be7af Fix test_global_redaction_off_with_dynamic_params 2026-02-05 18:26:03 +05:30
Sameer Kankute
34e5bb29a5
Merge pull request #20338 from Lucky-Lodhi2004/ttl-prompt-caching-bedrock
fix #20326 - [Feature]: Support TTL(1h) field in prompt caching for Bedrock Claude 4.5 models
2026-02-05 16:52:12 +05:30
Sameer Kankute
fa733c252e
Merge branch 'litellm_oss_staging_02_04_2026' into ttl-prompt-caching-bedrock 2026-02-05 16:52:02 +05:30
Sameer Kankute
7e8be5f542
Merge branch 'main' into ttl-prompt-caching-bedrock 2026-02-05 16:50:54 +05:30
Sameer Kankute
a21b625a59
Merge pull request #20105 from qiniu/fix/vertex-gemini-streaming-content-filter
Fix Vertex AI Gemini streaming content_filter handling
2026-02-05 16:48:48 +05:30
Sameer Kankute
453d1bd5e1
Merge branch 'main' into litellm_oss_staging_02_04_2026 2026-02-05 12:19:03 +05:30
Sameer Kankute
075b1b7921
Merge pull request #20341 from natimofeev/bugfix/remove-user-messages-merging
bugfix: Disable merging of consecutive user messages for GigaChat provider
2026-02-05 12:18:09 +05:30
nina-hu
5904fa159b fix(ui): adjust daily spend date filtering for user timezone
The daily spend tables store dates in UTC, but the UI sends dates in the
user's local timezone. This causes a mismatch where records from the
user's evening (stored as the next UTC day) don't appear when filtering
by "today".

Changes:
- Add `_adjust_dates_for_timezone()` helper to expand date range based
  on timezone offset
- Add `timezone` query parameter to `/user/daily/activity` and
  `/user/daily/activity/aggregated` endpoints
- Frontend sends `timezone` using `Date.getTimezoneOffset()`

For users west of UTC (e.g., PST), end_date is extended by 1 day.
For users east of UTC (e.g., IST), start_date is extended by 1 day earlier.
This ensures all records within the user's local date range are captured.
2026-02-04 21:42:34 -08:00
Sameer Kankute
1017c3a0e4
Merge pull request #20464 from BerriAI/litellm_cicd_5_feb_2026
Litellm cicd 5 feb 2026
2026-02-05 11:03:35 +05:30
Sameer Kankute
0ce1bfe26a bump: version 1.81.7 → 1.81.8 2026-02-05 10:44:48 +05:30