Sameer Kankute
1ec89b8a04
Feat: add inference_geo based pricing
2026-02-06 13:58:47 +05:30
Sameer Kankute
0934a4ab68
Correct litellm/litellm/llms/anthropic/chat/transformation.py
2026-02-06 13:08:43 +05:30
Sameer Kankute
358a081f63
Add compaction support for vertex ai
2026-02-06 12:52:28 +05:30
Sameer Kankute
1396813d74
The compact beta feature is not currently supported on the Converse and ConverseStream APIs
2026-02-06 12:52:28 +05:30
Sameer Kankute
d0444f402c
Add test for compaction in anthropic
2026-02-06 12:52:28 +05:30
Sameer Kankute
7a473f2954
Apply suggestion from @greptile-apps[bot]
...
Co-authored-by: greptile-apps[bot] <165735046+greptile-apps[bot]@users.noreply.github.com>
2026-02-06 12:07:13 +05:30
Sameer Kankute
ea518a7684
Apply suggestion from @greptile-apps[bot]
...
Co-authored-by: greptile-apps[bot] <165735046+greptile-apps[bot]@users.noreply.github.com>
2026-02-06 12:06:50 +05:30
Sameer Kankute
887a977ab4
Add doc on how to enable compaction via chat completion
2026-02-06 12:05:36 +05:30
Sameer Kankute
24dda99bd7
Handle compaction block in the input request
2026-02-06 11:55:12 +05:30
Sameer Kankute
c03ba8394e
Add compaction block in provider spcific fields streaming+ non streaming
2026-02-06 11:54:47 +05:30
Sameer Kankute
039b37fac1
Add compaction type block in the output
2026-02-06 11:31:35 +05:30
Sameer Kankute
1bcd407af6
Add adaptive thiking for bedrock converse
2026-02-06 09:40:16 +05:30
Sameer Kankute
f15dd691b4
Fix anthropic.claude-opus-4-6-v1 for bedrock
2026-02-06 09:28:50 +05:30
Sameer Kankute
186fd2e64e
Add adaptive thinking support for anthropic opus 4.6
2026-02-06 09:24:43 +05:30
Sameer Kankute
3923ef2857
Merge pull request #20491 from BerriAI/litellm_gemini_files_gcs
...
fix: make sure gcs_bucket_name passes
2026-02-06 08:26:14 +05:30
yuneng-jiang
3bececfa7a
Merge pull request #20539 from BerriAI/litellm_usage_failed_req
...
[Refactor] UI - Usage Page: Spend By Provider
2026-02-05 16:39:55 -08:00
yuneng-jiang
f64f949715
Spend by provider refactor
2026-02-05 16:22:07 -08:00
Shivam Rawat
93cf1ef517
Merge pull request #20535 from BerriAI/litellm_move_anthropic_test_script
...
[Chore] Move anthropic input/output test to right folder
2026-02-05 16:12:08 -08:00
yuneng-jiang
523a36ed53
Merge pull request #20530 from BerriAI/litellm_ui_team_soft_budget
...
[Feature] Add soft_budget to Team Table + Create/Update Endpoints
2026-02-05 15:52:17 -08:00
shivam
3dc70c3398
moved the test anthropic file
2026-02-05 15:44:24 -08:00
Ishaan Jaffer
bff5f8d7b0
doc fix
2026-02-05 15:31:41 -08:00
yuneng-jiang
45c6382667
cz bump + builds
2026-02-05 14:47:11 -08:00
yuneng-jiang
8b0d0b2610
bump: version 0.4.30 → 0.4.31
2026-02-05 14:46:38 -08:00
yuneng-jiang
a17efa1c8e
Add soft_budget to team table and create update endpoints
2026-02-05 14:43:48 -08:00
Ishaan Jaff
887a907e42
[Fix] Guardrails API - Ensure OpenAI Moderations Guard works with OpenAI Embeddings ( #20523 )
...
* init OpenAIEmbeddingsHandler
* init apply_guardrail
* use apply guardrails for OpenAI moderations
* test_embeddings_handler_string_input
* test_openai_moderation_guardrail_apply_guardrail
* fix typing
* test_openai_moderation_responses_api_input_field
* test fixes
2026-02-05 14:40:15 -08:00
shin-bot-litellm
0649720f79
Fix test isolation for test_log_langfuse_v2_handles_null_usage_values ( #20475 )
...
Create fresh mock objects within the test instead of reusing mocks
from setUp that have side_effect configured. The setUp's side_effect
on mock_langfuse_client.trace can interfere with return_value settings
when tests try to reset and reconfigure mocks.
Using dedicated mock objects for this test avoids state pollution
from setUp's side_effect configuration and makes the test more
deterministic in parallel execution environments.
2026-02-05 12:57:41 -08:00
shin-bot-litellm
1d92968e17
Fix test isolation for test_watsonx_gpt_oss_prompt_transformation ( #20474 )
...
Set cached tokenizer config directly and mock both sync and async
tokenizer functions to avoid race conditions when running with
parallel test execution (-n 16).
The issue was that parallel tests could populate the
litellm.known_tokenizer_config cache between clearing it and
when the code checked it. This caused the sync code path to be
used instead of the async path, bypassing the mocked async functions.
Fix:
1. Set cache directly instead of clearing it
2. Also mock sync versions _get_tokenizer_config and _get_chat_template_file
This ensures the test is deterministic regardless of test execution order.
2026-02-05 12:57:20 -08:00
yuneng-jiang
b5956cb020
Merge pull request #20513 from BerriAI/litellm_ui_test_cov_01
...
[Infra] UI - Adding Unit Tests for Coverage
2026-02-05 12:19:03 -08:00
yuneng-jiang
b41876b5f5
Merge pull request #20452 from BerriAI/litellm_non_root_pkglk
...
[Fix] Non Root Dockerfile: Keep package-lock.json
2026-02-05 12:18:29 -08:00
yuneng-jiang
a04a2de6a1
Merge pull request #20472 from nina-hu/fix/timezone-daily-spend-filtering
...
fix(ui): adjust daily spend date filtering for user timezone
2026-02-05 12:12:44 -08:00
yuneng-jiang
21d6025c01
adding test for converage
2026-02-05 12:02:12 -08:00
Ishaan Jaffer
ea35ab6d4e
day 0 blog post
2026-02-05 11:47:36 -08:00
Peter Dave Hello
dbd4bb5cf4
Add Claude Opus 4.6 ( #20508 )
...
Add Claude Opus 4.6 entries for Anthropic, Bedrock Converse, and Vertex AI.
Align pricing and capability metadata with Anthropic docs, including
long-context rates, above-200k prompt-caching rates, prefill removal,
and tool-use system prompt token counts.
Register the Bedrock Converse model ID in constants and add targeted
tests to validate model map values and converse registration.
2026-02-05 11:33:44 -08:00
Ishaan Jaff
eae5fa195e
[Feat] add claude-opus-4-6 to model cost map ( #20506 )
...
* add claude-opus-4-6 to model cost map
* fix azure_ai
* new model
2026-02-05 10:59:56 -08:00
Ishaan Jaff
d2bd029fa4
[Fix] 404 Not Found on /api/event_logging/batch endpoint ( #20504 )
...
* fix typing
* add event_logging_batch
* TestEventLoggingBatchEndpoint
2026-02-05 10:58:08 -08:00
Krish Dholakia
f65eec8fa8
docs: Update CLI arguments documentation with all available options ( #20437 )
...
- Reorganized documentation into logical sections (Server Configuration,
Server Backend Options, SSL/TLS Configuration, Model Configuration,
Model Parameters, Database Configuration, Debugging, Testing & Health Checks)
- Added missing CLI arguments:
- --iam_token_db_auth: RDS IAM token authentication for database connections
- --run_gunicorn: Start proxy with gunicorn backend
- --run_hypercorn: Start proxy with hypercorn backend (HTTP/2 support)
- --ssl_keyfile_path, --ssl_certfile_path, --ciphers: SSL/TLS configuration
- --keepalive_timeout: Uvicorn keepalive timeout setting
- --max_requests_before_restart: Worker recycling for memory management
- --use_prisma_db_push: Alternative database schema update method
- --test_async, --num_requests: Async endpoint testing
- --version/-v: Print version
- --max_budget: Set budget limits for API calls
- --headers, --add_key, --save: Model configuration options
- --use_queue: Celery workers for async endpoints
- --local: Local debugging flag
- Updated default value for --num_workers to reflect actual behavior
- Added environment variable documentation for applicable options
Co-authored-by: Cursor Agent <cursoragent@cursor.com>
2026-02-05 08:46:37 -08:00
Sameer Kankute
9376fcf771
Apply suggestion from @greptile-apps[bot]
...
Co-authored-by: greptile-apps[bot] <165735046+greptile-apps[bot]@users.noreply.github.com>
2026-02-05 19:53:01 +05:30
Sameer Kankute
eabe3b2e9c
fix: mypy issue
2026-02-05 19:50:16 +05:30
Sameer Kankute
b74afec2be
Merge pull request #20398 from BerriAI/litellm_oss_staging_02_04_2026
...
Litellm oss staging 02 04 2026
2026-02-05 19:42:36 +05:30
Sameer Kankute
26d25de577
fix: make sure gcs_bucket_name passes
2026-02-05 19:42:17 +05:30
Sameer Kankute
df590be7af
Fix test_global_redaction_off_with_dynamic_params
2026-02-05 18:26:03 +05:30
Sameer Kankute
34e5bb29a5
Merge pull request #20338 from Lucky-Lodhi2004/ttl-prompt-caching-bedrock
...
fix #20326 - [Feature]: Support TTL(1h) field in prompt caching for Bedrock Claude 4.5 models
2026-02-05 16:52:12 +05:30
Sameer Kankute
fa733c252e
Merge branch 'litellm_oss_staging_02_04_2026' into ttl-prompt-caching-bedrock
2026-02-05 16:52:02 +05:30
Sameer Kankute
7e8be5f542
Merge branch 'main' into ttl-prompt-caching-bedrock
2026-02-05 16:50:54 +05:30
Sameer Kankute
a21b625a59
Merge pull request #20105 from qiniu/fix/vertex-gemini-streaming-content-filter
...
Fix Vertex AI Gemini streaming content_filter handling
2026-02-05 16:48:48 +05:30
Sameer Kankute
453d1bd5e1
Merge branch 'main' into litellm_oss_staging_02_04_2026
2026-02-05 12:19:03 +05:30
Sameer Kankute
075b1b7921
Merge pull request #20341 from natimofeev/bugfix/remove-user-messages-merging
...
bugfix: Disable merging of consecutive user messages for GigaChat provider
2026-02-05 12:18:09 +05:30
nina-hu
5904fa159b
fix(ui): adjust daily spend date filtering for user timezone
...
The daily spend tables store dates in UTC, but the UI sends dates in the
user's local timezone. This causes a mismatch where records from the
user's evening (stored as the next UTC day) don't appear when filtering
by "today".
Changes:
- Add `_adjust_dates_for_timezone()` helper to expand date range based
on timezone offset
- Add `timezone` query parameter to `/user/daily/activity` and
`/user/daily/activity/aggregated` endpoints
- Frontend sends `timezone` using `Date.getTimezoneOffset()`
For users west of UTC (e.g., PST), end_date is extended by 1 day.
For users east of UTC (e.g., IST), start_date is extended by 1 day earlier.
This ensures all records within the user's local date range are captured.
2026-02-04 21:42:34 -08:00
Sameer Kankute
1017c3a0e4
Merge pull request #20464 from BerriAI/litellm_cicd_5_feb_2026
...
Litellm cicd 5 feb 2026
2026-02-05 11:03:35 +05:30
Sameer Kankute
0ce1bfe26a
bump: version 1.81.7 → 1.81.8
2026-02-05 10:44:48 +05:30