Ishaan Jaff
|
fd424a387c
|
Merge pull request #2913 from BerriAI/litellm_add_voyage_2
[New Models] Add Voyage 2 embedding models
|
2024-04-09 07:47:18 -07:00 |
|
Ishaan Jaff
|
9dcf671b55
|
docs- add voyage 2
|
2024-04-09 07:43:30 -07:00 |
|
CLARKBENHAM
|
e96d97d9e5
|
remove formating changes
|
2024-04-08 21:31:21 -07:00 |
|
CLARKBENHAM
|
6e20bb13b2
|
Revert "doc pre_call_check: enables router rate limits for concurrent calls"
This reverts commit 886c859519.
|
2024-04-08 21:27:38 -07:00 |
|
CLARKBENHAM
|
886c859519
|
doc pre_call_check: enables router rate limits for concurrent calls
|
2024-04-08 21:20:59 -07:00 |
|
Krrish Dholakia
|
b6cd200676
|
fix(llm_guard.py): enable request-specific llm guard flag
|
2024-04-08 21:15:33 -07:00 |
|
Krrish Dholakia
|
6dcf109a42
|
docs(demo.md): fix iframe link
|
2024-04-08 15:03:11 -07:00 |
|
Krrish Dholakia
|
009f548079
|
fix(proxy_server.py): allow /model/new feature flag to work via env
|
2024-04-08 14:57:19 -07:00 |
|
Krrish Dholakia
|
d099591a09
|
docs(sidebars.js): refactor ordering
|
2024-04-08 07:30:08 -07:00 |
|
Ishaan Jaff
|
3b6b497672
|
Merge pull request #2882 from BerriAI/litellm_docs_fix
docs fix gpt-3.5-turbo-instruct-0914
|
2024-04-06 20:26:24 -07:00 |
|
Ishaan Jaff
|
a2c63075ef
|
Merge pull request #2877 from BerriAI/litellm_fix_text_completion
[Feat] Text-Completion-OpenAI - Re-use OpenAI Client
|
2024-04-06 12:15:52 -07:00 |
|
Ishaan Jaff
|
c2f978fd5a
|
(docs) use text completion with litellm proxy
|
2024-04-06 12:07:20 -07:00 |
|
Ishaan Jaff
|
2cc364743c
|
docs fix gpt-3.5-turbo-instruct-0914
|
2024-04-06 09:12:01 -07:00 |
|
Krrish Dholakia
|
afa2e2eba9
|
docs(anthropic.md): update anthropic docs to show 'tool' param usage
|
2024-04-06 08:35:43 -07:00 |
|
Krrish Dholakia
|
dce96478c7
|
docs(vertex.md): add claude 3 on vertex ai to docs
|
2024-04-05 21:48:42 -07:00 |
|
Ishaan Jaff
|
faa0d38087
|
Merge pull request #2868 from BerriAI/litellm_add_command_r_on_proxy
Add Azure Command-r-plus on litellm proxy
|
2024-04-05 15:13:47 -07:00 |
|
Ishaan Jaff
|
2174b240d8
|
Merge pull request #2861 from BerriAI/litellm_add_azure_command_r_plust
[FEAT] add azure command-r-plus
|
2024-04-05 15:13:35 -07:00 |
|
Ishaan Jaff
|
26f5823c4e
|
docs - use azure command r
|
2024-04-05 14:55:28 -07:00 |
|
Ishaan Jaff
|
8df13306ea
|
docs call command-r through proxy
|
2024-04-05 14:22:29 -07:00 |
|
Krrish Dholakia
|
834f363b6a
|
docs(vertex.md): add safety settings tutorial to docs
|
2024-04-05 13:55:05 -07:00 |
|
Ishaan Jaff
|
22ac95b834
|
docs azure_ai command r
|
2024-04-05 13:50:56 -07:00 |
|
Ishaan Jaff
|
6b9c04618e
|
fix use azure_ai/mistral
|
2024-04-05 10:07:43 -07:00 |
|
Ishaan Jaff
|
ab60d7c8fb
|
docs azure ai command-r plust
|
2024-04-05 09:24:27 -07:00 |
|
Ishaan Jaff
|
cfe358abaa
|
simplify calling azure/commmand-r-plus
|
2024-04-05 09:18:11 -07:00 |
|
Ishaan Jaff
|
b25db0443a
|
docs - using command r on azure
|
2024-04-05 09:04:38 -07:00 |
|
Krish Dholakia
|
4ce8227e70
|
Merge pull request #2841 from Manouchehri/nuke-gemini-1.5-pro-vision
Fix: Remove non-existent gemini-1.5-pro-vision model.
|
2024-04-05 07:03:38 -07:00 |
|
Krrish Dholakia
|
b0d80de14d
|
docs(vertex.md): fix import routes
|
2024-04-04 21:32:44 -07:00 |
|
Krrish Dholakia
|
003cd3b102
|
docs(vertex.md): add tutorial for using vertex ai with gcp service account
|
2024-04-04 21:28:28 -07:00 |
|
Ishaan Jaff
|
d313f5bd61
|
Merge pull request #2847 from themrzmaster/feat/add_command_r_plus
Add command-r-plus
|
2024-04-04 21:06:42 -07:00 |
|
Ishaan Jaff
|
12e5118367
|
Merge pull request #2846 from BerriAI/litellm_docs_delete_cache_keys
docs - `delete` cache keys
|
2024-04-04 14:07:50 -07:00 |
|
Krrish Dholakia
|
4dbb46cf42
|
docs(vertex.md): add docs on setting google_application_credentials
|
2024-04-04 13:49:03 -07:00 |
|
lucca
|
be265fbb15
|
initial
|
2024-04-04 16:58:51 -03:00 |
|
Ishaan Jaff
|
9e9b617934
|
docs - delete cache keys
|
2024-04-04 12:20:14 -07:00 |
|
David Manouchehri
|
6044045b91
|
Fix: Remove non-existent gemini-1.5-pro-vision model.
The gemini-1.5-pro model handles both text and vision.
|
2024-04-04 17:33:08 +00:00 |
|
Krrish Dholakia
|
cbe4aa386b
|
docs(token_auth.md): update links
|
2024-04-03 13:23:30 -07:00 |
|
Krrish Dholakia
|
06b7d2608e
|
docs(token_auth.md): update docs
|
2024-04-03 13:21:25 -07:00 |
|
ryanwclark1
|
62caa7e858
|
Update gpt-4-turbo-preview pricing and context. Included in docs.
|
2024-04-03 10:48:14 -05:00 |
|
Ishaan Jaff
|
91269257f2
|
(docs) openai wildcard models
|
2024-04-01 19:53:34 -07:00 |
|
Krrish Dholakia
|
c52819d47c
|
fix(proxy_server.py): don't require scope for team-based jwt access
If team with the client_id exists then it should be allowed to make a request, if it doesn't then as we discussed it should return an error
|
2024-04-01 18:52:00 -07:00 |
|
Krrish Dholakia
|
cdae08f3c3
|
docs(openai.md): fix docs to include example of calling openai on proxy
|
2024-04-01 12:09:22 -07:00 |
|
Krrish Dholakia
|
a917fadf45
|
docs(routing.md): refactor docs to show how to use pre-call checks and fallback across model groups
|
2024-04-01 11:21:27 -07:00 |
|
Ishaan Jaff
|
18fec3ad8e
|
Merge pull request #2779 from DaxServer/update-proxy-dockerfile-branch
fix(docs): Correct Docker pull command in deploy.md
|
2024-04-01 07:10:45 -07:00 |
|
DaxServer
|
28f6caa04c
|
fix(docs): Correct Docker pull command in deploy.md
Corrected the Docker pull command in deploy.md to remove duplicated 'docker pull' command.
|
2024-03-31 20:10:00 +02:00 |
|
DaxServer
|
61b6f8be44
|
docs: Update references to Ollama repository url
Updated references to the Ollama repository URL from https://github.com/jmorganca/ollama to https://github.com/ollama/ollama.
|
2024-03-31 19:35:37 +02:00 |
|
Krrish Dholakia
|
a7aa6fae64
|
docs(deploy.md): fix docs for litlelm-database docker run example
|
2024-03-30 20:08:27 -07:00 |
|
Krish Dholakia
|
2ca303ec0e
|
Merge pull request #2748 from BerriAI/litellm_anthropic_tool_calling_list_parsing_fix
fix(factory.py): parse list in xml tool calling response (anthropic)
|
2024-03-30 11:27:02 -07:00 |
|
Krrish Dholakia
|
4826018756
|
docs(users.md): fix doc for end-user param
|
2024-03-29 21:54:07 -07:00 |
|
Vincelwt
|
1b84dfac91
|
Merge branch 'main' into main
|
2024-03-30 13:21:53 +09:00 |
|
Ishaan Jaff
|
e8ead49d29
|
Merge pull request #2628 from BerriAI/dependabot/npm_and_yarn/docs/my-website/webpack-dev-middleware-5.3.4
build(deps): bump webpack-dev-middleware from 5.3.3 to 5.3.4 in /docs/my-website
|
2024-03-29 16:12:29 -07:00 |
|
Ishaan Jaff
|
2974e0da31
|
Merge pull request #2689 from BerriAI/dependabot/npm_and_yarn/docs/my-website/express-4.19.2
build(deps): bump express from 4.18.2 to 4.19.2 in /docs/my-website
|
2024-03-29 16:12:17 -07:00 |
|
Ishaan Jaff
|
a78ed81cd9
|
(docs) grafana metrics
|
2024-03-29 14:38:37 -07:00 |
|
Ishaan Jaff
|
24570bc075
|
(docs) grafana / prometheus
|
2024-03-29 14:25:45 -07:00 |
|
Ishaan Jaff
|
c2283235a1
|
(docs) /metrics endpoint
|
2024-03-29 13:36:24 -07:00 |
|
Ishaan Jaff
|
ffa29ddfef
|
(docs) cleanup
|
2024-03-29 13:10:26 -07:00 |
|
Krrish Dholakia
|
81b4f47140
|
docs: show how tool calling parsing works + how to get raw model response
|
2024-03-29 11:58:49 -07:00 |
|
Krrish Dholakia
|
d547944556
|
fix(sagemaker.py): support 'model_id' param for sagemaker
allow passing inference component param to sagemaker in the same format as we handle this for bedrock
|
2024-03-29 08:43:17 -07:00 |
|
Krrish Dholakia
|
cdb940d504
|
docs(prod.md): update prod docs with batch writing info
|
2024-03-28 23:42:43 -07:00 |
|
Krrish Dholakia
|
85a5291142
|
docs(prod.md): doc improvements
|
2024-03-28 19:04:24 -07:00 |
|
Krrish Dholakia
|
da7a00d6d2
|
docs(prod.md): fix docker run commands
|
2024-03-28 18:51:53 -07:00 |
|
Krrish Dholakia
|
7c44b32cc2
|
refactor(proxy/utils.py): add more debug logs
|
2024-03-28 18:44:35 -07:00 |
|
Krrish Dholakia
|
eb318afe52
|
docs(prod.md): cleanup doc
|
2024-03-28 18:34:09 -07:00 |
|
Krrish Dholakia
|
ced902f822
|
docs(prod.md): improve docs
|
2024-03-28 15:35:07 -07:00 |
|
Krrish Dholakia
|
eb3806feba
|
docs(prod.md): update docs with litellm spend logs server machine spec
|
2024-03-28 15:26:26 -07:00 |
|
Krrish Dholakia
|
c15df27c1e
|
docs(prod.md): add litellm spend logs server to docs
|
2024-03-28 15:15:10 -07:00 |
|
Krish Dholakia
|
934a9ac2b4
|
Merge pull request #2722 from BerriAI/litellm_db_perf_improvement
feat(proxy/utils.py): enable updating db in a separate server
|
2024-03-28 14:56:14 -07:00 |
|
Krrish Dholakia
|
a09818e72e
|
build(ghcr_deploy.yml): deploy spend logs server docker image
make it easy for user to deploy a separate spend logs server
|
2024-03-28 13:39:52 -07:00 |
|
Krrish Dholakia
|
746cd3da11
|
docs(gemini.md): add link to google ai studio api key
|
2024-03-28 10:12:59 -07:00 |
|
Ishaan Jaff
|
5e5f3d5fd2
|
Merge pull request #2729 from BerriAI/litellm_show_better_error_msg_with_role
(fix) show user their role when rejecting /team/new requests
|
2024-03-28 07:42:31 -07:00 |
|
Ishaan Jaff
|
d5f6fe4eff
|
(docs) update UI
|
2024-03-27 22:24:56 -07:00 |
|
Krrish Dholakia
|
526aa9230f
|
docs(call_hooks.md): show result in docs
|
2024-03-27 21:04:51 -07:00 |
|
Krrish Dholakia
|
e6b929fff3
|
docs(call_hooks.md): show admin how to enforce user param
|
2024-03-27 20:58:26 -07:00 |
|
Krrish Dholakia
|
d08da5b05a
|
docs(instructor.md): improve default example
|
2024-03-27 12:51:05 -07:00 |
|
Krrish Dholakia
|
90b859ebcb
|
docs(token_auth.md): cleanup docs
|
2024-03-26 21:42:07 -07:00 |
|
Krrish Dholakia
|
282176c502
|
docs(token_auth.md): update docs
|
2024-03-26 21:41:08 -07:00 |
|
Krrish Dholakia
|
ca84e7a8e8
|
docs(token_auth.md): update jwt auth docs with new info
|
2024-03-26 21:33:03 -07:00 |
|
Krrish Dholakia
|
bf7cc943fb
|
docs(enterprise.md): update docs to turn on/off llm guard per key
|
2024-03-26 18:02:44 -07:00 |
|
Krish Dholakia
|
0ab708e6f1
|
Merge pull request #2704 from BerriAI/litellm_jwt_auth_improvements_3
fix(handle_jwt.py): enable team-based jwt-auth access
|
2024-03-26 16:06:56 -07:00 |
|
Krrish Dholakia
|
752516df1b
|
fix(handle_jwt.py): support public key caching ttl param
|
2024-03-26 14:32:55 -07:00 |
|
Ishaan Jaff
|
da503eab18
|
Merge branch 'main' into litellm_remove_litellm_telemetry
|
2024-03-26 11:35:02 -07:00 |
|
Ishaan Jaff
|
4d81df3d6f
|
(docs) switch of litellm telemetry
|
2024-03-26 11:19:55 -07:00 |
|
Ishaan Jaff
|
995c379a63
|
(fix) prod.md
|
2024-03-25 22:30:22 -07:00 |
|
Krrish Dholakia
|
16ade7e556
|
docs(proxy/caching.md): add ttl param to proxy/caching.md
|
2024-03-25 13:46:52 -07:00 |
|
dependabot[bot]
|
9323f1439f
|
build(deps): bump express from 4.18.2 to 4.19.2 in /docs/my-website
Bumps [express](https://github.com/expressjs/express) from 4.18.2 to 4.19.2.
- [Release notes](https://github.com/expressjs/express/releases)
- [Changelog](https://github.com/expressjs/express/blob/master/History.md)
- [Commits](https://github.com/expressjs/express/compare/4.18.2...4.19.2)
---
updated-dependencies:
- dependency-name: express
dependency-type: indirect
...
Signed-off-by: dependabot[bot] <support@github.com>
|
2024-03-25 20:31:36 +00:00 |
|
Krrish Dholakia
|
03b8444d3c
|
docs(token_auth.md): add renaming jwt scope string to docs
|
2024-03-25 12:49:44 -07:00 |
|
Krrish Dholakia
|
53695943e3
|
docs(instructor.md): tutorial on using litellm with instructor
|
2024-03-25 08:35:11 -07:00 |
|
Krrish Dholakia
|
9e9de7f6e2
|
docs(routing.md): add fallbacks being done in order
|
2024-03-24 12:13:19 -07:00 |
|
Krrish Dholakia
|
1c60fd0e78
|
docs(routing.md): add url
|
2024-03-23 20:03:42 -07:00 |
|
Krrish Dholakia
|
7c74ea8b77
|
docs(routing.md): add proxy example to pre-call checks in routing docs
|
2024-03-23 20:00:50 -07:00 |
|
Ishaan Jaff
|
92759ef055
|
Merge pull request #2670 from BerriAI/litellm_docs_best_practices_prod
(docs) best prod practices
|
2024-03-23 19:38:22 -07:00 |
|
Krish Dholakia
|
c92fa1af7c
|
Merge pull request #2669 from BerriAI/litellm_router_pre_call_checks
feat(router.py): enable pre-call checks
|
2024-03-23 19:38:09 -07:00 |
|
Ishaan Jaff
|
09992a6122
|
(docs) prod best perf
|
2024-03-23 19:36:26 -07:00 |
|
Ishaan Jaff
|
d04b4dea3e
|
(docs) best prod practices
|
2024-03-23 19:29:21 -07:00 |
|
Krrish Dholakia
|
e8e7964025
|
docs(routing.md): add pre-call checks to docs
|
2024-03-23 19:10:34 -07:00 |
|
Ishaan Jaff
|
2ae489c506
|
(docs) update config set_verbose
|
2024-03-23 18:54:31 -07:00 |
|
Ishaan Jaff
|
04a09830de
|
Merge pull request #2668 from BerriAI/litellm_update_deploy_docs
[Docs] Add Docs on deploying to EKS Cluster + K8
|
2024-03-23 18:42:36 -07:00 |
|
Ishaan Jaff
|
30ae52c21e
|
(docs) using litellm on EKS
|
2024-03-23 17:49:00 -07:00 |
|
Ishaan Jaff
|
f646a4612b
|
Merge pull request #2667 from BerriAI/litellm_update_gunicorn_instructions
[Docs] update gunicorn instructions - Uvicorn perf is significantly better on K8s
|
2024-03-23 17:43:09 -07:00 |
|
Ishaan Jaff
|
61d2e91632
|
(docs) update gunicorn usage
|
2024-03-23 17:39:07 -07:00 |
|
Vivek Aditya
|
efc90b04c7
|
minor fix
|
2024-03-23 12:50:46 +05:30 |
|
Vivek Aditya
|
6bd49c6087
|
Athina docs updated with information about additional fields and a minor fix in the callback
|
2024-03-23 12:42:07 +05:30 |
|
Krrish Dholakia
|
265dd5cd4f
|
docs(token_auth.md): add project based auth to docs
|
2024-03-22 17:27:40 -07:00 |
|
Krrish Dholakia
|
d06b9a5a47
|
fix(proxy_server.py): enable jwt-auth for users
allow a user to auth into the proxy via jwt's and call allowed routes
|
2024-03-22 17:08:10 -07:00 |
|
Krrish Dholakia
|
858fa07e07
|
docs(call_hooks.md): fix dead link
|
2024-03-22 09:06:01 -07:00 |
|
Krrish Dholakia
|
211a6887f3
|
docs(enterprise.md): fix llm guard api link
|
2024-03-22 09:03:11 -07:00 |
|
Krrish Dholakia
|
566d48d51b
|
docs(prompt_injection.md): fix dead link on docs
|
2024-03-22 08:24:47 -07:00 |
|
Krrish Dholakia
|
66e7345296
|
docs(virtual_keys.md): simplify virtual keys docs
|
2024-03-21 21:49:50 -07:00 |
|
Krrish Dholakia
|
425165dda9
|
docs(gemini.md): fix string for calling gemini 1.5
|
2024-03-21 18:04:11 -07:00 |
|
Krrish Dholakia
|
abe8d7c921
|
docs(configs.md): add disable swagger ui env tutorial to docs
|
2024-03-21 17:16:52 -07:00 |
|
dependabot[bot]
|
ad1054520c
|
build(deps): bump webpack-dev-middleware in /docs/my-website
Bumps [webpack-dev-middleware](https://github.com/webpack/webpack-dev-middleware) from 5.3.3 to 5.3.4.
- [Release notes](https://github.com/webpack/webpack-dev-middleware/releases)
- [Changelog](https://github.com/webpack/webpack-dev-middleware/blob/v5.3.4/CHANGELOG.md)
- [Commits](https://github.com/webpack/webpack-dev-middleware/compare/v5.3.3...v5.3.4)
---
updated-dependencies:
- dependency-name: webpack-dev-middleware
dependency-type: indirect
...
Signed-off-by: dependabot[bot] <support@github.com>
|
2024-03-21 23:56:48 +00:00 |
|
Krish Dholakia
|
33a433eb0a
|
Merge branch 'main' into litellm_llm_api_prompt_injection_check
|
2024-03-21 09:57:10 -07:00 |
|
Krrish Dholakia
|
0521e8a1d9
|
fix(prompt_injection_detection.py): fix type check
|
2024-03-21 08:56:13 -07:00 |
|
Vincelwt
|
29e8c144fb
|
Merge branch 'main' into main
|
2024-03-22 00:52:42 +09:00 |
|
Ishaan Jaff
|
8363bd4e7e
|
Merge pull request #2613 from BerriAI/litellm_fix_quick_start_docker
(docs) Litellm fix quick start docker
|
2024-03-20 21:34:19 -07:00 |
|
Ishaan Jaff
|
354d13c7e1
|
(docs) fix litellm docker quick start
|
2024-03-20 21:34:01 -07:00 |
|
Ishaan Jaff
|
3612e47d6b
|
(fix) clean up quick start
|
2024-03-20 21:24:42 -07:00 |
|
Ishaan Jaff
|
ee1a0a9409
|
(docs) add example using openai compatib model
|
2024-03-20 21:15:50 -07:00 |
|
Ishaan Jaff
|
f7945a945b
|
(docs) add example using vertex ai on litellm proxy
|
2024-03-20 21:07:55 -07:00 |
|
Ishaan Jaff
|
5842486810
|
(docs) using /cache/ping
|
2024-03-20 17:03:26 -07:00 |
|
Ishaan Jaff
|
c068a5fb74
|
(docs) deploying litellm
|
2024-03-20 11:47:55 -07:00 |
|
Krrish Dholakia
|
ca970a90c4
|
fix(handle_jwt.py): remove issuer check
|
2024-03-20 08:35:23 -07:00 |
|
Krrish Dholakia
|
88733fda5d
|
docs(prompt_injection.md): open sourcing prompt injection detection
|
2024-03-19 22:48:52 -07:00 |
|
Krrish Dholakia
|
f25b03326b
|
fix(proxy_server.py): allow user to disable scheduled reset budget task
|
2024-03-19 20:36:22 -07:00 |
|
Krish Dholakia
|
5171d7689f
|
Merge pull request #2592 from BerriAI/litellm_jwt_auth
feat(handle_jwt.py): support authenticating admins into the proxy via jwt's
|
2024-03-19 17:54:54 -07:00 |
|
Krrish Dholakia
|
3ed3d89a06
|
docs: doc cleanup
|
2024-03-19 17:50:31 -07:00 |
|
Krrish Dholakia
|
0c1d754e2c
|
docs(cost_tracking.md): cleanup docs
|
2024-03-19 17:40:52 -07:00 |
|
Krrish Dholakia
|
191c2256f6
|
docs(deploy.md): fix deploy quick start command
|
2024-03-19 17:35:08 -07:00 |
|
Krrish Dholakia
|
a8d3d51d21
|
docs(token_based_auth.md): add jwt auth to docs
|
2024-03-19 16:34:27 -07:00 |
|
Ishaan Jaff
|
36440f6b4a
|
Merge pull request #2521 from BerriAI/dependabot/npm_and_yarn/docs/my-website/follow-redirects-1.15.6
build(deps): bump follow-redirects from 1.15.4 to 1.15.6 in /docs/my-website
|
2024-03-19 13:13:55 -07:00 |
|
Krrish Dholakia
|
7c74a0e6e2
|
fix(proxy_server.py): expose disable_spend_logs flag in config general settings
Writing each spend log adds +300ms latency
https://github.com/BerriAI/litellm/issues/1714#issuecomment-1924727281
|
2024-03-19 12:08:37 -07:00 |
|
Krrish Dholakia
|
c63259ebcd
|
docs(deploy.md): update docker start command
|
2024-03-19 07:58:08 -07:00 |
|
Vincelwt
|
1cbfd312fe
|
Merge branch 'main' into main
|
2024-03-19 12:50:04 +09:00 |
|
Ishaan Jaff
|
038c9d5781
|
(docs) litellm + datadog
|
2024-03-18 17:06:00 -07:00 |
|
Ishaan Jaff
|
4a4c322278
|
(docs) easily find call hooks
|
2024-03-18 07:38:29 -07:00 |
|
Krrish Dholakia
|
136a58b84a
|
docs(secret.md): add aws secret manager to docs
|
2024-03-16 18:47:56 -07:00 |
|
Ishaan Jaff
|
27cd012c2d
|
(docs) litellm hel
|
2024-03-16 12:14:35 -07:00 |
|
Ishaan Jaff
|
c17e721278
|
(docs) update litellm helm docs
|
2024-03-16 12:01:38 -07:00 |
|
Krrish Dholakia
|
2c2db9ce89
|
fix(proxy_server.py): bug fix on getting user obj from cache
|
2024-03-16 11:07:38 -07:00 |
|
Krrish Dholakia
|
2d2731c3b5
|
docs(caching.md): add batch redis requests to docs
|
2024-03-15 23:01:08 -07:00 |
|
ishaan-jaff
|
92b198f6c5
|
(docs) litellm + helm chart
|
2024-03-15 17:04:51 -07:00 |
|
Ishaan Jaff
|
8cfa0b64ce
|
Merge pull request #2541 from udit-001/docs/chatlitellm-langfuse
docs(langfuse): add chatlitellm section
|
2024-03-15 16:32:01 -07:00 |
|
ishaan-jaff
|
3108c91ebd
|
(docs) using litellm + helm charts
|
2024-03-15 16:20:26 -07:00 |
|
ishaan-jaff
|
4a33c53619
|
(fix) docs litellm helm chart
|
2024-03-15 16:07:43 -07:00 |
|
Udit
|
4a232f4ab3
|
docs(langfuse): update langfuse casing
|
2024-03-16 03:34:47 +05:30 |
|
Udit
|
b8dbcd7ac3
|
docs(langfuse): update section titles
|
2024-03-16 03:33:14 +05:30 |
|
Udit
|
31fb2d0219
|
docs(langfuse): fix missing litellm import
|
2024-03-16 03:30:01 +05:30 |
|
Udit
|
1220eb3c7a
|
docs(langfuse): update chatlitellm section
|
2024-03-16 03:28:19 +05:30 |
|
Udit
|
acd56b174c
|
docs(langfuse): add chatlitellm section
|
2024-03-16 03:24:07 +05:30 |
|
Krrish Dholakia
|
e033e84720
|
docs(cohere.md): fix model name in cohere docs
|
2024-03-15 10:08:07 -07:00 |
|
Krish Dholakia
|
32ca306123
|
Merge pull request #2535 from BerriAI/litellm_fireworks_ai_support
feat(utils.py): add native fireworks ai support
|
2024-03-15 10:02:53 -07:00 |
|
Krrish Dholakia
|
5edf414a5f
|
docs(audio_transcription.md): add openai sdk usage example to audio transcription docs
|
2024-03-15 09:54:03 -07:00 |
|
Krrish Dholakia
|
fa0c8b7be6
|
docs(fireworks_ai.md): add fireworks ai to docs
|
2024-03-15 09:17:15 -07:00 |
|
ishaan-jaff
|
f07a652148
|
(docs) load test proxy
|
2024-03-15 08:10:45 -07:00 |
|
Ishaan Jaff
|
4c834714ab
|
Merge pull request #2529 from snekkenull/main
(feat) add groq/gemma-7b-it
|
2024-03-15 07:41:44 -07:00 |
|
ishaan-jaff
|
91a47dc17a
|
(docs) how to run a locust load test
|
2024-03-15 07:37:50 -07:00 |
|
USAGI
|
a9634b717c
|
Add groq/gemma-7b-it
|
2024-03-15 11:50:19 +08:00 |
|
Krrish Dholakia
|
6f1eb038bc
|
docs(bedrock.md): adding docs for calling bedrock models on proxy via config.yaml
|
2024-03-14 14:50:13 -07:00 |
|
Krrish Dholakia
|
cb221cc88b
|
docs(caching.md): add redis namespaces to docs
|
2024-03-14 13:38:33 -07:00 |
|
dependabot[bot]
|
913f19e046
|
build(deps): bump follow-redirects in /docs/my-website
Bumps [follow-redirects](https://github.com/follow-redirects/follow-redirects) from 1.15.4 to 1.15.6.
- [Release notes](https://github.com/follow-redirects/follow-redirects/releases)
- [Commits](https://github.com/follow-redirects/follow-redirects/compare/v1.15.4...v1.15.6)
---
updated-dependencies:
- dependency-name: follow-redirects
dependency-type: indirect
...
Signed-off-by: dependabot[bot] <support@github.com>
|
2024-03-14 20:02:59 +00:00 |
|
Krrish Dholakia
|
7876aa2d75
|
fix(parallel_request_limiter.py): handle metadata being none
|
2024-03-14 10:02:41 -07:00 |
|
ishaan-jaff
|
0269cddd55
|
(feat) add claude-3-haiku
|
2024-03-13 20:24:06 -07:00 |
|
Aaron Bach
|
daca0a93ce
|
Update docs
|
2024-03-13 17:37:59 -06:00 |
|
Krrish Dholakia
|
16e3aaced5
|
docs(enterprise.md): add prompt injection detection to docs
|
2024-03-13 12:37:32 -07:00 |
|
Krrish Dholakia
|
dbc7552d15
|
docs: refactor team based logging in docs
|
2024-03-13 12:26:39 -07:00 |
|
Krrish Dholakia
|
b3493269b3
|
fix(proxy_server.py): support checking openai user param
|
2024-03-13 12:00:27 -07:00 |
|
Krrish Dholakia
|
9e692d6cce
|
docs(sidebar.js): add dall e 3 cost tracking to docs
|
2024-03-12 22:26:10 -07:00 |
|
Krrish Dholakia
|
488c4b9939
|
docs(cost_tracking.md): add docs for cost tracking dall e 3 calls
|
2024-03-12 21:35:17 -07:00 |
|
Ishaan Jaff
|
5172fb1de9
|
Merge pull request #2474 from BerriAI/litellm_support_command_r
[New-Model] Cohere/command-r
|
2024-03-12 11:11:56 -07:00 |
|
ishaan-jaff
|
8fabaed543
|
(docs) cohere-comand-r
|
2024-03-12 10:34:08 -07:00 |
|
ishaan-jaff
|
ea83c8c9b0
|
(docs) using azure_text models
|
2024-03-12 09:54:34 -07:00 |
|
Krrish Dholakia
|
10f5f342bd
|
docs(virtual_keys.md): cleanup doc
|
2024-03-12 07:05:55 -07:00 |
|
Krrish Dholakia
|
47424b8c90
|
docs(routing.md): fix routing example on docs
|
2024-03-11 22:17:04 -07:00 |
|
ishaan-jaff
|
e46980c56c
|
(docs) using litellm router
|
2024-03-11 21:18:10 -07:00 |
|
Vince Loewe
|
7c38f992dc
|
Merge branch 'main' into main
|
2024-03-11 12:36:41 +09:00 |
|
Ishaan Jaff
|
a1784284bb
|
Merge pull request #2416 from BerriAI/litellm_use_consistent_port
(docs) LiteLLM Proxy - use port 4000 in examples
|
2024-03-09 16:32:08 -08:00 |
|
ishaan-jaff
|
a004240109
|
(docs) deploy litellm
|
2024-03-09 13:45:07 -08:00 |
|
ishaan-jaff
|
ba271e3e87
|
(docs) deploying litell
|
2024-03-09 13:41:20 -08:00 |
|
ishaan-jaff
|
7ae7e95da1
|
(docs) litellm getting started clarify sdk vs proxy
|
2024-03-09 13:04:52 -08:00 |
|
Krrish Dholakia
|
d8a6b8216d
|
docs(input.md): add docs on 'get_supported_openai_params'
|
2024-03-08 23:54:13 -08:00 |
|
Krrish Dholakia
|
b0fa25dfbd
|
docs(audio_transcription.md): add docs on audio transcription
|
2024-03-08 23:51:24 -08:00 |
|
ishaan-jaff
|
ea6f42216c
|
(docs) use port 4000
|
2024-03-08 21:59:00 -08:00 |
|
H0llyW00dzZ
|
33ba57a1b5
|
Fix Docs Formatting in Website
- [+] docs(deploy.md): move tip about versioning inside the tab item
|
2024-03-09 12:37:18 +07:00 |
|
Ishaan Jaff
|
8036b48f14
|
Merge pull request #2408 from BerriAI/litellm_no_store_reqs
[FEAT-liteLLM Proxy] Incognito Requests - Don't log anything
|
2024-03-08 21:11:43 -08:00 |
|
H0llyW00dzZ
|
7004940ce4
|
Update Docs for Kubernetes
- [+] docs(deploy.md): add tip about using versioning or SHA digests instead of latest tag
|
2024-03-09 11:55:31 +07:00 |
|
Guillermo
|
bb427b4659
|
Update deploy.md
|
2024-03-09 02:30:17 +01:00 |
|
Guillermo
|
37b4dde7fd
|
Add quickstart deploy with k8s
|
2024-03-09 02:24:07 +01:00 |
|
ishaan-jaff
|
4ff68c8562
|
(docs) no log requests
|
2024-03-08 16:26:25 -08:00 |
|
ishaan-jaff
|
2d71f54afb
|
(docs) load test litellm
|
2024-03-08 15:18:06 -08:00 |
|
Guillermo
|
a693045460
|
docs: fix yaml typo in proxy/configs.md
equals was used instead of : as key-value delimiter in yaml
|
2024-03-08 20:33:47 +01:00 |
|
Krrish Dholakia
|
0e7b30bec9
|
fix(utils.py): return function name for ollama_chat function calls
|
2024-03-08 08:01:10 -08:00 |
|
ishaan-jaff
|
b4e12fb8fd
|
(docs) litellm cloud formation stack
|
2024-03-07 21:01:28 -08:00 |
|
ishaan-jaff
|
3baa55210a
|
(docs) setting load balancing config
|
2024-03-07 12:22:39 -08:00 |
|
Krrish Dholakia
|
4508b4ede0
|
docs(team_based_routing.md): add docs on team based routing
|
2024-03-07 08:14:46 -08:00 |
|
ishaan-jaff
|
547f0a023d
|
(docs) best practices for high traffic
|
2024-03-06 16:36:35 -08:00 |
|
Krrish Dholakia
|
688b903d19
|
docs(bedrock.md): add anthropic function calling to docs
|
2024-03-05 21:34:46 -08:00 |
|
Krrish Dholakia
|
5c03109b6f
|
docs(configs.md): add load balancing to proxy config docs
|
2024-03-05 07:39:40 -08:00 |
|
Krrish Dholakia
|
8845430cdb
|
docs(anthropic.md): add proxy usage tutorial to anthropic docs
|
2024-03-04 22:00:33 -08:00 |
|
ishaan-jaff
|
1183e5f2e5
|
(feat) maintain anthropic text completion
|
2024-03-04 11:16:34 -08:00 |
|
Ishaan Jaff
|
14fc8355fb
|
Merge pull request #2315 from BerriAI/litellm_add_claude_3
[FEAT]- add claude 3
|
2024-03-04 09:23:13 -08:00 |
|
Ishaan Jaff
|
84415ef7b5
|
Merge pull request #2290 from ti3x/bedrock_mistral
Add support for Bedrock Mistral models
|
2024-03-04 08:42:47 -08:00 |
|
ishaan-jaff
|
d179ae376e
|
(feat) claude-3 test fixes
|
2024-03-04 07:53:06 -08:00 |
|
ishaan-jaff
|
25bbb73ce3
|
(docs) add claude-3
|
2024-03-04 07:40:18 -08:00 |
|
Krrish Dholakia
|
b283588eb8
|
refactor(budget_alerts.md): add spacing in doc
|
2024-03-02 23:12:49 -08:00 |
|
Krrish Dholakia
|
53326da1c9
|
refactor(budget_alerts.md): add spacing in doc
|
2024-03-02 23:12:34 -08:00 |
|
Krrish Dholakia
|
922b38301f
|
docs(budget_alerts.md): doc cleanup
|
2024-03-02 23:11:04 -08:00 |
|
ishaan-jaff
|
fa2c6fdebb
|
(docs) alerting slack budgets
|
2024-03-02 17:55:49 -08:00 |
|
ishaan-jaff
|
b2389a9b1d
|
(feat) improve test slack alert
|
2024-03-02 17:33:18 -08:00 |
|
ishaan-jaff
|
b7416fc946
|
(fix) improve example test slack alert
|
2024-03-02 17:33:07 -08:00 |
|
ishaan-jaff
|
0b85061086
|
(docs)
|
2024-03-02 16:32:33 -08:00 |
|
ishaan-jaff
|
fd9f8b7010
|
(docs) setting soft budgets
|
2024-03-02 13:05:00 -08:00 |
|
Krrish Dholakia
|
468995b288
|
docs(supported_embeddings.md): add vertex ai embeddings to docs
|
2024-03-02 09:34:01 -08:00 |
|
Ishaan Jaff
|
d494941dcc
|
Merge pull request #2285 from Linutux/patch-1
Update deploy.md / Update path to docker-compose.yml
|
2024-03-01 21:02:32 -08:00 |
|
Krrish Dholakia
|
d3a742efc8
|
docs(supported_embedding.md): add vertex ai embeddings to docs
|
2024-03-01 20:57:23 -08:00 |
|
Tim Xia
|
dc84e28c41
|
update docs for supported models
|
2024-03-01 23:34:04 -05:00 |
|
ishaan-jaff
|
d1465ed57e
|
(feat) add gpt-3.5-turbo-instruct-0914
|
2024-03-01 20:02:12 -08:00 |
|
Krrish Dholakia
|
5e516522fb
|
docs(azure.md): improve azure docs
|
2024-03-01 19:40:55 -08:00 |
|
Simon Zöllner
|
57bde371d6
|
Update deploy.md / Update path to docker-compose.yml
The file seems to have been moved
|
2024-03-01 23:32:47 +01:00 |
|
Vince Loewe
|
05c0fa8b9b
|
Merge branch 'main' into main
|
2024-03-01 13:37:17 -08:00 |
|
ishaan-jaff
|
edab94c0a6
|
(docs) load balancing on proxy
|
2024-03-01 12:28:59 -08:00 |
|
ishaan-jaff
|
295d36a383
|
(docs) load balancing
|
2024-03-01 11:46:42 -08:00 |
|
ishaan-jaff
|
4bfefd2b08
|
(feat) add groq api pricing, models
|
2024-02-29 16:11:22 -08:00 |
|
Ishaan Jaff
|
d2115d5a17
|
Merge pull request #2263 from BerriAI/litellm_gpt_gemini_base_64
[FEAT] Use Base64 images with vertex_ai/gemini-pro-vision
|
2024-02-29 16:01:19 -08:00 |
|
ishaan-jaff
|
02c3955c51
|
(docs) using gemini vision with base64
|
2024-02-29 15:55:56 -08:00 |
|
Ishaan Jaff
|
e044d6332f
|
Merge pull request #2255 from BerriAI/litellm_admin_ui_user_panel_load_time_high
[FEAT] proxy add pagination on /user/info endpoint (Admin UI does not load all users)
|
2024-02-29 13:22:00 -08:00 |
|
ishaan-jaff
|
2415dc7927
|
(docs) /user/info
|
2024-02-29 13:05:07 -08:00 |
|
ishaan-jaff
|
ed85d2de95
|
(docs) /user/info
|
2024-02-29 13:04:41 -08:00 |
|
Ishaan Jaff
|
3d2150afcc
|
Merge pull request #2247 from BerriAI/litellm_mistral_azure_ai_fixes
[FEAT] mistral azure - use with env vars + added pricing
|
2024-02-29 12:22:53 -08:00 |
|
ishaan-jaff
|
3fea348e4e
|
(docs) enterprise update
|
2024-02-29 11:42:19 -08:00 |
|
ishaan-jaff
|
817fa2483a
|
(docs) mistral on azure ai studio
|
2024-02-29 08:28:19 -08:00 |
|
ishaan-jaff
|
46134d3fba
|
(docs) using mistral azure ai studio
|
2024-02-29 08:14:24 -08:00 |
|
Vince Loewe
|
f98619e6f2
|
Merge branch 'BerriAI:main' into main
|
2024-02-28 22:18:14 -08:00 |
|
Krrish Dholakia
|
f120f94cc9
|
docs(ui.md): update
|
2024-02-28 21:25:39 -08:00 |
|
ishaan-jaff
|
76e7f8831f
|
(docs) check if supports function calling
|
2024-02-28 17:41:54 -08:00 |
|
Krrish Dholakia
|
a042092faa
|
test: removing bedrock claude-v1 testing - bedrock removed this
|
2024-02-28 11:08:17 -08:00 |
|
Vince Loewe
|
a9648613dc
|
feat: LLMonitor is now Lunary
|
2024-02-27 22:07:13 -08:00 |
|
Ishaan Jaff
|
990439c49c
|
Merge branch 'main' into litellm_daily_metrics
|
2024-02-27 20:33:35 -08:00 |
|
ishaan-jaff
|
4fb723f119
|
(docs) show /daily_metrics
|
2024-02-27 19:34:47 -08:00 |
|
ishaan-jaff
|
d19370083d
|
(feat) update /daily metrics
|
2024-02-27 18:33:59 -08:00 |
|
ishaan-jaff
|
7cabc6ac56
|
(docs) show daily metrics tab
|
2024-02-27 18:03:15 -08:00 |
|
ishaan-jaff
|
7485fa797c
|
(docs) vertex ai litellm proxy
|
2024-02-27 17:55:05 -08:00 |
|
ishaan-jaff
|
f3144dd9cf
|
(docs) vertex ai
|
2024-02-27 17:44:35 -08:00 |
|
ishaan-jaff
|
8075746078
|
(docs) fix clickhouse link
|
2024-02-26 18:35:33 -08:00 |
|
ishaan-jaff
|
9bb55a0671
|
add clickhouse docs
|
2024-02-26 18:31:10 -08:00 |
|
ishaan-jaff
|
c627ce6631
|
(docs) use azure ai studio + mistral litellm
|
2024-02-26 14:33:21 -08:00 |
|
ishaan-jaff
|
da4832d5a2
|
(feat) add mistral-large-latest
|
2024-02-26 09:25:44 -08:00 |
|
Krrish Dholakia
|
577a0f9440
|
docs: add gemini-1.5 to docs
|
2024-02-26 08:01:30 -08:00 |
|
Krish Dholakia
|
61d69b1efa
|
Merge pull request #2183 from BerriAI/litellm_team_rate_limits
fix(proxy_server.py): allow user to set team tpm/rpm limits/budget/models
|
2024-02-25 01:12:39 -08:00 |
|
ishaan-jaff
|
7c19783d32
|
(docs) writing custo add
|
2024-02-24 18:54:46 -08:00 |
|
Ishaan Jaff
|
c8daf61592
|
Merge pull request #2188 from BerriAI/litellm_fix_health_checks
[Fix] Fix health check when API base set for OpenAI compatible models
|
2024-02-24 18:47:13 -08:00 |
|
ishaan-jaff
|
a4d2d1c4e6
|
(docs) using openai compatible endpoints
|
2024-02-24 18:46:20 -08:00 |
|
ishaan-jaff
|
bf403dc02e
|
(docs) health
|
2024-02-24 18:38:39 -08:00 |
|