Commit Graph
9422 Commits
Author SHA1 Message Date
Krrish Dholakia c52819d47c fix(proxy_server.py): don't require scope for team-based jwt access
If team with the client_id exists then it should be allowed to make a request, if it doesn't then as we discussed it should return an error
2024-04-01 18:52:00 -07:00
Krrish Dholakia ceabf726b0 fix(main.py): support max retries for transcription calls 2024-04-01 18:37:53 -07:00
Krrish Dholakia c3e4af76cf refactor: fix linting issue 2024-04-01 18:11:38 -07:00
Krrish Dholakia ca54b62656 refactor(main.py): trigger new build 2024-04-01 18:03:46 -07:00
Krrish Dholakia 6467dd4e11 fix(tpm_rpm_limiter.py): fix cache init logic 2024-04-01 18:01:38 -07:00
Krrish Dholakia 9c0aecf9b8 fix(proxy/utils.py): support redis caching for alerting 2024-04-01 16:13:59 -07:00
Krrish Dholakia cdae08f3c3 docs(openai.md): fix docs to include example of calling openai on proxy 2024-04-01 12:09:22 -07:00
Krish DholakiaandGitHub d3e61a0adc Merge pull request #2783 from BerriAI/litellm_context_window_Fallback_fix
fix(router.py): fix check for context window fallbacks
2024-04-01 11:55:49 -07:00
Krrish Dholakia a917fadf45 docs(routing.md): refactor docs to show how to use pre-call checks and fallback across model groups 2024-04-01 11:21:27 -07:00
Ishaan Jaff d5d800e141 (fix) _update_end_user_cache 2024-04-01 11:18:00 -07:00
Krrish Dholakia 52b1538b2e fix(router.py): support context window fallbacks for pre-call checks 2024-04-01 10:51:54 -07:00
Krrish Dholakia f46a9d09a5 fix(router.py): fix check for context window fallbacks
fallback if list is not none
2024-04-01 10:41:12 -07:00
Krrish Dholakia c9e6b05cfb test(test_max_tpm_rpm_limiter.py): add unit testing for redis namespaces working for tpm/rpm limits 2024-04-01 10:39:03 -07:00
Ishaan JaffandGitHub f58436d6a0 Merge pull request #2782 from phact/patch-1
support cohere_chat in get_api_key
2024-04-01 10:33:43 -07:00
Sebastián EstévezandGitHub e50e76bbd5 support cohere_chat in get_api_key 2024-04-01 13:24:03 -04:00
Ishaan Jaff 53d7b95364 (fix) load testing key used 2024-04-01 08:28:19 -07:00
Ishaan JaffandGitHub bbfd850e12 Merge pull request #2774 from BerriAI/litellm_async_perf
(fix) improve async perf by 100ms
2024-04-01 08:12:34 -07:00
Krrish DholakiaandIshaan Jaff f3e47323b9 test(test_max_tpm_rpm_limiter.py): unit tests for key + team based tpm rpm limits on proxy 2024-04-01 08:11:30 -07:00
Krrish DholakiaandIshaan Jaff 19fc120081 fix(proxy/utils.py): uncomment max parallel request limit check 2024-04-01 08:11:30 -07:00
Krrish DholakiaandIshaan Jaff 1bb4f3ad6d fix(utils.py): set redis_usage_cache to none by default 2024-04-01 08:11:30 -07:00
Krrish DholakiaandIshaan Jaff 2dd5f2bc8c fix(tpm_rpm_limiter.py): enable redis caching for tpm/rpm checks on keys/user/teams
allows tpm/rpm checks to work across instances

https://github.com/BerriAI/litellm/issues/2730
2024-04-01 08:11:30 -07:00
Krrish Dholakia 383f12bbd3 test(test_max_tpm_rpm_limiter.py): unit tests for key + team based tpm rpm limits on proxy 2024-04-01 08:00:01 -07:00
Ishaan Jaff ddb35facc0 ci/cd run again 2024-04-01 07:40:05 -07:00
DaxServerandIshaan Jaff 3f25049dc8 fix(docs): Correct Docker pull command in deploy.md
Corrected the Docker pull command in deploy.md to remove duplicated 'docker pull' command.
2024-04-01 07:29:56 -07:00
DaxServerandIshaan Jaff a2c7455c3d docs: Update references to Ollama repository url
Updated references to the Ollama repository URL from https://github.com/jmorganca/ollama to https://github.com/ollama/ollama.
2024-04-01 07:29:56 -07:00
Krrish DholakiaandIshaan Jaff ea2356fd95 bump: version 1.34.17 → 1.34.18 2024-04-01 07:29:56 -07:00
Krrish DholakiaandIshaan Jaff 5800a5095a refactor(main.py): trigger new build 2024-04-01 07:29:56 -07:00
Krrish DholakiaandIshaan Jaff a1365f6035 test: cleanup 2024-04-01 07:29:56 -07:00
Krrish DholakiaandIshaan Jaff d4dd6d0cdc fix(proxy/utils.py): uncomment max parallel request limit check 2024-04-01 07:29:56 -07:00
Krrish DholakiaandIshaan Jaff aebb0e489c test: fix test 2024-04-01 07:29:56 -07:00
Krrish DholakiaandIshaan Jaff 17cabf013c fix(caching.py): respect redis namespace for all redis get/set requests 2024-04-01 07:29:56 -07:00
Krrish DholakiaandIshaan Jaff 583e334bd2 fix(utils.py): set redis_usage_cache to none by default 2024-04-01 07:29:56 -07:00
Krrish DholakiaandIshaan Jaff 0c77f75ce9 fix(tpm_rpm_limiter.py): enable redis caching for tpm/rpm checks on keys/user/teams
allows tpm/rpm checks to work across instances

https://github.com/BerriAI/litellm/issues/2730
2024-04-01 07:29:56 -07:00
Krrish DholakiaandIshaan Jaff ea76c546ff docs(deploy.md): fix docs for litlelm-database docker run example 2024-04-01 07:29:56 -07:00
Ishaan Jaff 15b5cb1612 (fix) check size of data to predict 2024-04-01 07:29:56 -07:00
Ishaan Jaff e62e83c42a (ui) dont let prediction block spend view 2024-04-01 07:29:56 -07:00
Ishaan JaffandGitHub 18fec3ad8e Merge pull request #2779 from DaxServer/update-proxy-dockerfile-branch
fix(docs): Correct Docker pull command in deploy.md
2024-04-01 07:10:45 -07:00
Ishaan JaffandGitHub 7e24461ae9 Merge pull request #2778 from DaxServer/main-1
docs: Update references to Ollama repository url
2024-04-01 07:10:23 -07:00
DaxServerandGitButler 28f6caa04c fix(docs): Correct Docker pull command in deploy.md
Corrected the Docker pull command in deploy.md to remove duplicated 'docker pull' command.
2024-03-31 20:10:00 +02:00
DaxServerandGitButler 61b6f8be44 docs: Update references to Ollama repository url
Updated references to the Ollama repository URL from https://github.com/jmorganca/ollama to https://github.com/ollama/ollama.
2024-03-31 19:35:37 +02:00
Krrish Dholakia 006ea0abe0 bump: version 1.34.17 → 1.34.18 v1.34.18 v1.34.17 2024-03-30 22:10:21 -07:00
Krish DholakiaandGitHub 1356f6cd32 Merge pull request #2775 from BerriAI/litellm_redis_user_api_key_cache_v3
fix(tpm_rpm_limiter.py): enable redis caching for tpm/rpm checks on keys/user/teams
2024-03-30 22:07:05 -07:00
Krrish Dholakia f5d920e314 refactor(main.py): trigger new build 2024-03-30 21:41:14 -07:00
Krrish Dholakia 60f89faf1c test: cleanup 2024-03-30 21:40:43 -07:00
Krrish Dholakia 3b8e7241b4 fix(proxy/utils.py): uncomment max parallel request limit check 2024-03-30 20:51:59 -07:00
Krrish Dholakia 364526d0bc test: fix test 2024-03-30 20:22:48 -07:00
Krrish Dholakia 5926792de6 fix(caching.py): respect redis namespace for all redis get/set requests 2024-03-30 20:20:29 -07:00
Krrish Dholakia d9ff13b624 fix(utils.py): set redis_usage_cache to none by default 2024-03-30 20:10:56 -07:00
Krrish Dholakia a7aa6fae64 docs(deploy.md): fix docs for litlelm-database docker run example 2024-03-30 20:08:27 -07:00
Krrish Dholakia f58fefd589 fix(tpm_rpm_limiter.py): enable redis caching for tpm/rpm checks on keys/user/teams
allows tpm/rpm checks to work across instances

https://github.com/BerriAI/litellm/issues/2730
2024-03-30 20:01:36 -07:00