Krrish Dholakia
|
c52819d47c
|
fix(proxy_server.py): don't require scope for team-based jwt access
If team with the client_id exists then it should be allowed to make a request, if it doesn't then as we discussed it should return an error
|
2024-04-01 18:52:00 -07:00 |
|
Krrish Dholakia
|
ceabf726b0
|
fix(main.py): support max retries for transcription calls
|
2024-04-01 18:37:53 -07:00 |
|
Krrish Dholakia
|
c3e4af76cf
|
refactor: fix linting issue
|
2024-04-01 18:11:38 -07:00 |
|
Krrish Dholakia
|
ca54b62656
|
refactor(main.py): trigger new build
|
2024-04-01 18:03:46 -07:00 |
|
Krrish Dholakia
|
6467dd4e11
|
fix(tpm_rpm_limiter.py): fix cache init logic
|
2024-04-01 18:01:38 -07:00 |
|
Krrish Dholakia
|
9c0aecf9b8
|
fix(proxy/utils.py): support redis caching for alerting
|
2024-04-01 16:13:59 -07:00 |
|
Krrish Dholakia
|
cdae08f3c3
|
docs(openai.md): fix docs to include example of calling openai on proxy
|
2024-04-01 12:09:22 -07:00 |
|
 Krish DholakiaandGitHub
|
d3e61a0adc
|
Merge pull request #2783 from BerriAI/litellm_context_window_Fallback_fix
fix(router.py): fix check for context window fallbacks
|
2024-04-01 11:55:49 -07:00 |
|
Krrish Dholakia
|
a917fadf45
|
docs(routing.md): refactor docs to show how to use pre-call checks and fallback across model groups
|
2024-04-01 11:21:27 -07:00 |
|
Ishaan Jaff
|
d5d800e141
|
(fix) _update_end_user_cache
|
2024-04-01 11:18:00 -07:00 |
|
Krrish Dholakia
|
52b1538b2e
|
fix(router.py): support context window fallbacks for pre-call checks
|
2024-04-01 10:51:54 -07:00 |
|
Krrish Dholakia
|
f46a9d09a5
|
fix(router.py): fix check for context window fallbacks
fallback if list is not none
|
2024-04-01 10:41:12 -07:00 |
|
Krrish Dholakia
|
c9e6b05cfb
|
test(test_max_tpm_rpm_limiter.py): add unit testing for redis namespaces working for tpm/rpm limits
|
2024-04-01 10:39:03 -07:00 |
|
 Ishaan JaffandGitHub
|
f58436d6a0
|
Merge pull request #2782 from phact/patch-1
support cohere_chat in get_api_key
|
2024-04-01 10:33:43 -07:00 |
|
 Sebastián EstévezandGitHub
|
e50e76bbd5
|
support cohere_chat in get_api_key
|
2024-04-01 13:24:03 -04:00 |
|
Ishaan Jaff
|
53d7b95364
|
(fix) load testing key used
|
2024-04-01 08:28:19 -07:00 |
|
 Ishaan JaffandGitHub
|
bbfd850e12
|
Merge pull request #2774 from BerriAI/litellm_async_perf
(fix) improve async perf by 100ms
|
2024-04-01 08:12:34 -07:00 |
|
 Krrish DholakiaandIshaan Jaff
|
f3e47323b9
|
test(test_max_tpm_rpm_limiter.py): unit tests for key + team based tpm rpm limits on proxy
|
2024-04-01 08:11:30 -07:00 |
|
 Krrish DholakiaandIshaan Jaff
|
19fc120081
|
fix(proxy/utils.py): uncomment max parallel request limit check
|
2024-04-01 08:11:30 -07:00 |
|
 Krrish DholakiaandIshaan Jaff
|
1bb4f3ad6d
|
fix(utils.py): set redis_usage_cache to none by default
|
2024-04-01 08:11:30 -07:00 |
|
 Krrish DholakiaandIshaan Jaff
|
2dd5f2bc8c
|
fix(tpm_rpm_limiter.py): enable redis caching for tpm/rpm checks on keys/user/teams
allows tpm/rpm checks to work across instances
https://github.com/BerriAI/litellm/issues/2730
|
2024-04-01 08:11:30 -07:00 |
|
Krrish Dholakia
|
383f12bbd3
|
test(test_max_tpm_rpm_limiter.py): unit tests for key + team based tpm rpm limits on proxy
|
2024-04-01 08:00:01 -07:00 |
|
Ishaan Jaff
|
ddb35facc0
|
ci/cd run again
|
2024-04-01 07:40:05 -07:00 |
|
 DaxServerandIshaan Jaff
|
3f25049dc8
|
fix(docs): Correct Docker pull command in deploy.md
Corrected the Docker pull command in deploy.md to remove duplicated 'docker pull' command.
|
2024-04-01 07:29:56 -07:00 |
|
 DaxServerandIshaan Jaff
|
a2c7455c3d
|
docs: Update references to Ollama repository url
Updated references to the Ollama repository URL from https://github.com/jmorganca/ollama to https://github.com/ollama/ollama.
|
2024-04-01 07:29:56 -07:00 |
|
 Krrish DholakiaandIshaan Jaff
|
ea2356fd95
|
bump: version 1.34.17 → 1.34.18
|
2024-04-01 07:29:56 -07:00 |
|
 Krrish DholakiaandIshaan Jaff
|
5800a5095a
|
refactor(main.py): trigger new build
|
2024-04-01 07:29:56 -07:00 |
|
 Krrish DholakiaandIshaan Jaff
|
a1365f6035
|
test: cleanup
|
2024-04-01 07:29:56 -07:00 |
|
 Krrish DholakiaandIshaan Jaff
|
d4dd6d0cdc
|
fix(proxy/utils.py): uncomment max parallel request limit check
|
2024-04-01 07:29:56 -07:00 |
|
 Krrish DholakiaandIshaan Jaff
|
aebb0e489c
|
test: fix test
|
2024-04-01 07:29:56 -07:00 |
|
 Krrish DholakiaandIshaan Jaff
|
17cabf013c
|
fix(caching.py): respect redis namespace for all redis get/set requests
|
2024-04-01 07:29:56 -07:00 |
|
 Krrish DholakiaandIshaan Jaff
|
583e334bd2
|
fix(utils.py): set redis_usage_cache to none by default
|
2024-04-01 07:29:56 -07:00 |
|
 Krrish DholakiaandIshaan Jaff
|
0c77f75ce9
|
fix(tpm_rpm_limiter.py): enable redis caching for tpm/rpm checks on keys/user/teams
allows tpm/rpm checks to work across instances
https://github.com/BerriAI/litellm/issues/2730
|
2024-04-01 07:29:56 -07:00 |
|
 Krrish DholakiaandIshaan Jaff
|
ea76c546ff
|
docs(deploy.md): fix docs for litlelm-database docker run example
|
2024-04-01 07:29:56 -07:00 |
|
Ishaan Jaff
|
15b5cb1612
|
(fix) check size of data to predict
|
2024-04-01 07:29:56 -07:00 |
|
Ishaan Jaff
|
e62e83c42a
|
(ui) dont let prediction block spend view
|
2024-04-01 07:29:56 -07:00 |
|
 Ishaan JaffandGitHub
|
18fec3ad8e
|
Merge pull request #2779 from DaxServer/update-proxy-dockerfile-branch
fix(docs): Correct Docker pull command in deploy.md
|
2024-04-01 07:10:45 -07:00 |
|
 Ishaan JaffandGitHub
|
7e24461ae9
|
Merge pull request #2778 from DaxServer/main-1
docs: Update references to Ollama repository url
|
2024-04-01 07:10:23 -07:00 |
|
 DaxServerandGitButler
|
28f6caa04c
|
fix(docs): Correct Docker pull command in deploy.md
Corrected the Docker pull command in deploy.md to remove duplicated 'docker pull' command.
|
2024-03-31 20:10:00 +02:00 |
|
 DaxServerandGitButler
|
61b6f8be44
|
docs: Update references to Ollama repository url
Updated references to the Ollama repository URL from https://github.com/jmorganca/ollama to https://github.com/ollama/ollama.
|
2024-03-31 19:35:37 +02:00 |
|
Krrish Dholakia
|
006ea0abe0
|
bump: version 1.34.17 → 1.34.18
v1.34.18
v1.34.17
|
2024-03-30 22:10:21 -07:00 |
|
 Krish DholakiaandGitHub
|
1356f6cd32
|
Merge pull request #2775 from BerriAI/litellm_redis_user_api_key_cache_v3
fix(tpm_rpm_limiter.py): enable redis caching for tpm/rpm checks on keys/user/teams
|
2024-03-30 22:07:05 -07:00 |
|
Krrish Dholakia
|
f5d920e314
|
refactor(main.py): trigger new build
|
2024-03-30 21:41:14 -07:00 |
|
Krrish Dholakia
|
60f89faf1c
|
test: cleanup
|
2024-03-30 21:40:43 -07:00 |
|
Krrish Dholakia
|
3b8e7241b4
|
fix(proxy/utils.py): uncomment max parallel request limit check
|
2024-03-30 20:51:59 -07:00 |
|
Krrish Dholakia
|
364526d0bc
|
test: fix test
|
2024-03-30 20:22:48 -07:00 |
|
Krrish Dholakia
|
5926792de6
|
fix(caching.py): respect redis namespace for all redis get/set requests
|
2024-03-30 20:20:29 -07:00 |
|
Krrish Dholakia
|
d9ff13b624
|
fix(utils.py): set redis_usage_cache to none by default
|
2024-03-30 20:10:56 -07:00 |
|
Krrish Dholakia
|
a7aa6fae64
|
docs(deploy.md): fix docs for litlelm-database docker run example
|
2024-03-30 20:08:27 -07:00 |
|
Krrish Dholakia
|
f58fefd589
|
fix(tpm_rpm_limiter.py): enable redis caching for tpm/rpm checks on keys/user/teams
allows tpm/rpm checks to work across instances
https://github.com/BerriAI/litellm/issues/2730
|
2024-03-30 20:01:36 -07:00 |
|