Commit Graph
17174 Commits
Author SHA1 Message Date
Ishaan JaffandGitHub 1bd2b2fc92 Merge pull request #5449 from BerriAI/litellm_Fix_vertex_multimodal
[Fix-Proxy] Allow running /health checks on vertex multimodal embedding requests
2024-08-30 10:21:42 -07:00
Ishaan Jaff a6273a29fe add test for test_vertexai_multimodal_embedding_text_input 2024-08-30 09:19:48 -07:00
8d6a0bdc81 - merge - fix TypeError: 'CompletionUsage' object is not subscriptable #5441 (#5448)
* fix TypeError: 'CompletionUsage' object is not subscriptable (#5441)

* test(test_team_logging.py): mark flaky test

---------

Co-authored-by: yafei lee <yafei@dao42.com>
2024-08-30 08:54:42 -07:00
JooHo KimandGitHub 5b1d9712c5 chore: Clarify support-related Exceptions in utils.py (#5447)
Improved the clarity of Exceptions in supports_system_messages, supports_response_schema, supports_function_calling, and supports_parallel_function_calling. Previously, it was difficult to determine the cause of Exception logs due to vague messaging. Each case now includes a more specific and appropriate Exception message.
2024-08-30 08:29:05 -07:00
Krrish Dholakia 7f1531006c docs(routing.md): add weight-based shuffling to docs 2024-08-30 08:24:12 -07:00
Krrish Dholakia 94db4ec830 test: mark flaky tests 2024-08-30 07:53:04 -07:00
Ishaan JaffandGitHub a9ca183021 Merge pull request #5438 from BerriAI/litellm_show_error_types_swagger
[Feat-Proxy] Show all exceptioons types on swagger for LiteLLM Proxy
v1.44.12 v1.44.12-stable
2024-08-30 07:21:23 -07:00
Ishaan JaffandGitHub ef7835f7f3 Merge pull request #5442 from kiriloman/main
[Pricing] Add pricing for Openai ft:gpt-4o
2024-08-30 07:19:30 -07:00
Kyrylo Yefimenko a100b01b90 Add pricing for Openai ft:gpt-4o 2024-08-30 08:16:24 +01:00
Krrish Dholakia 856ed40a07 bump: version 1.44.11 → 1.44.12 2024-08-29 22:41:10 -07:00
Krish DholakiaandGitHub dd7b008161 fix: Minor LiteLLM Fixes + Improvements (29/08/2024) (#5436)
* fix(model_checks.py): support returning wildcard models on `/v1/models`

Fixes https://github.com/BerriAI/litellm/issues/4903

* fix(bedrock_httpx.py): support calling bedrock via api_base

Closes https://github.com/BerriAI/litellm/pull/4587

* fix(litellm_logging.py): only leave last 4 char of gemini key unmasked

Fixes https://github.com/BerriAI/litellm/issues/5433

* feat(router.py): support setting 'weight' param for models on router

Closes https://github.com/BerriAI/litellm/issues/5410

* test(test_bedrock_completion.py): add unit test for custom api base

* fix(model_checks.py): handle no "/" in model
2024-08-29 22:40:25 -07:00
Ishaan Jaff f70b7575d2 update docs 2024-08-29 21:00:10 -07:00
Ishaan Jaff 26c03c9c8b add pricing for vertex ai 21 2024-08-29 19:03:38 -07:00
Ishaan Jaff ad88c7d0a8 show all error types on swagger 2024-08-29 18:50:41 -07:00
Ishaan Jaff ef47b2bc87 mark test_cost_tracking_with_caching as flaky v1.44.11 v1.44.11-stable 2024-08-29 17:44:21 -07:00
Ishaan Jaff d57f5a955e bump: version 1.44.10 → 1.44.11 2024-08-29 17:37:01 -07:00
Ishaan JaffandGitHub 9444b34711 Merge pull request #5432 from BerriAI/litellm_add_tag_control_team
[Feat-Proxy] Set tags per team - (use tag based routing for team)
2024-08-29 17:34:58 -07:00
Ishaan JaffandGitHub e329c4509a Merge branch 'main' into litellm_add_tag_control_team 2024-08-29 17:34:40 -07:00
Ishaan JaffandGitHub ef16738720 Merge pull request #5435 from BerriAI/litellm_fwd_vtx_sdk_headers
[Feat-Proxy] Pass through Vertex Endpoint - allow forwarding vertex credentials
2024-08-29 17:24:35 -07:00
Ishaan JaffandGitHub 010f526226 Merge branch 'main' into litellm_fwd_vtx_sdk_headers 2024-08-29 17:24:31 -07:00
Ishaan Jaff a4b88c16dc fix indentation 2024-08-29 17:01:23 -07:00
Ishaan Jaff 748cc80783 fix auth checks for provider routes 2024-08-29 16:40:46 -07:00
Ishaan Jaff e449bf062d add docs on pass thtough 2024-08-29 16:12:14 -07:00
Ishaan Jaff 41a5daf20c add test for vertex sdk foward headers 2024-08-29 15:50:54 -07:00
Ishaan Jaff 9fb188a3b6 vertex add vertex endpoints 2024-08-29 15:49:25 -07:00
Krrish Dholakia 601945d114 docs(docker_quick_start.md): add new quick start doc for litellm proxy 2024-08-29 15:35:39 -07:00
Ishaan Jaff 4b7ceade64 mark test_key_info_spend_values_streaming as flaky 2024-08-29 14:39:53 -07:00
Ishaan Jaff da2cefc45a fix team based tag routing 2024-08-29 14:37:44 -07:00
Ishaan JaffandGitHub 208fe6cb90 Merge pull request #5430 from Manouchehri/bedrock-cross-inference-1
(bedrock): Add new cross-region inference support for Bedrock.
2024-08-29 14:30:11 -07:00
Ishaan JaffandGitHub 5851a8f901 Merge pull request #5431 from BerriAI/litellm_Add_fireworks_ai_health_check
[Fix-Proxy] /health check for provider wildcard models (fireworks/*)
2024-08-29 14:25:05 -07:00
Ishaan Jaff 308377fbe2 docs tag based routing per team 2024-08-29 14:23:55 -07:00
Ishaan Jaff 242f66054d enable_tag_filtering 2024-08-29 14:14:46 -07:00
Ishaan Jaff d9433d9f94 doc Tag Based Routing 2024-08-29 14:14:37 -07:00
Ishaan Jaff 944c7ac3fa fix missing link on docs 2024-08-29 14:00:16 -07:00
Ishaan Jaff f592aeaa38 add test_chat_completion_with_no_tags 2024-08-29 13:54:11 -07:00
Ishaan Jaff 84bda9cc80 fix get_deployments_for_tag 2024-08-29 13:51:36 -07:00
Ishaan Jaff 34f1d32799 add test for tag based routing 2024-08-29 13:45:24 -07:00
Ishaan Jaff 2cb4882e70 define tags on model list 2024-08-29 13:06:39 -07:00
Ishaan Jaff ff42962750 add_team_based_tags_to_metadata 2024-08-29 13:06:03 -07:00
Ishaan Jaff ffeb5ce22a add set / update tags for a team 2024-08-29 13:05:00 -07:00
Ishaan Jaff 5e590e7326 allow settings tags per team 2024-08-29 13:03:49 -07:00
Ishaan Jaff d2ddc5aba9 add test_team_tags to set / update tags 2024-08-29 13:02:57 -07:00
Ishaan Jaff 284b8b3418 add test for health check 2024-08-29 11:16:20 -07:00
David Manouchehri 19db80ffeb (bedrock): Add new cross-region inference support for Bedrock. 2024-08-29 17:49:16 +00:00
Ishaan Jaff b576233e66 add support for fireworks ai health check 2024-08-29 09:29:16 -07:00
Ishaan Jaff 45774010f7 add util to pick_cheapest_model_from_llm_provider 2024-08-29 09:27:20 -07:00
Ishaan Jaff 4b6a2fa4f3 add fireworks_ai_models 2024-08-29 09:23:11 -07:00
Krish DholakiaandGitHub 559a6ad826 fix(google_ai_studio): working context caching (#5421)
* fix(google_ai_studio): working context caching

* feat(vertex_ai_context_caching.py): support async cache check calls

* fix(vertex_and_google_ai_studio_gemini.py): fix setting headers

* fix(vertex_ai_parter_models): fix import

* fix(vertex_and_google_ai_studio_gemini.py): fix input

* test(test_amazing_vertex_completion.py): fix test
2024-08-29 07:00:30 -07:00
Krish DholakiaandGitHub 8ce1e49fbe fix(utils.py): correctly log streaming cache hits (#5417) (#5426)
Fixes https://github.com/BerriAI/litellm/issues/5401
2024-08-28 22:50:33 -07:00
Krrish Dholakia e9957f1265 bump: version 1.44.9 → 1.44.10 v1.44.10 v1.44.10-stable 2024-08-28 22:21:38 -07:00