Commit Graph
25030 Commits
Author SHA1 Message Date
Krish DholakiaandGitHub f4e2870490 Merge pull request #14532 from timelfrink/feat/issue-14476-compactifai-provider
Add CompactifAI provider support
2025-09-15 21:15:00 -07:00
Krish DholakiaandGitHub e368a0895f Merge pull request #14542 from uc4w6c/fix/mcp-server-delete-refresh
fix: recompute filters after deleting an MCP Server
2025-09-15 21:12:55 -07:00
Krish DholakiaandGitHub 784eb18366 Merge pull request #14558 from iabhi4/fix-14536
fix(proxy): Correctly parse multi-part MCP server aliases from URL paths
2025-09-15 21:05:22 -07:00
Krish DholakiaandGitHub 9d2cce43ba Merge pull request #14583 from pazevedo-hyland/fix/bedrock_convert_layer
Fix: handle empty arguments in Bedrock tool call invocation
2025-09-15 20:57:45 -07:00
Ishaan JaffandGitHub 8e22cf5d65 [Fix] /responses API - add cancel endpoint + allow non-admins to use this as an llm api endpoint (#14594)
* fix: ensure /responses/cancel works for non admins

* test: cancel endpoint

* fix responses API  cancel endpoint

* test fix

* TestGoogleAIStudioResponsesAPITest
2025-09-15 18:49:54 -07:00
Krish DholakiaandGitHub e1a6b9f858 Merge pull request #14582 from timelfrink/fix/issue-14573-aws-external-id-support
Add AWS external ID parameter support for Bedrock authentication
2025-09-15 17:32:00 -07:00
Ishaan Jaffer eb3e159b7c docs update 2025-09-15 17:22:03 -07:00
Tim Elfrink afd720a62f Fix CompactifAI provider tests and implementation
- Add missing provider_config parameter in main.py for proper HTTP handler integration
- Update tests to use correct respx mocking pattern with litellm.disable_aiohttp_transport
- Add get_error_class method to CompactifAI transformation for proper error handling
- Fix authentication error test to expect APIConnectionError instead of AuthenticationError
- All 8 CompactifAI tests now pass successfully
2025-09-15 22:03:42 +02:00
Tim ElfrinkandGitHub 9d7942eb35 Fix: Vertex AI Gemini labels field provider-aware filtering (#14563)
* Add comprehensive tests for Vertex AI Gemini labels provider filtering

- Test Google GenAI endpoints exclude labels even when explicitly provided
- Test Vertex AI endpoints include labels when provided
- Cover provider detection logic for different endpoint URLs
- Verify metadata-to-labels conversion only happens for Vertex AI
- Ensure edge cases are handled properly (null/empty api_base)

* Fix Vertex AI Gemini labels field provider-aware filtering

- Add _is_google_genai_endpoint() function to detect Google GenAI vs Vertex AI endpoints
- Update _transform_request_body() to accept api_base parameter
- Only include labels field for Vertex AI endpoints (not Google GenAI)
- Pass api_base through sync/async transform functions
- Maintain backward compatibility with existing usage
- Fixes issue where Google GenAI requests failed with unsupported labels field

* Refactor labels filtering to use custom_llm_provider instead of URL parsing

Replace URL-based endpoint detection with custom_llm_provider parameter
checking for cleaner, more reliable provider identification.

Changes:
- Remove _is_google_genai_endpoint() helper function
- Update labels condition to use custom_llm_provider != "gemini"
- Remove api_base parameter from _transform_request_body()
- Simplify sync/async transform function signatures
- Update tests to reflect new parameter structure
- Remove obsolete test_provider_detection test

This approach aligns with existing codebase patterns where
custom_llm_provider="gemini" identifies Google AI Studio endpoints
that don't support labels, while vertex_ai/vertex_ai_beta identify
Vertex AI endpoints that do support labels.

* Use LlmProviders.GEMINI constant instead of hardcoded string
2025-09-15 12:43:07 -07:00
Mubashir OsmaniandGitHub 321d5299b2 s3_endpoint_url returned 404 (#14559)
* added spend metrics

* feat: Add Spend metrics in datadog

* fix: lint errors

* fix: s3 endpoint url logging

* fixed lint errors

* remove from branch

This reverts commit e123cae06e.

* Remove from branch

This reverts commit e694cc102a.

* remove "added spend metrics"

This reverts commit 6156590190.
2025-09-15 12:08:18 -07:00
Ishaan JaffandGitHub cebacd65cf [Bug Fix] SCIM v2 - ensure group PUSH and PUT ops allow creating non-existent members (#14581)
* fix: scim handle non existent members

* test - scim v2

* test fix

* fix: NewUserResponse
2025-09-15 11:27:05 -07:00
pazevedo-hyland 6558156642 Fix: handle empty arguments in Bedrock tool call invocation 2025-09-15 19:23:39 +01:00
Tim Elfrink 5bd94cccb9 Add AWS external ID parameter support for Bedrock authentication
- Add aws_external_id to authentication parameters list
- Update get_credentials method to accept and propagate external ID
- Modify all STS assume_role calls to conditionally include ExternalId parameter
- Support both assume_role and assume_role_with_web_identity flows
- Handle IRSA cross-account and same-account role assumption scenarios
- Add external ID support to Bedrock Converse API authentication
- Maintain full backward compatibility with existing authentication flows
- Support AWS_EXTERNAL_ID environment variable

Fixes cross-account role assumption security requirements per AWS best practices.
2025-09-15 19:57:04 +02:00
Tim Elfrink f6ff7042ba Add comprehensive tests for AWS external ID support
- Test external ID parameter propagation through authentication chain
- Cover both standard Bedrock and Converse API authentication flows
- Verify assume_role STS calls include ExternalId when provided
- Ensure backward compatibility when external ID not specified
- Add specific test for BedrockConverseLLM parameter extraction
- Extend existing dynamic parameter tests to include aws_external_id
2025-09-15 19:56:31 +02:00
Tim ElfrinkandGitHub 30c3e7b3d3 Fix: Bedrock cross-region inference profile cost calculation (#14566)
* Add tests for Bedrock cross-region inference profile mapping

- Test model mapping lookup works correctly
- Test proxy cost calculation scenario reproduces original issue
- Verify cost calculation returns expected values
- Ensure compatibility with existing test patterns

* Fix Bedrock cross-region inference profile cost calculation

- Add mapping for bedrock/us.anthropic.claude-3-5-haiku-20241022-v1:0
- Sync backup file for local testing consistency
- Resolve proxy spend tracking failures for cross-region profiles
- Maintain identical configuration with standalone profile

Fixes #14458
2025-09-15 07:10:20 -07:00
Sameer KankuteandGitHub 110ce543c2 [Feat]Add cancel endpoint support for openai and azure (#14561)
* Add cancel endpoint support for openai
 and azure

* fix lint error

* fix cancel url contruction azure

* readd changes
2025-09-15 07:08:56 -07:00
Sameer KankuteandGitHub 7fd6e62570 Fix unsupported stop param for grok-code models (#14565) 2025-09-15 07:00:54 -07:00
iabhi4 4ba3a21042 fix(proxy): Correctly parse multi-part MCP server aliases from URL paths 2025-09-14 15:17:08 -07:00
Tim Elfrink 9521414efa Resolve merge conflict by including both CompactifAI and OVHCloud providers
- Keep CompactifAI provider detection logic
- Include new OVHCloud provider from main branch
- Both providers now work correctly with model prefix detection
2025-09-14 23:03:18 +02:00
Tim Elfrink 6ac37093e5 Update CompactifAI model references and move tests to unit test directory
- Update all model references from llama-2-7b-compressed to cai-llama-3-1-8b-slim
- Move CompactifAI tests from tests/llm_translation to tests/test_litellm/llms/compactifai/
- Update documentation examples to use the new model name
- Remove integration test inheritance to make tests pure mock tests

This addresses review feedback to use mock tests and updated model naming.
2025-09-14 23:00:27 +02:00
Krrish Dholakia 03f2be1e20 fix: fix race conditions 2025-09-14 09:41:04 -07:00
Krrish Dholakia 2c6481fa33 fix: remove incorrect test 2025-09-14 09:34:31 -07:00
Krrish Dholakia fc2d1f2646 fix: fix import errors 2025-09-14 09:32:21 -07:00
Krish DholakiaandGitHub 510332b886 Merge pull request #14491 from Rasmusafj/main
Resolve cache key collision issue where all soft budget alerts use identical cache keys
2025-09-14 00:51:27 -07:00
Krish DholakiaandGitHub 56fd60b140 Merge pull request #14494 from eliasto/feat/ovhcloud-ai-edpoints-provider
feat: Add OVHCloud AI Endpoints as a provider
2025-09-14 00:45:08 -07:00
Krish DholakiaandGitHub 37aa8fff5e Merge pull request #14548 from nearai-cloud/fix/completion-chat-id
fix: completion chat id
2025-09-14 00:42:04 -07:00
Krrish Dholakia 7ef8c808cf docs: update doc 2025-09-14 00:28:11 -07:00
Krish DholakiaandGitHub db644b6edd Merge pull request #14500 from luisfucros/feat/update-sambanova-models
Add sambanova deepseek v3.1 and gpt-oss-120b models
2025-09-13 23:38:20 -07:00
Krish DholakiaandGitHub 11822e63f1 Merge pull request #14519 from uc4w6c/feat/add_tools_permission_guardrail
feat: add tool-permission guardrail
2025-09-13 23:22:31 -07:00
Coffee ac0386ae1f fix: completion chat id 2025-09-14 14:21:57 +08:00
Krish DholakiaandGitHub dc4bbba0a5 Merge pull request #14520 from boopesh07/email_prometheus
Added user_email labels to the prometheus monitoring.
2025-09-13 23:19:46 -07:00
Krish DholakiaandGitHub 2ec4b2953c Merge pull request #14531 from mubashir1osmani/main
fix: DD tool calls passed in metadata
2025-09-13 23:16:15 -07:00
Krish DholakiaandGitHub 2338dd952e Merge pull request #14546 from BerriAI/filter-on-logs-bug
The 'last 24 hours' button shows up above the end user dropdown on Logs page
v1.77.2.rc.1
2025-09-13 23:12:39 -07:00
Krish DholakiaandGitHub 1622d03ecc Merge pull request #14545 from BerriAI/litellm_ui_qa_09_13_2025_p1
Litellm UI qa 09 13 2025 p1 - fix end user filtering + fix load mcp tool call error + prevent setting max user budget on scroll in edit user settings
2025-09-13 18:48:43 -07:00
Krrish Dholakia fe546b936a fix(fetch_mcp_tools.tsx): fix load mcp tools 2025-09-13 18:42:42 -07:00
Ishaan JaffandGitHub f37dd6bb95 Litellm 1.77.2 stable notes (#14544)
* fix release notes instructions

* docs v1

* fix doc

* fix highlights

* docs fix

* docs fix
2025-09-13 18:41:34 -07:00
Krrish Dholakia 84f934bf36 fix(user_edit_view.tsx): use shared component to prevent accidental setting of edit user budget 2025-09-13 18:20:07 -07:00
Krrish Dholakia bb5e71447d fix(spend_management_endpoints.py): add end user filtering 2025-09-13 18:15:34 -07:00
Yuta Saito 0c1abf1a55 fix: recompute filters after deleting an MCP Server 2025-09-14 09:19:23 +09:00
Boopesh ShanmugamandGitHub 95da14cb96 Docs update on user header mapping (#14527) 2025-09-13 16:58:10 -07:00
Ishaan JaffandGitHub 6172145014 fix: org budget update fix (#14541) 2025-09-13 16:34:26 -07:00
Krrish Dholakia b6bca6369f fix(constants.py): make default num workers 1 2025-09-13 13:57:58 -07:00
Ishaan Jaff 9110af37d0 bump: version 1.77.1 → 1.77.2 2025-09-13 13:55:30 -07:00
Krrish Dholakia dc4b09e26e build(ui/): new ui build 2025-09-13 13:48:27 -07:00
Krish DholakiaandGitHub 6efc898407 Merge pull request #14523 from BerriAI/litellm_dev_09_12_2025_p1
VLLM - transcription endpoint support + Ollama_chat/ - images, thinking, and content as list handling +
2025-09-13 13:39:28 -07:00
Ishaan Jaff 6c27e5ce43 fix: DEFAULT_NUM_WORKERS_LITELLM_PROXY v1.77.1-nightly 2025-09-13 12:11:33 -07:00
Ishaan Jaff 252ec8e1ae test_normal_router_call_tpm_v3 2025-09-13 12:04:56 -07:00
Alexsander HamirandGitHub 44d209622b fix: remove dynamic creation of static value (#14538) 2025-09-13 11:58:38 -07:00
Krrish Dholakia 8443000ca4 fix(main.py): route vllm calls via the openai sdk route
consistent with other openai-like implementations
2025-09-13 11:49:05 -07:00
Ishaan Jaff 26dafdc493 test fix: note this does not play nice with circleCI, it passes on local 2025-09-13 11:37:19 -07:00