Krish Dholakia and GitHub
8cbc85578e
Merge pull request #15351 from BerriAI/litellm_dev_10_08_2025_p3
...
SSO - support EntraID app roles
2025-10-08 19:04:00 -07:00
Krish Dholakia and GitHub
adbdf9dae9
Merge pull request #14470 from BerriAI/litellm_dev_09_11_2025_p1
...
AzureAD Default credentials - select credential type based on environment
2025-10-08 19:03:07 -07:00
Krish Dholakia and GitHub
12a1d081ee
Merge branch 'main' into litellm_dev_09_11_2025_p1
2025-10-08 19:02:58 -07:00
Krrish Dholakia
697f99c1d5
docs: doc improvements
2025-10-08 18:59:00 -07:00
Krrish Dholakia
8eafb7a8fa
docs: doc improvements
2025-10-08 18:50:42 -07:00
Krrish Dholakia
ff3f356416
docs(docs/): add app role support to docs
2025-10-08 18:46:28 -07:00
Krrish Dholakia
d28ecbc900
feat(ui_sso.py): support mapping app roles from azure entra id to litellm user roles
...
Closes LIT-1228
2025-10-08 18:45:13 -07:00
Ishaan Jaffer
c4022ade49
test mapped tests MCP
2025-10-08 18:34:30 -07:00
4226314096
Add native Responses API support for litellm_proxy provider ( #15347 )
...
* Initial plan
* Add native Responses API support for litellm_proxy provider
Co-authored-by: ishaan-jaff <29436595+ishaan-jaff@users.noreply.github.com >
---------
Co-authored-by: copilot-swe-agent[bot] <198982749+Copilot@users.noreply.github.com >
Co-authored-by: ishaan-jaff <29436595+ishaan-jaff@users.noreply.github.com >
2025-10-08 18:31:26 -07:00
Ishaan Jaff and GitHub
2f42c806cb
[Fix] x-litellm-cache-key header not being returned on cache hit ( #15348 )
...
* fix: x-cache-key
* test_cache_key_in_hidden_params_acompletion
* fix: remove_cache_control_flag_from_messages_and_tools
2025-10-08 18:10:43 -07:00
Ishaan Jaff and GitHub
97031dc8ee
Fix - (openrouter): move cache_control to content blocks for claude/gemini ( #15345 )
...
* test_openrouter_transform_request_with_cache_control
* fix CacheControlSupportedModels
* test_openrouter_transform_request_with_cache_control_gemini
2025-10-08 17:41:04 -07:00
Krish Dholakia and GitHub
0a507c37ea
Merge pull request #15344 from sandeshghanta/patch-1
...
Add gpt-5-pro-2025-10-06 to model costs
2025-10-08 17:23:09 -07:00
Pierre-Emmanuel MERCIER and Ishaan Jaffer
861750790b
feat: add redis ssl and username support ( #11319 )
2025-10-08 16:48:07 -07:00
Sandesh Ghanta and GitHub
e4317030bf
Add gpt-5-pro-2025-10-06 to model costs
2025-10-08 16:44:40 -07:00
Krish Dholakia and GitHub
5c6fb5eb29
Merge pull request #15309 from BerriAI/litellm_cookie_spasm_mitigation
...
potentially fixes a UI spasm issue with an expired cookie
2025-10-08 16:13:38 -07:00
Ishaan Jaff and GitHub
1c56a0d856
[Fix] Watsonx - Apply correct prompt templates for openai/gpt-oss model family ( #15341 )
...
* fix: apply_prompt_template
* Revert "fix: apply_prompt_template"
This reverts commit 3e0e40b497 .
* add apply_prompt_template for WatsonX
* feat: add apply_prompt_template
* test_watsonx_gpt_oss_prompt_transformation
* Revert "add apply_prompt_template for WatsonX"
This reverts commit 3e80903796e39d4fe7206d63445e024c3ad8d0c4.
* add apply_prompt_template for WatsonX
* fix apply_prompt_template
* fix: add hf template handler
* fix hf_chat_template
* fix _get_tokenizer_config
* fix hf_chat_template
* add WatsonXModelPattern
* fix aapply_prompt_template
2025-10-08 15:39:36 -07:00
Achintya Rajan
fe2d4addfa
Update CreateKeyPage.expiredToken.test.tsx
2025-10-08 14:10:09 -07:00
Ishaan Jaffer
9d84a7cc8b
Revert "fix: apply_prompt_template"
...
This reverts commit 3e0e40b497 .
2025-10-08 13:56:47 -07:00
Ishaan Jaffer
3e0e40b497
fix: apply_prompt_template
2025-10-08 13:56:25 -07:00
Felipe Garé and GitHub
f12b8bf2c1
feat(files): add @client decorator to file operations ( #15339 )
...
- Decorate sync/async file APIs (retrieve, delete, list, content) with @client
- Ensures calls route through the configured client, honoring provider/client config
- Standardizes behavior across file endpoints and aligns with existing patterns
- No API signature changes; improves consistency and client-bound usage
2025-10-08 13:42:34 -07:00
Achintya Rajan
d6852b11ae
Create CreateKeyPage.expiredToken.test.tsx
2025-10-08 13:41:00 -07:00
Achintya Rajan
7595acff42
empty commit for CI/CD
2025-10-08 12:03:44 -07:00
0b6b69cd5b
fix issue with parsing assistant messages ( #15320 )
...
* Update factory.py
* reinclude the test cases
* fixed test
---------
Co-authored-by: Weijie <weijie-tan3@github.com >
2025-10-08 09:49:02 -07:00
Krish Dholakia and GitHub
13703f289b
Merge pull request #15292 from timelfrink/fix/bedrock-prompt-caching-cost-calculation
...
fix(bedrock): include cacheWriteInputTokens in prompt_tokens calculation
2025-10-07 19:11:45 -07:00
Krish Dholakia and GitHub
3f4d2c6ada
Add Cohere Embed v4 support for AWS Bedrock
...
Add Cohere Embed v4 support for AWS Bedrock
2025-10-07 19:10:22 -07:00
Ishaan Jaffer
5d6cd3ffed
UI new build
2025-10-07 18:59:11 -07:00
Ishaan Jaffer
a8243079d3
doc fix
2025-10-07 18:51:27 -07:00
Ishaan Jaff and GitHub
8a8cc5a1d3
[QA/Fixes] - Dynamic Rate Limiter v3 - final QA ( #15311 )
...
* fix: PriorityReservationSettings
* fix: use correct casting
* fix type casting
2025-10-07 18:42:17 -07:00
Krish Dholakia and GitHub
1238464e20
Merge pull request #15303 from BerriAI/litellm_tenacity_upgrade
...
Upgrades tenacity version to 8.5.0
2025-10-07 18:19:52 -07:00
Krish Dholakia and GitHub
60fb8cdb87
Merge pull request #15308 from BerriAI/litellm_models_page_crash_fix
...
fix: model + endpoints page crash when config file contains router_settings.model_group_alias
2025-10-07 18:19:15 -07:00
Achintya Rajan
4de4e4af1a
Update page.tsx
2025-10-07 18:14:20 -07:00
Ishaan Jaffer
7b82473bfb
fix gpt-image-1-mini
2025-10-07 17:56:51 -07:00
Ishaan Jaffer
e1ab3620ee
fix: mapped tests
2025-10-07 17:55:52 -07:00
Ishaan Jaffer
8590646b84
fix code QA check
2025-10-07 17:49:57 -07:00
Ishaan Jaffer
f2a96e6830
fix: e2eUI testing
2025-10-07 17:45:30 -07:00
Achintya Rajan
9fa8c6099e
Update model_group_alias_settings.tsx
2025-10-07 17:42:04 -07:00
Ishaan Jaff and GitHub
36c971a6fd
[MCP Gateway] QA/Fixes - Ensure Team/Key level enforcement works for MCPs ( #15305 )
...
* fix: _set_object_permission
* fix: _set_object_permission on teams
* fix: _set_object_permission
* fixes for team/key permissions
* statsh: object permission view
* fix: MCPServerPermissions
* fix: _get_team_object_permission
* test mcp checks for permissions
* fix server checks with prefix names
* test_list_tools_strips_prefix_when_matching_permissions
* ruff fix
* docs - refactor MCP
* docs update MCP docs
* docs allowed tools
2025-10-07 17:34:48 -07:00
Ishaan Jaff and GitHub
7b56ba240e
[MCP Gateway] Litellm mcp fixes team control ( #15304 )
...
* fix: _set_object_permission
* fix: _set_object_permission on teams
* fix: _set_object_permission
* fixes for team/key permissions
* statsh: object permission view
* fix: MCPServerPermissions
2025-10-07 16:48:00 -07:00
Alexsander Hamir and GitHub
49e04e0217
[Fix] Networking: remove limitations ( #15302 )
...
* fix: remove limitations
* fix: linter issues
2025-10-07 16:45:42 -07:00
Achintya Rajan
f72f21b261
Update requirements.txt
2025-10-07 16:18:48 -07:00
Ishaan Jaff and GitHub
bc26eff98f
Fix: Make PATCH /model/{model_id}/update handle team_id consistently with POST /model/new ( #15297 )
...
* fix: _update_team_model_in_db
* test_patch_model_with_team_id_creates_proper_setup
2025-10-07 14:04:08 -07:00
Tim Elfrink
d71d801e4d
Add Cohere Embed v4 support for AWS Bedrock
...
- Add cohere.embed-v4:0 to model pricing configs
- Update bedrock_embedding_models constant
- Update documentation with v4 model support
Fixes #15272
2025-10-07 22:08:12 +02:00
Ishaan Jaff and GitHub
07a17d6d6b
[Feat] Proxy CLI - dont store existing key in the URL, store it in the state param ( #15290 )
...
* Feat: CLI Auth fixes for UI SSO
* fix auth.py
* fix test ui sso.py
2025-10-07 12:36:17 -07:00
Tim Elfrink
c5eb22381d
fix(bedrock): include cacheWriteInputTokens in prompt_tokens calculation
...
Fixes #15263
This PR fixes the cost calculation for Bedrock Anthropic models with prompt caching.
**Root Cause:**
PR #9838 incorrectly removed adding `cacheWriteInputTokens` to `prompt_tokens`
for Bedrock, based on the assumption that it would cause double counting (similar
to an Anthropic API issue). However, Bedrock's token structure is different:
- **Bedrock API**: `inputTokens`, `cacheReadInputTokens`, and `cacheWriteInputTokens`
are ALL separate values that should be summed for total input tokens
- **Anthropic API**: Same structure - all three token types are separate
The fix in #9838 was later reverted for Anthropic (correctly re-adding
`cache_creation_input_tokens` to `prompt_tokens`), but Bedrock was never fixed.
**Changes:**
1. Re-add `cacheWriteInputTokens` to `input_tokens` in Bedrock transformation
2. Update test assertions to reflect correct behavior
3. Add regression test for prompt caching cost calculation
4. Fix typo in Anthropic transformation where `cache_creation_tokens` was
incorrectly set to `cache_read_input_tokens`
**Testing:**
- All existing Bedrock transformation tests pass
- New test validates correct cost calculation with prompt caching
- Verified costs are non-negative and accurate
2025-10-07 20:28:46 +02:00
Sameer Kankute and GitHub
51971f4750
Add gpt-realtime-mini support ( #15283 )
2025-10-07 11:27:04 -07:00
Sameer Kankute and GitHub
e4892735f0
fix gemini cli by actually streaming the response ( #15264 )
...
* fix gemini cli by actually streaming the response
* fix cost tracking
* fix test
2025-10-07 11:26:39 -07:00
Sameer Kankute and GitHub
73f96712f5
fix the reasoningresponse id ( #15265 )
2025-10-07 11:24:29 -07:00
Krrish Dholakia
421d38c94a
build(ui/): build new ui
v1.77.7.dev.3
2025-10-07 10:48:19 -07:00
Krish Dholakia and GitHub
f044eb80de
Merge pull request #15285 from BerriAI/litellm_infinity_new_provider_ui
...
feature: adds Infinity as a provider in the UI
2025-10-07 10:46:05 -07:00
Achintya Rajan
e2f21beb7f
added Infinity as a provider in the UI
2025-10-07 10:21:18 -07:00