Carlo Alberto Ferraris and GitHub
b50fcc4b56
vertex ai: use the correct domain for the global location when counting tokens ( #17116 )
2025-11-25 19:22:20 -08:00
Sameer Kankute and GitHub
cd65a84abd
Merge pull request #16844 from Chesars/fix/response-format-to-text-format-bridge-conversion
...
fix: Support response_format parameter in completion -> responses bridge
2025-11-26 08:51:09 +05:30
Ishaan Jaff and GitHub
5c192a23c3
[Feat] Add new RAG API on LiteLLM AI Gateway ( #17109 )
...
* init RAG api types
* add RAG endpoints
* init main.py for RAG ingest API
* init RecursiveCharacterTextSplitter
* add BaseRAGIngestion
* fix OpenAIRAGIngestion
* fix img handler
* init OpenAIRAGIngestion
* init BedrockRAGIngestion
* init BedrockRAGIngestion
* init rag tests
* init BedrockVectorStoreOptions
* implement BedrockRAGIngestion
* add BaseRAGAPI
* add endpoint for RAG ingest
* add ingest RAG endpoints
* add test doc
* add parse_rag_ingest_request
* update endpoints
* docs add docs for new RAG API
* fix qa check
* fix linting
* docs ficx
* docs
* add max depth checks
* docs anthropic
2025-11-25 17:54:29 -08:00
Otavio Brito and GitHub
6e5c7c0008
fix transcription exception handling - /audio/transcriptions ( #16791 )
...
* fix transcription exception handling
* reraise the exception
2025-11-25 16:41:35 -08:00
Krrish Dholakia
5cb5c2a7b7
docs: more doc cleanup
2025-11-25 16:04:27 -08:00
Krrish Dholakia
8ee6812edf
docs: cleanup launch post
2025-11-25 15:58:51 -08:00
Krrish Dholakia
70a1325847
docs: more doc cleanup
2025-11-25 15:01:22 -08:00
Kerem Turgutlu and GitHub
8637d74e17
include server_tool_use in streaming usage ( #16826 )
...
* include server_tool_use in streaming usage
* add test
2025-11-25 14:50:17 -08:00
Sam Chou and GitHub
c0288d81aa
Fix bedrock claude opus 4.5 inference profile - only global currently ( #17101 )
2025-11-25 14:49:12 -08:00
Krrish Dholakia
f3d5775920
fix: fix doc load issue
2025-11-25 14:40:26 -08:00
YutaSaito and GitHub
52f1bf1a80
fix: missing await ( #17103 )
2025-11-25 14:33:38 -08:00
Ishaan Jaff and GitHub
be712908a3
[Feat] Add OpenAI compatible bedrock imported models. - qwen etc ( #17097 )
...
* test_bedrock_openai_imported_model
* AmazonBedrockOpenAIConfig
* add openai route for bedrock
* docs fix
* fix code qa check
2025-11-25 12:20:39 -08:00
Krrish Dholakia
db2c8e3631
docs: initial doc cleanup
2025-11-25 11:57:51 -08:00
Sameer Kankute and GitHub
67622fb040
Add day 0 support for anthropic new feat ( #17091 )
...
* Added tool search support for anthropic
* Add programtic tool calling support
* Add tool use input examples support
* Add anthropic effort param support
* Add anthropic effort param support
* Add blog for new features
* fix mypy and lint errors
* fix mypy and lint errors
* fix mypy and lint errors
* fix mypy and lint errors
* Add better handling
* Add better handling
2025-11-25 11:28:47 -08:00
Sameer Kankute and GitHub
3249f6dd2d
Merge pull request #17070 from BerriAI/litellm_add_vertex_ai_image_support
...
Add vertex ai image gen support for both gemini and imagen models
2025-11-26 00:04:03 +05:30
Sameer Kankute and GitHub
83a9dcd2d2
Merge pull request #16886 from BerriAI/litellm_anthopic_azure_support
...
Added support for azure anthopic models via chat completion
2025-11-26 00:03:52 +05:30
Sameer Kankute
59b4b9a07c
fix documentation of anthropic azure
2025-11-26 00:02:48 +05:30
Sameer Kankute and GitHub
59bcf079fb
Merge pull request #17078 from BerriAI/litellm_add_search_logging
...
Add search API logging and cost tracking in LiteLLM Proxy
2025-11-25 23:59:41 +05:30
Sameer Kankute and GitHub
2e50db81a5
Merge pull request #17071 from BerriAI/litellm_azure_gpt_5_reasoning
...
Fix `reasoning_effort="none"` not working on Azure for GPT-5.1
2025-11-25 23:59:25 +05:30
00e17c81a1
Add enforce user param functionality ( #17088 )
...
* feat: Add reject_metadata_tags to proxy config
Co-authored-by: krrishdholakia <krrishdholakia@gmail.com >
* Refactor: Rename reject_metadata_tags to reject_clientside_metadata_tags
Co-authored-by: krrishdholakia <krrishdholakia@gmail.com >
---------
Co-authored-by: Cursor Agent <cursoragent@cursor.com >
2025-11-25 09:36:24 -08:00
Sameer Kankute
1c612288bc
fix lint errors
2025-11-25 20:20:09 +05:30
Sameer Kankute
255d1bc239
fix lint errors
2025-11-25 20:20:09 +05:30
Sameer Kankute and GitHub
e0396e5fa7
Merge pull request #17082 from BerriAI/main
...
merge main
2025-11-25 18:49:52 +05:30
Sameer Kankute
e2f2ccd913
Add tests related messages api
2025-11-25 18:45:51 +05:30
Sameer Kankute
dd4c8ecbef
Add v1/messages support for azure anthropic models
2025-11-25 18:36:39 +05:30
Sameer Kankute
afe540e88d
Fix auth issue
2025-11-25 18:26:25 +05:30
Sameer Kankute
67d69d12b0
Add cost tracking and logging support
2025-11-25 17:14:59 +05:30
Sameer Kankute
c149ade6a8
Add tests related to reasoning param none
2025-11-25 13:57:15 +05:30
Sameer Kankute
a50083a87b
Remove none support from reasoning param
2025-11-25 13:56:30 +05:30
Sameer Kankute
b0d511143c
remove unsused imports
2025-11-25 13:36:20 +05:30
Sameer Kankute
883cfaeeaf
Add tests
2025-11-25 13:32:13 +05:30
Sameer Kankute
f52f05748d
Update docs related to vertex ai image gen
2025-11-25 13:31:50 +05:30
Sameer Kankute
29ab291cf5
Add vertex ai image support
2025-11-25 13:31:16 +05:30
wcyat and GitHub
6dcb5425a5
fix(vertex): fix CreateCachedContentRequest enum error ( #16965 )
...
* feat: add _fix_enum_types function to remove enums from non-string fields in schema
* test: add test for _fix_enum_types function to validate enum removal from non-string fields
2025-11-24 21:24:29 -08:00
yuneng-jiang and GitHub
babee43dde
Merge pull request #17068 from BerriAI/litellm_additional_delete_resource_modal
...
[Feature] Change Delete Modals to Common Component
2025-11-24 20:58:46 -08:00
Dmitrii Komarov and GitHub
046b7efbbe
Make Bedrock image generation more consistent ( #17021 )
2025-11-24 20:58:01 -08:00
Saar wintrov and GitHub
cfd35d3b14
Metadata: fix 401 when audio/transcriptions ( #17023 )
...
* Metadata: fix 401 when audio/transcriptions
* check if str, CR fixes
2025-11-24 20:56:27 -08:00
Cesar Garcia and GitHub
650b18974f
fix(gemini): skip thinking config for image models ( #17027 )
...
* fix(gemini): exclude image models from automatic thinking_level parameter (#17013 )
- gemini-3-pro-image-preview does not support thinking_level parameter
- Added check to skip adding thinkingConfig for models containing "image"
- Fixes BadRequestError: "Thinking level is not supported for this model"
- Only affects automatic default behavior, user can still pass reasoning_effort explicitly
Fixes #17013
* test: add tests for gemini-3 image models thinking_level exclusion
* update docs
2025-11-24 20:54:12 -08:00
yuneng-jiang and GitHub
3aba6d96fd
[Fix] UI - Add No Default Models for Team and User Settings ( #17037 )
...
* Add No Default Models to Team and User settings
* Removing unused imports
* Adding to Create User and Team flow
2025-11-24 20:53:17 -08:00
Saar wintrov and GitHub
777ef628d2
Enhancement(helm): ServiceMonitor template rendering ( #17038 )
...
* Metadata: fix 401 when audio/transcriptions
* check if str, CR fixes
* Added new helmchart functionality
* .
* .
* adding new tests
2025-11-24 20:53:02 -08:00
597fa4d35c
Fix image edit endpoint ( #17046 )
...
* Fix image edit endpoint
* Update litellm/proxy/image_endpoints/endpoints.py
Co-authored-by: Copilot <175728472+Copilot@users.noreply.github.com >
---------
Co-authored-by: Copilot <175728472+Copilot@users.noreply.github.com >
2025-11-24 20:52:35 -08:00
yuneng-jiang and GitHub
d2b3ef0667
Add aws_bedrock_runtime_endpoint into Credential Types ( #17053 )
2025-11-24 20:48:51 -08:00
1ae80955e8
Docs: Add link to logging payload spec ( #17049 )
...
Co-authored-by: Cursor Agent <cursoragent@cursor.com >
2025-11-24 20:48:10 -08:00
yuneng-jiang and GitHub
3f5a34d72c
Deleting a user from team deletes key user created for team ( #17057 )
2025-11-24 20:47:43 -08:00
yuneng-jiang and GitHub
e371ff454a
Non root docker build fix ( #17060 )
2025-11-24 20:45:56 -08:00
yuya_matsuba and GitHub
262fb742d2
Fix: Distinguish permission errors from idempotent errors in Prisma migrations ( #17064 )
...
* fix: distinguish permission errors from idempotent errors in Prisma migrations
* style: apply Black formatting and fix line length issues
2025-11-24 20:41:44 -08:00
Raghav Jhavar and GitHub
bd8196f982
(fix) propagate x-litellm-model-id in responses ( #16986 )
...
* propagate model id on errors too
* make it work for messages and streaming
* fix
* cleanup
* cleanup
* final
* cleanup
* clean up method name and fix responses api streaming
* remove comment
2025-11-24 20:40:43 -08:00
yuneng-jiang
d53bc7b9a0
Change modals to reusable component
2025-11-24 20:37:33 -08:00
Sameer Kankute and GitHub
282ac87617
Add temperature support for 5.1 models ( #17011 )
2025-11-24 18:54:22 -08:00
Sameer Kankute and GitHub
fc219c7db8
Integrate eleven labs text-to-speech ( #16573 )
...
* Add elevenlaps tts support
* fix mypy error
* add simple usage in docs
2025-11-24 18:49:30 -08:00