Ishaan Jaffer
d9b85ab276
fix: rename search_provider
2025-10-21 17:42:18 -07:00
Ishaan Jaffer
bea8e13a94
fix: GuardrailConfigModel
2025-10-21 17:11:56 -07:00
wangjifeng and GitHub
8cbaec0310
feat: Add imageConfig parameter for gemini-2.5-flash-image ( #15530 )
...
* Add imageConfig parameter support for Vertex AI to enable gemini-2.5-flash-image model requirements
* Add test for imageConfig parameter support in Vertex AI Gemini transformation
2025-10-21 17:08:30 -07:00
Ishaan Jaff and GitHub
7b939b4558
[Feat] Add EXA AI Search API to LiteLLM ( #15774 )
...
* add BaseSearchConfig
* add BaseSearchConfig
* validate_environment
* fix handlers
* add PerplexitySearchConfig
* add PerplexitySearchConfig
* add LiteLLM Search API module.
* add BaseSearchConfig
* add _build_search_optional_params
* add search_testing
* add BaseSearchTest
* add TestPerplexitySearch
* fix BASE
* fix handler
* add search API
* add to init
* fix: working perplexity search API
* add _hidden_params to search
* add TAVILY to LlmProviders
* add TavilySearchConfig
* add TavilySearchConfig
* TestTavilySearch
* add tavily transform
* TestParallelAISearch
* add LlmProviders
* add ParallelAISearchConfig
* add ParallelAISearchConfig
* ParallelAISearchConfig
* add EXA AI Search API
* add ExaAISearchConfig
* TestExaAISearch
* add get_supported_perplexity_optional_params
* add Exa AI Search API
* add transform_search_request
* add ExaAISearchConfig
* fix linting errors
* transform_search_request
2025-10-21 17:06:23 -07:00
Ishaan Jaff and GitHub
208f76f8ad
[Feat] Add Parallel AI - Search API ( #15772 )
...
* add BaseSearchConfig
* add BaseSearchConfig
* validate_environment
* fix handlers
* add PerplexitySearchConfig
* add PerplexitySearchConfig
* add LiteLLM Search API module.
* add BaseSearchConfig
* add _build_search_optional_params
* add search_testing
* add BaseSearchTest
* add TestPerplexitySearch
* fix BASE
* fix handler
* add search API
* add to init
* fix: working perplexity search API
* add _hidden_params to search
* add TAVILY to LlmProviders
* add TavilySearchConfig
* add TavilySearchConfig
* TestTavilySearch
* add tavily transform
* TestParallelAISearch
* add LlmProviders
* add ParallelAISearchConfig
* add ParallelAISearchConfig
* ParallelAISearchConfig
2025-10-21 17:00:05 -07:00
Ishaan Jaff and GitHub
b9f3f9fb79
[Feat] Add Tavily Search API ( #15770 )
...
* add BaseSearchConfig
* add BaseSearchConfig
* validate_environment
* fix handlers
* add PerplexitySearchConfig
* add PerplexitySearchConfig
* add LiteLLM Search API module.
* add BaseSearchConfig
* add _build_search_optional_params
* add search_testing
* add BaseSearchTest
* add TestPerplexitySearch
* fix BASE
* fix handler
* add search API
* add to init
* fix: working perplexity search API
* add _hidden_params to search
* add TAVILY to LlmProviders
* add TavilySearchConfig
* add TavilySearchConfig
* TestTavilySearch
* add tavily transform
2025-10-21 16:59:29 -07:00
wenhua and GitHub
b0ccc35a9c
fix(ollama): Enhance chunk parsing for empty responses without 'thinking' and improve error logging ( #13333 ) ( #15717 )
2025-10-21 16:59:01 -07:00
Ishaan Jaff and GitHub
e1cb92862e
[Feat] Add def search() APIs for Web Search - Perplexity API ( #15769 )
...
* add BaseSearchConfig
* add BaseSearchConfig
* validate_environment
* fix handlers
* add PerplexitySearchConfig
* add PerplexitySearchConfig
* add LiteLLM Search API module.
* add BaseSearchConfig
* add _build_search_optional_params
* add search_testing
* add BaseSearchTest
* add TestPerplexitySearch
* fix BASE
* fix handler
* add search API
* add to init
* fix: working perplexity search API
* add _hidden_params to search
2025-10-21 16:58:51 -07:00
Ishaan Jaff and GitHub
9135e748a0
[Feat ] /ocr - Add mode + Health check support for OCR models ( #15767 )
...
* get_mode_handlers
* use get_mode_handlers
* test_ahealth_check_ocr
* Add OCR mode to test models
* docs OCR Health Checks
* fix connection endpoint
2025-10-21 16:58:37 -07:00
Javier Garcia and GitHub
b0a3a7c4fb
Add details in docs ( #15721 )
...
* Add details in docs
* add logic to set span attributes and unit tests
* Restore html files
* Remove html files
* Remove html files
2025-10-21 16:57:51 -07:00
Thomas Mildner and GitHub
1cfc4624c3
[Feat] Add SENTRY_ENVIRONMENT configuration for Sentry integration ( #15760 )
...
* [Feat] Add SENTRY_ENVIRONMENT configuration for Sentry integration and corresponding tests
* [Refactor] Enhance test_sentry_environment by mocking sentry_sdk and improving environment handling
* [Fix] Update default SENTRY_ENVIRONMENT to 'production' and enhance test for Sentry integration
* [Fix] Update test_sentry_environment to verify correct handling of SENTRY_ENVIRONMENT values
* [Fix] Update test_sentry_environment to assert correct handling of production environment
2025-10-21 16:40:55 -07:00
Kowyo and GitHub
1fcadd6c05
feat(ollama): set 'think' to False when reasoning effort is not high/medium/low ( #15763 )
2025-10-21 16:39:08 -07:00
Krrish Dholakia
2a1dbb5b9e
docs(creating_adapters.md): document how to write an adapter
2025-10-21 16:20:31 -07:00
nuernber and GitHub
353dfb1238
Add AWS us-gov-west-1 Claude 3.7 Sonnet costs ( #15775 )
...
* add us-gov-west-1 claude 3.7 sonnet to prices
* add to _backup file as well
2025-10-21 16:17:07 -07:00
YutaSaito and GitHub
39641e7e68
chore: rename GraySwan to Gray Swan ( #15771 )
2025-10-21 15:18:55 -07:00
Vinod Singh and GitHub
d4aadda692
Auth Header Fix for MCP Tool Call ( #15736 )
...
* fixed the Auth header for MCP Tool Call
* Final fix for Auth header
* testcase for mcp_auth_header_extraction, insensitive_alias_matching, insensitive_servername_matching added
2025-10-21 13:58:03 -07:00
Krrish Dholakia
1e0368521e
refactor: cleanup
2025-10-21 13:46:19 -07:00
Ishaan Jaffer
3741c43396
docs fix
2025-10-21 13:21:59 -07:00
Ishaan Jaff and GitHub
8ad9bbbd02
[Docs] Add Azure AI - OCR to docs ( #15768 )
...
* add Azure OCR to docs
* docs fix
* docs fix
* docs fix
* docs OCR
2025-10-21 13:10:45 -07:00
Ishaan Jaffer
185182bebc
Revert "add Azure OCR to docs"
...
This reverts commit a3699e28a4 .
2025-10-21 13:02:31 -07:00
Ishaan Jaffer
a3699e28a4
add Azure OCR to docs
2025-10-21 13:02:21 -07:00
Ishaan Jaffer
6605aba307
docs grayswan
2025-10-21 11:16:25 -07:00
YutaSaito and GitHub
d79bdd491f
feat: add GraySwan Guardrails support ( #15756 )
2025-10-21 11:13:50 -07:00
46d55bd92a
fix: Add response_type + PKCE parameters to OAuth authorization endpoint ( #15720 )
...
* fix: Add response_type parameter to OAuth authorization endpoint
Fixes #15684
OAuth providers like Google require the response_type parameter during
the authorization flow. This commit adds response_type=code to the
authorization redirect parameters, which is required by the OAuth 2.0
specification (RFC 6749 Section 4.1.1).
Changes:
- Added response_type=code to authorization params in discoverable_endpoints.py
- Added test coverage for the response_type parameter
🤖 Generated with [Claude Code](https://claude.com/claude-code )
Co-Authored-By: Claude <noreply@anthropic.com >
* fix oauth flow by forwarding code_challenge and forwarding code_verifier
---------
Co-authored-by: Claude <noreply@anthropic.com >
2025-10-21 09:43:19 -07:00
Tom Haynes and GitHub
98f1d63508
use correct otel logger, and normalise otel paths ( #15645 )
v1.78.6-nightly
2025-10-21 09:16:03 -07:00
Ishaan Jaffer
8b522d88a2
is_llm_api_route
2025-10-20 18:05:35 -07:00
Ishaan Jaff and GitHub
92335d991c
[Feat] Add Azure AVA (Speech AI) Cost Tracking ( #15754 )
...
* add azure/speech/ cost tracking
* test_azure_ava_tts_async
* add azure/speech to model cost map
* docs cost tracking
* docs tts AVA
* add azure/speech/azure-tts
2025-10-20 18:01:51 -07:00
Ishaan Jaffer
60fab591db
rename test files
2025-10-20 18:00:17 -07:00
Ishaan Jaff and GitHub
157739da01
[Bug]: Fix Incorrect status value in responses api with gemini ( #15753 )
...
* _map_chat_completion_finish_reason_to_responses_status
* test_transform_chat_completion_response_with_reasoning_content
* test_transform_chat_completion_response_output_item_status
2025-10-20 17:58:56 -07:00
Ishaan Jaffer
5ce2be732e
get_provider_text_to_speech_config
2025-10-20 17:10:09 -07:00
Ishaan Jaffer
9a25eeccb2
docs fix
2025-10-20 17:02:38 -07:00
Ishaan Jaffer
c9152003bd
bump V
2025-10-20 16:55:03 -07:00
Ishaan Jaff and GitHub
73a23a6c78
[Feat] Add Azure AVA TTS integration ( #15749 )
...
* add AzureBaseIssueTokenHandler
* add BaseTextToSpeechConfig
* async_text_to_speech_handler
* add AzureAVATextToSpeechConfig
* add get_provider_text_to_speech_config
* add AzureAVATextToSpeechConfig
* fixes for base_llm_http_handler
* fix transform_text_to_speech_request
* test_azure_ava_tts_async
* test_azure_ava_tts_async
* fix TextToSpeechRequestData
* fix transform_text_to_speech_request
* add text_to_speech_handler in LLMHttpHandler
* remove old file
* fix transform_text_to_speech_request
* fix dispatch_text_to_speech
* fix azure TTS
* fix AVA TTS
* fix transform
* fix linting
* ci/cd - use one job for audio testing
* fix tests
* fix llm http handler debugging
* unit tests azure tts
* docs Azure speech
* docs fix
* docs azure AVA
* docs azure AVA
* fix handlers
* test_async_realtime_uses_max_size_parameter
2025-10-20 16:52:23 -07:00
akraines and GitHub
41a6ecd5b6
Change max_tokens value to match max_output_tokens for claude sonnet 4.5: 64000 ( #15715 )
...
See https://github.com/RooCodeInc/Roo-Code/issues/8454
2025-10-20 16:11:36 -07:00
Ishaan Jaff and GitHub
0c25b1a256
[Fix] OpenAI Realtime API integration fails due to websockets.exceptions.PayloadTooBig error ( #15751 )
...
* fix REALTIME_WEBSOCKET_MAX_MESSAGE_SIZE_BYTES
* edit max_size for websockets
* fix AzureOpenAIRealtime
2025-10-20 15:54:14 -07:00
Sameer Kankute and GitHub
1fb798f81d
(Bug) Fix JSON serialization error in Helicone logging by removing OpenTelemetry span from metadata ( #15728 )
...
* remove span object from helicon metadata
* Add test
2025-10-20 08:53:22 -07:00
Sameer Kankute and GitHub
3955a3de5d
fix the wrong request body in json mode doc ( #15729 )
2025-10-20 08:44:14 -07:00
Timothée Lecomte and GitHub
3ef9b2015a
feat: read from custom-llm-provider header ( #15528 )
2025-10-18 22:04:53 -07:00
4a74190c12
Fix: Add gpt 4.1 pricing for response endpoint ( #15593 )
...
* Add gpt41, gpt-41-mini, and gpt-41-nano to pricing and context window json
* Add gpt-41s to azure_llms dict
* Undo json changes
---------
Co-authored-by: IQHL (Hans Jacob Landelius) <iqhl@novnordisk.com >
2025-10-18 22:04:14 -07:00
Lucas Sugi and GitHub
ae86862e74
fix: Add function responsible to call precall ( #15636 )
...
* fix: Add function responsible to call precall
* fix: Set correct route_type
2025-10-18 22:01:14 -07:00
Lucas Sugi and GitHub
ce9e22688d
fix: Add pre and post call for list batches ( #15673 )
2025-10-18 21:52:35 -07:00
f55745fc5e
[Fix] Forward anthropic-beta headers to Bedrock, VertexAI ( #15700 )
...
* [Fix] Forward anthropic-beta headers to Bedrock and other cross-provider scenarios (#15623 )
* add_provider_specific_headers_to_request
* fix add_provider_specific_headers_to_request
* test_provider_specific_header_multi_provider
* test_provider_specific_header_in_request
---------
Co-authored-by: Jack Venberg <jack.venberg@rover.com >
2025-10-18 16:26:32 -07:00
Alexsander Hamir and GitHub
441aed2c87
fix: update worker recommendation ( #15702 )
2025-10-18 16:24:32 -07:00
Ishaan Jaffer
3fc49a029f
docs fix
2025-10-18 15:28:49 -07:00
Ishaan Jaffer
eed1ddba49
docs v1.78.5-stable
2025-10-18 15:28:07 -07:00
Ishaan Jaff and GitHub
6ab9b0af9f
[Fix] Anthropic cache_control incorrectly applied to all content items instead of last item only ( #15699 )
...
* fix: _safe_insert_cache_control_in_message
* test_anthropic_cache_control_hook_system_message
* docs prompt cache injection
* docs fix
2025-10-18 15:18:08 -07:00
Jason Roberts and GitHub
c471bf1f16
feat(guardrails): Add content masking and streaming support to PANW Prisma AIRS guardrail ( #15666 )
...
* feat(guardrails): Add content masking and streaming support to PANW Prisma AIRS
- Add mask_request_content and mask_response_content parameters
- Implement content masking for prompts and responses
- Add streaming support with real-time masking
- Add comprehensive test coverage (28 tests)
- Update documentation with masking examples and security notes
* fix(guardrails): Fix PANW Prisma AIRS env var fallback and text completion support
v1.78.5-stable
v1.78.5.rc.1
v1.78.5-nightly
2025-10-18 13:57:51 -07:00
YutaSaito and GitHub
645f84c02e
fix: add imagePullSecrets to migrations-job ( #15681 )
2025-10-18 13:56:31 -07:00
katsuhiro muto and GitHub
d5e686b3e8
[Fix] Support service_tier in chat completion ( #15693 )
...
* Support service_tier
* fix test
2025-10-18 13:55:54 -07:00
Krish Dholakia and GitHub
c1355e92dc
fix(proxy_server.py): re-encrypt env var on config save + use original value on decrypt error ( #15671 )
...
* fix(proxy_server.py): re-encrypt env var on config save + use original value on decrypt error
Closes https://github.com/BerriAI/litellm/issues/14854
Fixes https://github.com/BerriAI/litellm/issues/13406
* docs: email.md
document PROXY_BASE_URL param
* fix(proxy_server.py): pop model list before writing to db
2025-10-18 13:39:25 -07:00