Commit Graph
26697 Commits
Author SHA1 Message Date
Ishaan Jaffer d9b85ab276 fix: rename search_provider 2025-10-21 17:42:18 -07:00
Ishaan Jaffer bea8e13a94 fix: GuardrailConfigModel 2025-10-21 17:11:56 -07:00
wangjifengandGitHub 8cbaec0310 feat: Add imageConfig parameter for gemini-2.5-flash-image (#15530)
* Add imageConfig parameter support for Vertex AI to enable gemini-2.5-flash-image model requirements

* Add test for imageConfig parameter support in Vertex AI Gemini transformation
2025-10-21 17:08:30 -07:00
Ishaan JaffandGitHub 7b939b4558 [Feat] Add EXA AI Search API to LiteLLM (#15774)
* add BaseSearchConfig

* add BaseSearchConfig

* validate_environment

* fix handlers

* add PerplexitySearchConfig

* add PerplexitySearchConfig

* add LiteLLM Search API module.

* add BaseSearchConfig

* add _build_search_optional_params

* add search_testing

* add BaseSearchTest

* add TestPerplexitySearch

* fix BASE

* fix handler

* add search API

* add to init

* fix: working perplexity search API

* add _hidden_params to search

* add TAVILY to LlmProviders

* add TavilySearchConfig

* add TavilySearchConfig

* TestTavilySearch

* add tavily transform

* TestParallelAISearch

* add LlmProviders

* add ParallelAISearchConfig

* add ParallelAISearchConfig

* ParallelAISearchConfig

* add EXA AI Search API

* add ExaAISearchConfig

* TestExaAISearch

* add get_supported_perplexity_optional_params

* add Exa AI Search API

* add transform_search_request

* add ExaAISearchConfig

* fix linting errors

* transform_search_request
2025-10-21 17:06:23 -07:00
Ishaan JaffandGitHub 208f76f8ad [Feat] Add Parallel AI - Search API (#15772)
* add BaseSearchConfig

* add BaseSearchConfig

* validate_environment

* fix handlers

* add PerplexitySearchConfig

* add PerplexitySearchConfig

* add LiteLLM Search API module.

* add BaseSearchConfig

* add _build_search_optional_params

* add search_testing

* add BaseSearchTest

* add TestPerplexitySearch

* fix BASE

* fix handler

* add search API

* add to init

* fix: working perplexity search API

* add _hidden_params to search

* add TAVILY to LlmProviders

* add TavilySearchConfig

* add TavilySearchConfig

* TestTavilySearch

* add tavily transform

* TestParallelAISearch

* add LlmProviders

* add ParallelAISearchConfig

* add ParallelAISearchConfig

* ParallelAISearchConfig
2025-10-21 17:00:05 -07:00
Ishaan JaffandGitHub b9f3f9fb79 [Feat] Add Tavily Search API (#15770)
* add BaseSearchConfig

* add BaseSearchConfig

* validate_environment

* fix handlers

* add PerplexitySearchConfig

* add PerplexitySearchConfig

* add LiteLLM Search API module.

* add BaseSearchConfig

* add _build_search_optional_params

* add search_testing

* add BaseSearchTest

* add TestPerplexitySearch

* fix BASE

* fix handler

* add search API

* add to init

* fix: working perplexity search API

* add _hidden_params to search

* add TAVILY to LlmProviders

* add TavilySearchConfig

* add TavilySearchConfig

* TestTavilySearch

* add tavily transform
2025-10-21 16:59:29 -07:00
wenhuaandGitHub b0ccc35a9c fix(ollama): Enhance chunk parsing for empty responses without 'thinking' and improve error logging (#13333) (#15717) 2025-10-21 16:59:01 -07:00
Ishaan JaffandGitHub e1cb92862e [Feat] Add def search() APIs for Web Search - Perplexity API (#15769)
* add BaseSearchConfig

* add BaseSearchConfig

* validate_environment

* fix handlers

* add PerplexitySearchConfig

* add PerplexitySearchConfig

* add LiteLLM Search API module.

* add BaseSearchConfig

* add _build_search_optional_params

* add search_testing

* add BaseSearchTest

* add TestPerplexitySearch

* fix BASE

* fix handler

* add search API

* add to init

* fix: working perplexity search API

* add _hidden_params to search
2025-10-21 16:58:51 -07:00
Ishaan JaffandGitHub 9135e748a0 [Feat ] /ocr - Add mode + Health check support for OCR models (#15767)
* get_mode_handlers

* use get_mode_handlers

* test_ahealth_check_ocr

* Add OCR mode to test models

* docs OCR Health Checks

* fix connection endpoint
2025-10-21 16:58:37 -07:00
Javier GarciaandGitHub b0a3a7c4fb Add details in docs (#15721)
* Add details in docs

* add logic to set span attributes and unit tests

* Restore html files

* Remove html files

* Remove html files
2025-10-21 16:57:51 -07:00
Thomas MildnerandGitHub 1cfc4624c3 [Feat] Add SENTRY_ENVIRONMENT configuration for Sentry integration (#15760)
* [Feat] Add SENTRY_ENVIRONMENT configuration for Sentry integration and corresponding tests

* [Refactor] Enhance test_sentry_environment by mocking sentry_sdk and improving environment handling

* [Fix] Update default SENTRY_ENVIRONMENT to 'production' and enhance test for Sentry integration

* [Fix] Update test_sentry_environment to verify correct handling of SENTRY_ENVIRONMENT values

* [Fix] Update test_sentry_environment to assert correct handling of production environment
2025-10-21 16:40:55 -07:00
KowyoandGitHub 1fcadd6c05 feat(ollama): set 'think' to False when reasoning effort is not high/medium/low (#15763) 2025-10-21 16:39:08 -07:00
Krrish Dholakia 2a1dbb5b9e docs(creating_adapters.md): document how to write an adapter 2025-10-21 16:20:31 -07:00
nuernberandGitHub 353dfb1238 Add AWS us-gov-west-1 Claude 3.7 Sonnet costs (#15775)
* add us-gov-west-1 claude 3.7 sonnet to prices

* add to _backup file as well
2025-10-21 16:17:07 -07:00
YutaSaitoandGitHub 39641e7e68 chore: rename GraySwan to Gray Swan (#15771) 2025-10-21 15:18:55 -07:00
Vinod SinghandGitHub d4aadda692 Auth Header Fix for MCP Tool Call (#15736)
* fixed the Auth header for MCP Tool Call

* Final fix for Auth header

* testcase for mcp_auth_header_extraction, insensitive_alias_matching, insensitive_servername_matching added
2025-10-21 13:58:03 -07:00
Krrish Dholakia 1e0368521e refactor: cleanup 2025-10-21 13:46:19 -07:00
Ishaan Jaffer 3741c43396 docs fix 2025-10-21 13:21:59 -07:00
Ishaan JaffandGitHub 8ad9bbbd02 [Docs] Add Azure AI - OCR to docs (#15768)
* add Azure OCR to docs

* docs fix

* docs fix

* docs fix

* docs OCR
2025-10-21 13:10:45 -07:00
Ishaan Jaffer 185182bebc Revert "add Azure OCR to docs"
This reverts commit a3699e28a4.
2025-10-21 13:02:31 -07:00
Ishaan Jaffer a3699e28a4 add Azure OCR to docs 2025-10-21 13:02:21 -07:00
Ishaan Jaffer 6605aba307 docs grayswan 2025-10-21 11:16:25 -07:00
YutaSaitoandGitHub d79bdd491f feat: add GraySwan Guardrails support (#15756) 2025-10-21 11:13:50 -07:00
46d55bd92a fix: Add response_type + PKCE parameters to OAuth authorization endpoint (#15720)
* fix: Add response_type parameter to OAuth authorization endpoint

Fixes #15684

OAuth providers like Google require the response_type parameter during
the authorization flow. This commit adds response_type=code to the
authorization redirect parameters, which is required by the OAuth 2.0
specification (RFC 6749 Section 4.1.1).

Changes:
- Added response_type=code to authorization params in discoverable_endpoints.py
- Added test coverage for the response_type parameter

🤖 Generated with [Claude Code](https://claude.com/claude-code)

Co-Authored-By: Claude <noreply@anthropic.com>

* fix oauth flow by forwarding code_challenge and forwarding code_verifier

---------

Co-authored-by: Claude <noreply@anthropic.com>
2025-10-21 09:43:19 -07:00
Tom HaynesandGitHub 98f1d63508 use correct otel logger, and normalise otel paths (#15645) v1.78.6-nightly 2025-10-21 09:16:03 -07:00
Ishaan Jaffer 8b522d88a2 is_llm_api_route 2025-10-20 18:05:35 -07:00
Ishaan JaffandGitHub 92335d991c [Feat] Add Azure AVA (Speech AI) Cost Tracking (#15754)
* add azure/speech/ cost tracking

* test_azure_ava_tts_async

* add azure/speech to model cost map

* docs cost tracking

* docs tts AVA

* add azure/speech/azure-tts
2025-10-20 18:01:51 -07:00
Ishaan Jaffer 60fab591db rename test files 2025-10-20 18:00:17 -07:00
Ishaan JaffandGitHub 157739da01 [Bug]: Fix Incorrect status value in responses api with gemini (#15753)
* _map_chat_completion_finish_reason_to_responses_status

* test_transform_chat_completion_response_with_reasoning_content

* test_transform_chat_completion_response_output_item_status
2025-10-20 17:58:56 -07:00
Ishaan Jaffer 5ce2be732e get_provider_text_to_speech_config 2025-10-20 17:10:09 -07:00
Ishaan Jaffer 9a25eeccb2 docs fix 2025-10-20 17:02:38 -07:00
Ishaan Jaffer c9152003bd bump V 2025-10-20 16:55:03 -07:00
Ishaan JaffandGitHub 73a23a6c78 [Feat] Add Azure AVA TTS integration (#15749)
* add AzureBaseIssueTokenHandler

* add BaseTextToSpeechConfig

* async_text_to_speech_handler

* add AzureAVATextToSpeechConfig

* add get_provider_text_to_speech_config

* add AzureAVATextToSpeechConfig

* fixes for base_llm_http_handler

* fix transform_text_to_speech_request

* test_azure_ava_tts_async

* test_azure_ava_tts_async

* fix TextToSpeechRequestData

* fix transform_text_to_speech_request

* add text_to_speech_handler in LLMHttpHandler

* remove old file

* fix transform_text_to_speech_request

* fix dispatch_text_to_speech

* fix azure TTS

* fix AVA TTS

* fix transform

* fix linting

* ci/cd - use one job for audio testing

* fix tests

* fix llm http handler debugging

* unit tests azure tts

* docs Azure speech

* docs fix

* docs azure AVA

* docs azure AVA

* fix handlers

* test_async_realtime_uses_max_size_parameter
2025-10-20 16:52:23 -07:00
akrainesandGitHub 41a6ecd5b6 Change max_tokens value to match max_output_tokens for claude sonnet 4.5: 64000 (#15715)
See https://github.com/RooCodeInc/Roo-Code/issues/8454
2025-10-20 16:11:36 -07:00
Ishaan JaffandGitHub 0c25b1a256 [Fix] OpenAI Realtime API integration fails due to websockets.exceptions.PayloadTooBig error (#15751)
* fix REALTIME_WEBSOCKET_MAX_MESSAGE_SIZE_BYTES

* edit max_size for websockets

* fix AzureOpenAIRealtime
2025-10-20 15:54:14 -07:00
Sameer KankuteandGitHub 1fb798f81d (Bug) Fix JSON serialization error in Helicone logging by removing OpenTelemetry span from metadata (#15728)
* remove span object from helicon metadata

* Add test
2025-10-20 08:53:22 -07:00
Sameer KankuteandGitHub 3955a3de5d fix the wrong request body in json mode doc (#15729) 2025-10-20 08:44:14 -07:00
Timothée LecomteandGitHub 3ef9b2015a feat: read from custom-llm-provider header (#15528) 2025-10-18 22:04:53 -07:00
4a74190c12 Fix: Add gpt 4.1 pricing for response endpoint (#15593)
* Add gpt41, gpt-41-mini, and gpt-41-nano to pricing and context window json

* Add gpt-41s to azure_llms dict

* Undo json changes

---------

Co-authored-by: IQHL (Hans Jacob Landelius) <iqhl@novnordisk.com>
2025-10-18 22:04:14 -07:00
Lucas SugiandGitHub ae86862e74 fix: Add function responsible to call precall (#15636)
* fix: Add function responsible to call precall

* fix: Set correct route_type
2025-10-18 22:01:14 -07:00
Lucas SugiandGitHub ce9e22688d fix: Add pre and post call for list batches (#15673) 2025-10-18 21:52:35 -07:00
f55745fc5e [Fix] Forward anthropic-beta headers to Bedrock, VertexAI (#15700)
* [Fix] Forward anthropic-beta headers to Bedrock and other cross-provider scenarios (#15623)

* add_provider_specific_headers_to_request

* fix add_provider_specific_headers_to_request

* test_provider_specific_header_multi_provider

* test_provider_specific_header_in_request

---------

Co-authored-by: Jack Venberg <jack.venberg@rover.com>
2025-10-18 16:26:32 -07:00
Alexsander HamirandGitHub 441aed2c87 fix: update worker recommendation (#15702) 2025-10-18 16:24:32 -07:00
Ishaan Jaffer 3fc49a029f docs fix 2025-10-18 15:28:49 -07:00
Ishaan Jaffer eed1ddba49 docs v1.78.5-stable 2025-10-18 15:28:07 -07:00
Ishaan JaffandGitHub 6ab9b0af9f [Fix] Anthropic cache_control incorrectly applied to all content items instead of last item only (#15699)
* fix: _safe_insert_cache_control_in_message

* test_anthropic_cache_control_hook_system_message

* docs prompt cache injection

* docs fix
2025-10-18 15:18:08 -07:00
Jason RobertsandGitHub c471bf1f16 feat(guardrails): Add content masking and streaming support to PANW Prisma AIRS guardrail (#15666)
* feat(guardrails): Add content masking and streaming support to PANW Prisma AIRS

- Add mask_request_content and mask_response_content parameters
- Implement content masking for prompts and responses
- Add streaming support with real-time masking
- Add comprehensive test coverage (28 tests)
- Update documentation with masking examples and security notes

* fix(guardrails): Fix PANW Prisma AIRS env var fallback and text completion support
v1.78.5-stable v1.78.5.rc.1 v1.78.5-nightly
2025-10-18 13:57:51 -07:00
YutaSaitoandGitHub 645f84c02e fix: add imagePullSecrets to migrations-job (#15681) 2025-10-18 13:56:31 -07:00
katsuhiro mutoandGitHub d5e686b3e8 [Fix] Support service_tier in chat completion (#15693)
* Support service_tier

* fix test
2025-10-18 13:55:54 -07:00
Krish DholakiaandGitHub c1355e92dc fix(proxy_server.py): re-encrypt env var on config save + use original value on decrypt error (#15671)
* fix(proxy_server.py): re-encrypt env var on config save + use original value on decrypt error

Closes https://github.com/BerriAI/litellm/issues/14854

Fixes https://github.com/BerriAI/litellm/issues/13406

* docs: email.md

document PROXY_BASE_URL param

* fix(proxy_server.py): pop model list before writing to db
2025-10-18 13:39:25 -07:00