Sameer Kankute and GitHub
b2feedc469
Merge pull request #20318 from BerriAI/litellm_oss_staging_02_03_2026
...
feat(guardrails): implement team-based isolation guardrails mgmnt (#1…
2026-02-04 17:49:30 +05:30
Sameer Kankute and GitHub
bfc0c4b4ab
Merge pull request #20396 from BerriAI/litellm_copilot_kit_sdk
...
[DOCS]Add copilotkit sdk doc as supported agents sdk
2026-02-04 13:41:08 +05:30
Sameer Kankute
8109413a54
Add copilotkit sdk doc as supported agents sdk
2026-02-04 11:52:17 +05:30
Ishaan Jaff and GitHub
da4cf4942f
[Feat] Add xAI /realtime API Support - works with LiveKitSDK ( #20381 )
...
* init: _realtime_health_check + routing
* refactor: OpenAIRealtime
* refactor: XAI_API_BASE
* feat: XAIRealtime
* init feat: XAIRealtime
* OpenAIRealtime
* TestXAIRealtime
* test fixes
* test OAI
* TEST xAI, OAI
* clean realtime jobs
* refactor
* test XAI
* docs xAI
* fix xAI
* fix lint errors
* test_async_realtime_url_contains_model
* test fix
* document test changes
* _realtime_health_check
* docs xai realtime
* fix handlers
* add additional_headers
* fix
2026-02-03 19:58:28 -08:00
Krish Dholakia and GitHub
7056d9984e
Custom Code Guardrails UI Playground ( #20377 )
...
* feat(guardrails/): allow custom code execution for guardrails
first step in allowing teams to submit custom code for guardrails
* feat: custom_code_guardrail.md
support passing custom code for guardrails
* feat: initial commit adding ui for custom code guardrails
allows users to write guardrails based on custom code
* feat: expose new test custom code guardrail endpoint
allows ui testing playground to sanity check if guardrail is working as expected
* fix: fix linting errors
* fix: fix max recursion check
* fix: fix linting error
2026-02-03 19:57:24 -08:00
Ishaan Jaff and GitHub
d267c69086
[Feat] Use A2A registered agents with /chat/completions ( #20362 )
...
* test_a2a_registry_integration
* fix: render agents on model dropdown on UI
* init append_agents_to_model_group
* route_a2a_agent_request
* is_a2a_agent_model
* route_a2a_agent_request
* fix: error handling
* docs A2A usage
* docs fix
* feat: working A2a streaming
* fix transform
2026-02-03 15:25:38 -08:00
Sameer Kankute
333419b4d2
Add documentation correctly for nova sonic
2026-02-03 09:03:27 +05:30
0ef506a54a
Litellm docs mcp filtering semantic ( #20316 )
...
* init: SemanticMCPToolFilter
* init: SemanticToolFilterHook
* test_e2e_semantic_filter
* mock tests: test_semantic_filter_basic_filtering
* Update litellm/proxy/_experimental/mcp_server/semantic_tool_filter.py
Co-authored-by: greptile-apps[bot] <165735046+greptile-apps[bot]@users.noreply.github.com>
* refactor folder/file organization
* docs fix
* fix filter
* fix: filter_tools
* fix linting tool filrer
* initialize_from_config
* fix: _expand_mcp_tools
* _initialize_semantic_tool_filter
* working: async_post_call_response_headers_hook
* clean up semantic tool filter
* add _initialize_semantic_tool_filter
* build_router_from_mcp_registry
* _get_tools_by_names
* fiix config
* async_post_call_response_headers_hook
* docs mcp filter
* docs fix
---------
Co-authored-by: greptile-apps[bot] <165735046+greptile-apps[bot]@users.noreply.github.com>
2026-02-02 18:29:07 -08:00
shin-bot-litellm and GitHub
0614ff9fda
docs: add Prisma migration troubleshooting guide ( #20300 )
...
* docs: add Prisma migration troubleshooting guide
Add troubleshooting documentation for common Prisma migration errors
encountered when upgrading/downgrading LiteLLM proxy versions.
Covers:
- 'relation does not exist' errors after version rollback
- Blocked migrations from previous failures
- Migration state mismatch after version rollback
- General tips for prisma migrate resolve, db push, and migrate deploy
* docs: simplify prisma migration troubleshooting - focus on delete + restart
2026-02-02 14:39:18 -08:00
73691fb373
Model request tags documentation ( #20290 )
...
* Add request tags documentation for spend tracking
- Add new concise doc explaining how to tag model requests
- Include Python SDK and cURL examples
- Show where tags appear in spend logs
- Add common use cases table (AWS accounts, teams, projects)
- Include how to set default tags on API keys
- Add to Spend Tracking section in sidebar
Co-authored-by: ishaan <ishaan@berri.ai >
* Simplify request tags doc for AI Gateway usage
- Focus on config.yaml setup with default_key_generate_params
- Show both request body and header methods for sending tags
- Remove SDK examples, keep concise cURL examples
- Streamline for quick reference
Co-authored-by: ishaan <ishaan@berri.ai >
* Update request tags doc to show model-level config
- Set tags directly on model deployments in litellm_params
- Requests just specify model, tags applied automatically
- Use clear naming: AWS_IAM_PROD, AWS_IAM_DEV
Co-authored-by: ishaan <ishaan@berri.ai >
---------
Co-authored-by: Cursor Agent <cursoragent@cursor.com >
Co-authored-by: ishaan <ishaan@berri.ai >
2026-02-02 11:32:00 -08:00
yuneng-jiang
af015fe4f0
UI spend logs setting docs
2026-01-31 15:16:59 -08:00
Ishaan Jaff and GitHub
c9658f877e
[Docs] Claude Agents SDK x LiteLLM Guide ( #20036 )
...
* docs claude agent SDK
* docs fix
* docs
* docs
2026-01-29 18:04:54 -08:00
Ishaan Jaff and GitHub
3080e04180
[Feat] UI: Allow Admins to control what pages are visible on LeftNav ( #19907 )
...
* feat: enabled_ui_pages_internal_users
* init ui for internal user controsl
* fix ui settings
* fix build
* fix leftnav
* fix leftnav
* test fixes
* fix leftnav
* isPageAccessibleToInternalUsers
* docs fix
* docs ui viz
2026-01-27 19:31:24 -08:00
Harshit Jain and GitHub
0f0b71e6d9
feat: add feature to make silent calls ( #19544 )
...
* feat: add feature to make silent calls
* add test or silent feat
* add docs for silent feat
* fix lint issues and UI logs
* add docs of ab testing and deep copy
2026-01-27 09:16:53 -08:00
Sameer Kankute
faf9c9ba76
Add sarvam doc
2026-01-27 13:11:29 +05:30
yuneng-jiang and GitHub
8094aff8c5
Merge pull request #19715 from BerriAI/key_teams_fallback_docs
...
[Docs] UI Keys Teams Router Settings docs
2026-01-24 16:26:28 -08:00
yuneng-jiang
937ccf1977
UI Keys Teams Router Settings docs
2026-01-24 16:23:46 -08:00
Ishaan Jaffer
bbeb007f4e
docs fix
2026-01-24 12:09:16 -08:00
milan-berri and GitHub
37b7dff194
add spend-queue-troubleshooting docs ( #19659 )
...
* add spend-queue-troubleshooting docs
* adjust spend-queue-troubleshooting docs
2026-01-23 13:44:41 -08:00
Ishaan Jaff and GitHub
c23e4b87dc
[Feat] New LiteLLM Policy engine - create policies to manage guardrails, conditions - permissions per Key, Team ( #19612 )
...
* init PolicyMatcher
* TestPolicyMatcherGetMatchingPolicies
* TestPolicyMatcherGetMatchingPolicies
* feat: init PolicyResolver
* init resolver types
* init policy from config
* inint PolicyValidator
* validate policy
* init Architecture Diagram
* test_add_guardrails_from_policy_engine
* init _init_policy_engine
* test updates
* test fixws
* new attachment config
* simplify types
* TestPolicyResolverInheritance
* fix policy resolver
* fix policies
* fix applied policy
* docs fix
* docs fix
* fix linting + QA checks
* fix linting + QA fixes
* test fixes
2026-01-22 19:49:53 -08:00
milan-berri and GitHub
c9516d68d4
add opencode tutorial ( #19602 )
2026-01-22 15:29:19 -08:00
Sameer Kankute and GitHub
ad1edd38d5
Merge branch 'main' into litellm_staging_01_21_2026
2026-01-22 17:56:40 +05:30
Ishaan Jaff and GitHub
ab606c9a73
[Feat] Add Structured output for /v1/messages with Anthropic API, Azure Anthropic API, Bedrock Converse ( #19545 )
...
* fix: add AnthropicMessagesRequestOptionalParams
* add _update_headers_with_anthropic_beta
* fix output format tests
* test_structured_output_e2e
* TestAnthropicAPIStructuredOutput
* test_structured_output_e2e
* fix BASE
* TestAzureAnthropicStructuredOutput
* fix: Bedrock Converse
* add nthropic Messages Pass-Through Architecture
* fix: bedrock invoke output_format
* fix: transform_anthropic_messages_request for vertex anthropic
* TestBedrockInvokeStructuredOutput
* docs anthropic vertex
* docs fix
* docs fix
2026-01-21 20:09:18 -08:00
4106d24215
feat: add GMI Cloud provider support ( #19376 )
...
* feat: add GMI Cloud provider support
Add GMI Cloud as an OpenAI-compatible provider with:
- Provider configuration in providers.json
- Documentation page with usage examples
- Model pricing for 16 models (Claude, GPT, DeepSeek, Gemini, etc.)
- Sidebar entry for docs navigation
* Add gmi_cloud to provider_endpoints_support.json
Add provider entry to pass CI validation check that ensures all
providers in openai_like/providers.json are documented.
* Fix provider key: gmi_cloud -> gmi
Match the provider key with providers.json
---------
Co-authored-by: Krish Dholakia <krrishdholakia@gmail.com >
2026-01-21 15:48:15 -08:00
Ishaan Jaff and GitHub
a02c43d300
Litellm cc docs max ( #19466 )
...
* docs claude code max
* docs fix
* docs
* docs fix
* docs fix
2026-01-20 20:08:36 -08:00
Sampson and GitHub
09941dd1d1
add search provider for brave search api ( #19433 )
...
* add search provider for brave search api
Introduces a minimal implementation of the Brave Search API as a search provider. Additionally, this PR introduces a test file to ensure the provider works properly, and numerous other smaller changes (e.g., changes to docs to mention the new option).
* Update transformation.py
2026-01-20 19:23:29 -08:00
Sameer Kankute and GitHub
deb9142117
Merge pull request #19400 from BerriAI/main
...
merge main iin 19/1 staging
2026-01-20 16:45:01 +05:30
Ishaan Jaffer
f865f92bec
docs plugin marketplaces
2026-01-19 19:42:15 -08:00
Manuel Schweigert and GitHub
29adf34313
Add ChatGPT subscription support and responses bridge ( #19030 )
...
* Add ChatGPT subscription support and responses bridge
* Fix typing import for responses bridge
* Guard device code timestamp parsing
* add /v1/messages endpoint to chatgpt model
2026-01-19 05:37:45 -08:00
Ishaan Jaff and GitHub
1417b002a3
[Feat] Claude Code x LiteLLM WebSearch - QA Fixes to work with Claude Code ( #19294 )
...
* fix websearch_interception_converted_stream
* test_websearch_interception_no_tool_call_streaming
* FakeAnthropicMessagesStreamIterator
* LITELLM_WEB_SEARCH_TOOL_NAME
* fixes tools def for litellm web search
* fixes FakeAnthropicMessagesStreamIterator
* test_litellm_standard_websearch_tool
* use new hook for modfying before any transfroms from litellm
* init WebSearchInterceptionLogger + ARCHITECTURE
* fix config.yaml
* init doc for claude code web search
* docs fix
* doc fix
* fix mypy linting
2026-01-17 16:30:31 -08:00
yuneng-jiang
19a69a89f0
deleted keys and teams docs
2026-01-17 13:19:48 -08:00
Ishaan Jaffer
eb26ebc926
docs ui usage
2026-01-17 12:15:37 -08:00
YutaSaito and GitHub
66d67ae356
Revert "Add sanititzation for anthropic messages"
2026-01-17 06:01:12 +09:00
Sameer Kankute and GitHub
fb3b4e6b33
Merge pull request #19196 from BerriAI/litellm_sanitise_anthropic_mesages
...
Add sanititzation for anthropic messages
2026-01-16 17:48:55 +05:30
Sameer Kankute
8be3712e82
Add docs for message sanitisation
2026-01-16 12:52:13 +05:30
Sameer Kankute
d585b760c9
Add fallback endpoints support
2026-01-16 10:51:33 +05:30
Ishaan Jaff and GitHub
b5b9c39beb
[Docs Guide] Litellm claude code end user tracking ( #19176 )
...
* add to sidebar
* v1 guide
* guide claude granular cost tracking
* docs fix
2026-01-15 18:32:58 -08:00
Krish Dholakia and GitHub
664ee27ef5
Litellm dev 01 15 2026 p1 ( #19153 )
...
* fix: safely handle unmapped call type
* docs: cleanup links for ai coding tools
* docs(claude_non_anthropic_models.md): add tutorial showing non anthropic model connection to claude code
* docs: link to non-anthropic model tutorial for claude code
2026-01-16 00:48:41 +05:30
Krrish Dholakia
c7ca2dd4d8
docs(claude_mcp.md): separate claude mcp tutorial into a separate doc
...
easier to surface
2026-01-15 20:52:31 +05:30
YutaSaito and GitHub
e3e6fc2806
Merge pull request #19122 from BerriAI/litellm_docs_mcp_troubleshooting
...
[doc] add MCP troubleshooting guide
2026-01-15 09:34:35 +09:00
Alexsander Hamir and GitHub
f442b57848
docs: Add structured issue reporting guides for CPU and memory issues ( #19117 )
2026-01-14 15:28:20 -08:00
Yuta Saito
0b04145c11
doc: add MCP troubleshooting guide
2026-01-15 08:01:59 +09:00
Ishaan Jaff and GitHub
1b00576711
[Feat] New Model - Azure Model Router on LiteLLM AI Gateway ( #19054 )
...
* fix - azure model router integration
* fix:_check_provider_match
* fix:_get_response_model
* tests azure model router
* test_azure_ai_model_router_streaming_model_in_chunk
* fix LlmProviders.AZURE.value
* test_azure_ai_model_router_streaming_cost_with_stream_options
* def test_get_model_from_chunks_azure_model_router():
* _get_model_from_chunks
* docs azure model router
* azure model router
2026-01-13 18:31:43 -08:00
Sameer Kankute and GitHub
844c766c65
Merge pull request #18763 from BerriAI/litellm_staging_01_07_2026
...
Staging - 01/07/2026
2026-01-09 17:01:58 +05:30
Ishaan Jaff and GitHub
cbac70a4ec
MANUS docs ( #18817 )
2026-01-08 18:58:10 +05:30
dc4ce7c5a2
feat: Add abliteration.ai provider ( #18678 )
...
* feat: Add abliteration.ai provider
* adding signoz integration to observability docs
* Fixing build
* Adding timeout for flaky test
* Fixing e2e
* add team member budget duration in team/update
* Reusable Duration Select and update team member budget UI
---------
Co-authored-by: Goutham Karthi <goutham@signoz.io >
Co-authored-by: yuneng-jiang <yuneng.jiang@gmail.com >
Co-authored-by: YutaSaito <36355491+uc4w6c@users.noreply.github.com >
2026-01-07 21:46:54 +05:30
Krish Dholakia and GitHub
80ead21c3a
Litellm improve endpoint discovery ( #18762 )
...
* docs: document all endpoints in .json and add consistency checks against docs + providers.json
* docs: add more tests + improve coverage
2026-01-07 17:35:01 +05:30
drorIvry and GitHub
000913fa12
Hotfix - docs qualifire ( #18724 )
...
* Hotfix - docs qualifire
* Hotfix - docs qualifire
* Hotfix - docs qualifire
* Hotfix - docs qualifire
* Hotfix - docs qualifire
* Hotfix - docs qualifire
* Hotfix - docs qualifire
2026-01-07 17:23:12 +05:30
Lundin Matthews and GitHub
762345172c
Add LlamaGate as a new provider ( #18673 )
...
Adds LlamaGate (https://llamagate.dev ) as an OpenAI-compatible provider with:
- Provider configuration in providers.json
- Documentation page with usage examples
- Model pricing for 17 models across categories:
- General purpose (Llama 3.1/3.2, Mistral, Qwen, Dolphin)
- Reasoning (DeepSeek R1, OpenThinker)
- Code (Qwen Coder, DeepSeek Coder, CodeLlama)
- Vision (Qwen VL, LLaVA, Gemma 3)
- Embeddings (Nomic, Qwen3 Embedding)
Provider details:
- Base URL: https://api.llamagate.dev/v1
- Auth: Bearer token via LLAMAGATE_API_KEY
- Pricing: $0.02-$0.55 per 1M tokens
- All models are open-weights
2026-01-07 00:00:30 +05:30
Ishaan Jaff and GitHub
76eda472be
[Feat] New API Endpoint - Responses API (v1/responses/compact) ( #18697 )
...
* init transform_compact_response_api_request
* init acompact_responses
* init async_compact_response_api_handler in llm http handler
* init transform_compact_response_api_request for openai
* init acompact_responses
* fix acompact_responses
* add OAI Compact API
* docs responses API Compact
* code qa checks
* test_openai_compact_responses_api
* fix mypy linting
2026-01-06 16:24:04 +05:30