Commit Graph
30990 Commits
Author SHA1 Message Date
yuneng-jiang 937ccf1977 UI Keys Teams Router Settings docs 2026-01-24 16:23:46 -08:00
yuneng-jiang 47810f1523 Model and Team filtering 2026-01-24 14:45:14 -08:00
Ishaan Jaffer a7e26460d0 fix unstable tests 2026-01-24 11:15:16 -08:00
Ishaan Jaffer bd38374a45 fix: FLAKY tests 2026-01-24 11:13:44 -08:00
Alexsander HamirandGitHub 9cdd7a8fd2 Fix: log duplication when json_logs is enabled (#19705) 2026-01-24 11:09:04 -08:00
Ishaan Jaffer 6710abd1fa test_web_search 2026-01-24 11:06:23 -08:00
Ishaan Jaffer 6587cd228b test_partner_models_httpx_streaming 2026-01-24 10:58:35 -08:00
Ishaan Jaffer 1a7274aa4e fix: _apply_search_filter_to_models mypy linting 2026-01-24 09:24:50 -08:00
Ishaan Jaffer 489d587332 CI/CD fixes - split local testing 2026-01-24 09:23:03 -08:00
yuneng-jiangandGitHub 387cb8fde8 Merge pull request #19703 from BerriAI/yj_ui_build_jan24
[Infra] Build UI for Release
2026-01-24 09:20:26 -08:00
yuneng-jiang 17aec96186 chore: update Next.js build artifacts (2026-01-24 17:18 UTC, node v22.16.0) 2026-01-24 09:18:58 -08:00
Ishaan Jaffer 006b810370 bump: version 1.81.2 → 1.81.3 2026-01-24 09:15:06 -08:00
yuneng-jiangandGitHub ed1b9529f7 Merge pull request #19601 from BerriAI/litellm_ui_update_org_model
[Feature] UI - Organization Edit Page: Reusable Model Select
2026-01-24 09:10:15 -08:00
yuneng-jiangandGitHub 5e395db1dc Merge pull request #19604 from BerriAI/litellm_team_update_org
[Fix] Team Update with Organization having All Proxy Models
2026-01-24 09:09:55 -08:00
yuneng-jiangandGitHub f88a32de05 Merge pull request #19622 from BerriAI/litellm_ui_model_backend
[Feature] UI - Models Page: Model Search
2026-01-24 09:09:03 -08:00
yuneng-jiang b44ac6c682 Fixing ruff check 2026-01-24 09:08:29 -08:00
Ishaan JaffandGitHub 0bdb68dea7 Update OSS Adopters section with new table format 2026-01-24 09:07:03 -08:00
yuneng-jiangandGitHub 09c7d67539 Merge pull request #19701 from BerriAI/litellm_ci_fix_yj_03
[Infra] Fixing CircleCI Config
2026-01-24 08:55:50 -08:00
yuneng-jiang 1c9731527f fixing circleci config 2026-01-24 08:54:22 -08:00
yuneng-jiangandGitHub cec367065e Merge pull request #19695 from BerriAI/litellm_ci_fix_yj_03
[Infra] CI/CD - Fixing failing tests
2026-01-23 23:31:31 -08:00
yuneng-jiang acd8f21f68 fixing circleci config 2026-01-23 23:30:54 -08:00
yuneng-jiang 00ec939ea6 cache tests serial 2026-01-23 23:13:45 -08:00
yuneng-jiang 63166c3acc fixing arize tests 2026-01-23 23:13:21 -08:00
yuneng-jiang e1bb4ae280 deactivating non root tests 2026-01-23 22:55:36 -08:00
yuneng-jiang 86676142c9 Fixing failing tests 2026-01-23 22:33:00 -08:00
yuneng-jiangandGitHub e181f2daa4 Merge pull request #19694 from BerriAI/litellm_ci_fix_yj_02
[Infra] CI/CD - Fixing UI Build
2026-01-23 22:12:38 -08:00
yuneng-jiang 1dbb6e0d3f fixing build 2026-01-23 22:11:49 -08:00
Cesar GarciaandGitHub 31a8d76d11 Update Gemini 2.0 Flash deprecation dates to March 31, 2026 (#19592)
Google announced that Gemini 2.0 Flash and Flash Lite models will be discontinued on March 31, 2026. Updated deprecation_date field for all affected model variants across different providers (vertex_ai, gemini, deepinfra, openrouter, vercel_ai_gateway).

Models updated:
- gemini-2.0-flash (added deprecation date)
- gemini-2.0-flash-001 (updated from 2026-02-05)
- gemini-2.0-flash-lite (added deprecation date)
- gemini-2.0-flash-lite-001 (updated from 2026-02-25)

All variants now correctly reflect the March 31, 2026 shutdown date.
2026-01-23 20:36:36 -08:00
Harshit JainandGitHub f4ba5b9209 docs: add litellm-enterprise requirement for managed files (#19689) 2026-01-23 19:51:39 -08:00
Ishaan JaffandGitHub a870722f65 [Feat] UI + Backend - Allow adding policies on Keys/Teams + Viewing on Info panels (#19688)
* ui for policy mgmt

* test_add_guardrails_from_policy_engine_accepts_dynamic_policies_and_pops_from_data
2026-01-23 19:03:44 -08:00
yuneng-jiangandGitHub 4ed5aa5de0 Merge pull request #19687 from BerriAI/litellm_ui_refresh_mcp
[Fix] UI - Redirect to ui/login on expired JWT
2026-01-23 18:09:44 -08:00
yuneng-jiang a34664d8b0 redirect to login on expired jwt 2026-01-23 18:03:10 -08:00
Ishaan Jaffer 46ef001150 UI: new build 2026-01-23 17:40:54 -08:00
yuneng-jiangandGitHub 26a6b86ea9 Merge pull request #19265 from naaa760/fix/guar-patt-edi
fix: ensure guardrail patterns persist on edit and mode toggle
2026-01-23 17:36:09 -08:00
ryan-crabbeandGitHub d67d12fc54 perf: Add LRU caching to get_model_info for faster cost lookups (#19606)
- Add @lru_cache decorator to get_model_info() and _cached_get_model_info_helper()
- Update _invalidate_model_cost_lowercase_map() to clear these caches when model_cost changes
- Update test to call cache invalidation after modifying litellm.model_cost

Reduces get_model_cost_information from 46% to <1% of request handling time.
2026-01-23 17:26:45 -08:00
yuneng-jiangandGitHub b5dfb57073 Merge pull request #19686 from BerriAI/litellm_key_team_create_routing_setting_ui
[Feature] UI - Create Team and Key Router Settings
2026-01-23 17:26:13 -08:00
ryan-crabbeandGitHub 6e930c9724 perf: skip pattern_router.route() for non-wildcard models (#19664)
Check "*" in model before calling pattern_router.route() to avoid
unnecessary pattern matching for non-wildcard model configurations.
2026-01-23 17:21:41 -08:00
ryan-crabbeandGitHub 54f9ad370f perf: Optimize use_custom_pricing_for_model with set intersection (#19677)
* perf: Optimize use_custom_pricing_for_model with set intersection

Cache CustomPricingLiteLLMParams.model_fields.keys() as a module-level
frozenset and use set intersection to reduce loop iterations from 882k
to 90k (only iterating over keys that exist in both sets).

Performance improvement: 84% faster (6.3x speedup)
- Before: 1.17s total, 65µs per call
- After: 0.19s total, 10µs per call

* Use .get() for defensive dictionary access
2026-01-23 17:18:16 -08:00
yuneng-jiang de9802578b Fixing tests 2026-01-23 17:15:50 -08:00
ryan-crabbeandGitHub 0133d50a45 perf: Optimize strip_trailing_slash with O(1) index check (#19679)
* perf: Optimize strip_trailing_slash with O(1) index check

Replace rstrip("/") with direct index check for O(1) performance
instead of O(n) string scanning.

Results:
- strip_trailing_slash: 311ms → 13ms (96% faster)
- get_standard_logging_object_payload: 6.11s → 5.80s (5% faster)

* Handle multiple trailing slashes in strip_trailing_slash

Use rstrip for correctness when URL ends with "//" or more,
otherwise use O(1) index check for single trailing slash.
2026-01-23 17:12:08 -08:00
yuneng-jiang f9bdc20be2 fixing tests 2026-01-23 17:07:51 -08:00
yuneng-jiang ee1fd1c6c2 fixing build 2026-01-23 17:04:52 -08:00
yuneng-jiang 804567d681 Merge remote-tracking branch 'origin' into litellm_key_team_create_routing_setting_ui 2026-01-23 16:56:30 -08:00
yuneng-jiang 9850dbe934 Adding router settings to create team and key 2026-01-23 16:56:19 -08:00
Alexsander HamirandGitHub 5c61586e65 Add GCS mock mode for testing without API calls (#19683) 2026-01-23 16:25:32 -08:00
Alexsander HamirandGitHub 56883add3c Add Langfuse mock mode for testing without API calls (#19676) 2026-01-23 15:33:40 -08:00
yuneng-jiangandGitHub 22a268c544 Merge pull request #19673 from BerriAI/litellm_ui_router_fallbacks_02
[Feature] UI - Fallbacks: New Add Fallbacks Modal
2026-01-23 14:33:30 -08:00
yuneng-jiang 3ba9b13390 adding tests 2026-01-23 14:28:47 -08:00
yuneng-jiang 7e6fc6af2c New add fallbacks modal 2026-01-23 14:21:32 -08:00
Ishaan Jaffer 5b341ee842 fix linting 2026-01-23 13:46:02 -08:00