litellm

mirror of https://github.com/tiennm99/litellm.git synced 2026-06-18 17:28:19 +00:00

Author	SHA1	Message	Date
Ryan Crabbe	75bc8329e2	Merge origin/main into litellm_perf_skip_usage_roundtrip Resolve conflict in litellm_logging.py: take main's version and re-apply get_usage_as_dict optimization on top.	2026-02-21 12:55:55 -08:00
Ishaan Jaff	a5e886de79	fix(tests): read CI_CD_DEFAULT_ANTHROPIC_MODEL env var instead of hardcoding model (#21781 ) * fix(tests): read CI_CD_DEFAULT_ANTHROPIC_MODEL env var in bedrock KB tests * fix(tests): read CI_CD_DEFAULT_ANTHROPIC_MODEL env var in test_router * fix(tests): read CI_CD_DEFAULT_ANTHROPIC_MODEL env var in test_router_retries * fix(tests): read CI_CD_DEFAULT_ANTHROPIC_MODEL env var in test_router_timeout	2026-02-21 10:46:49 -08:00
Ishaan Jaff	0726bdb67c	fix(tests): update gcs pubsub v1 fixture with new SpendLogsMetadata fields (#21779 ) SpendLogsMetadata added new fields (user_api_key, status, error_information, etc.) that weren't in the expected spend_logs_payload.json fixture, causing test_async_gcs_pub_sub_v1 to fail.	2026-02-21 10:40:26 -08:00
Emerson Gomes	cba3bcf1a9	fix(logging): avoid shared callback list references (#20984 )	2026-02-13 18:32:41 +05:30
Sameer Kankute	59d6ab8a00	Merge branch 'main' into litellm_oss_staging_02_11_2026	2026-02-12 20:04:46 +05:30
Ryan Crabbe	7dda6b0cbb	perf: skip Usage Pydantic round-trip in logging payload (6.8x faster)	2026-02-11 12:39:45 -08:00
Sameer Kankute	375ebb333e	Fix: phoenix tests issues	2026-02-11 16:13:20 +05:30
shin-bot-litellm	c7a22921dc	feat: add standard_logging_payload_excluded_fields config option (#20831 ) Adds a new config option to exclude specific fields from StandardLoggingPayload before any callback receives it. This provides a general approach to control what data is logged across ALL integrations (S3, GCS, Datadog, etc.). ## Changes 1. litellm/__init__.py: Added new global setting `standard_logging_payload_excluded_fields: Optional[List[str]] = None` 2. litellm/integrations/custom_logger.py: Modified `redact_standard_logging_payload_from_model_call_details()` to: - Remove specified fields entirely from the StandardLoggingPayload - Works alongside existing `turn_off_message_logging` feature - Excluded fields take precedence (removed rather than redacted) 3. tests/: Added comprehensive test suite with 17 tests covering: - Single/multiple field exclusion - Interaction with turn_off_message_logging - Original payload immutability - Config loading via setattr (proxy pattern) - Edge cases (empty list, non-existent fields, None standard_logging_object) ## Usage ```yaml litellm_settings: success_callback: ["s3"] standard_logging_payload_excluded_fields: ["response", "messages"] ``` This removes the `response` and `messages` fields from logs before any callback processes them, reducing log size and improving privacy compliance. ## Available Fields The fields match StandardLoggingPayload TypedDict keys including: - messages, response (large payload fields) - metadata, hidden_params, model_parameters - error_str, error_information - And all other StandardLoggingPayload fields Closes the need for per-integration flags like `s3_log_response`.	2026-02-10 22:16:41 -08:00
Shivam Rawat	dd5c14baf8	posthog serilization fix (#20668 )	2026-02-07 14:24:32 -08:00
ryan-crabbe	8c7051686b	perf: optimize get_standard_logging_metadata with set intersection (#19685 ) * perf: Optimize get_standard_logging_metadata with set intersection - Cache StandardLoggingMetadata.__annotations__.keys() as module-level frozenset - Use set intersection to iterate only keys present in both metadata and supported keys - Single lookup for user_api_key instead of 3 separate .get() calls Results: - get_standard_logging_metadata: 1.55s → 1.41s (9.2% faster) * test: add unit tests for get_standard_logging_metadata non-string user_api_key handling	2026-02-07 09:35:03 -08:00
Harshit Jain	13130ea3e1	Litellm fix langfuse otel trace (#20382 ) * fix: support multi-project keys and fix trace leakage * fix: Langfuse otel handle * fix lint errors mypy * passing all test case --------- Co-authored-by: Krish Dholakia <krrishdholakia@gmail.com>	2026-02-03 22:40:19 -08:00
mubashir1osmani	c41963c949	fix: add openinference span kinds to arize phoenix fix: add openinference span kinds to arize phoenix	2026-01-23 16:32:49 -05:00
Yuta Saito	898cc3ff4f	test: update langfuse trace_id tests to use litellm_trace_id	2026-01-22 06:19:43 +09:00
mubashir1osmani	410daf6e6d	added tests	2026-01-21 14:41:35 -05:00
Sameer Kankute	eb49adb201	Add user auth in standard logging object for bedrock passthrough	2026-01-15 18:36:06 +05:30
Sameer Kankute	7dbf09cb12	Fix all 130126 tests	2026-01-14 17:47:03 +05:30
Sameer Kankute	de6330b6b6	Fix test_async_otel_callback[False]	2026-01-13 16:59:17 +05:30
Dima-Mediator	7c61933bc5	Fix image tokens spend logging for /images/generations	2026-01-12 23:07:08 -05:00
Ishaan Jaffer	38ccfb4234	test_completion_claude_3_function_call_with_otel	2026-01-07 14:22:54 +05:30
Ishaan Jaffer	b7798b7f7c	fix metadata.cost_breakdown	2026-01-07 14:17:08 +05:30
Sameer Kankute	3c0248edb9	Merge pull request #18663 from BerriAI/litellm_staging_01_05_2026 Staging 01/05/2026	2026-01-06 10:46:04 +05:30
YutaSaito	ccdcb20048	Merge pull request #18279 from mangabits/fix-otel-provider Use already configured opentelemetry providers	2026-01-06 12:49:48 +09:00
Shivam Rawat	8c21fcb957	added the option of adding langsmith tenant id in the env (#18623 )	2026-01-06 01:19:27 +05:30
Chetan Choudhary	687adc6024	Add log_format parameter to GenericAPILogger (#18587 ) Adds log_format parameter supporting json_array (default), ndjson, and single formats. NDJSON format enables webhook integrations like Sumo Logic to parse individual log records at ingest time. Defaults to json_array for backward compatibility.	2026-01-02 23:28:30 +05:30
Yuta Saito	3ae38c2a6d	fix: test_langfuse_logging_completion_with_bedrock_llm_response	2026-01-02 15:55:02 +09:00
mangabits	78693bb9d0	Use already configured opentelemetry providers Users that instrument using opentelemetry-instrument can now setup exporters as per their environment.	2026-01-01 19:06:55 -08:00
Yuta Saito	b343d15157	fix: prevent LiteLLM from closing external OTEL spans	2026-01-01 08:28:48 +09:00
Alexsander Hamir	5534038e93	Fix CI: Revert security scan changes and add GitGuardian ignore rules (#18358 )	2025-12-22 17:03:53 -08:00
Yuta Saito	41bbb3a6a5	feat: datadog log trace linking	2025-12-22 06:53:44 +09:00
Yuta Saito	01aa082d16	fix: call datadog_handler	2025-12-22 05:33:39 +09:00
Ishaan Jaffer	6112160a16	Revert "[Fix] Security - Remove example API keys with high entropy (#18255 )" This reverts commit `24edbccf5c`.	2025-12-20 20:48:11 +05:30
Alexsander Hamir	24edbccf5c	[Fix] Security - Remove example API keys with high entropy (#18255 )	2025-12-19 10:09:50 -08:00
Alexsander Hamir	2e7b554747	3[Fix] CI/CD - logging_testing (#18204 ) * fix: enforce team member budget check in common_checks - Add missing team member budget validation in common_checks() function - Checks team membership budget when team key is used - Raises BudgetExceededError when team member spend exceeds max_budget_in_team - Follows same pattern as other budget checks (team, user, end_user) - Uses cached get_team_membership() for performance - Fix AttributeError in lowest_tpm_rpm.py - Add null check for model_info before accessing .get() method - Prevents 'NoneType' object has no attribute 'get' error - Add unit tests for team member budget enforcement - Test budget exceeded scenario - Test within budget scenario - Test edge cases (no budget, no membership, personal keys) - Tests run without requiring proxy server Fixes failing test: test_users_in_team_budget * fix: mock get_async_httpx_client in test_langsmith_key_based_logging - Mock get_async_httpx_client to return a mock AsyncHTTPHandler instance - Fixes test failure where mock_post was never called - LangsmithLogger creates its own httpx client instance via get_async_httpx_client, so we need to mock the factory function rather than the class method - Use MagicMock for response.raise_for_status (sync method) instead of AsyncMock * fix: resolve linting errors (PLR0915, F401) - Remove unused imports (datetime, ServiceLoggerPayload) from arize_phoenix.py - Extract health ping setup logic from RedisCache.__init__ to reduce statement count - Extract team member budget check from common_checks to reduce statement count * fix: resolve type errors in ChatCompletionToolCallChunk construction - Cast type field to Literal['function'] to satisfy TypedDict requirements - Ensure arguments field is explicitly str type to match TypedDict signature - Fixes pyright errors for incompatible types in transformation.py	2025-12-18 10:52:24 -08:00
yuneng-jiang	c2f79681b6	Fixing test 2	2025-12-16 13:44:53 -08:00
Alexsander Hamir	fab1b81b7f	fix: add agent_id field to GCS PubSub spend_logs_payload.json test expectation (#17938 ) - Add agent_id: null to expected JSON to match actual payload structure - Fixes test_async_gcs_pub_sub_v1 test failure - agent_id is an optional field in SpendLogsPayload that is always included (as null when not provided)	2025-12-13 13:35:20 -08:00
Cesar Garcia	d693596e87	feat(langfuse): Add support for custom masking function (#17826 ) * feat(langfuse): Add support for custom masking function Allow users to pass a custom masking function via metadata to selectively redact sensitive data (credit cards, emails, PII) before sending to Langfuse. Usage: ```python def mask_pii(data): if isinstance(data, str): data = re.sub(r'\b\d{4}[\s-]?\d{4}[\s-]?\d{4}[\s-]?\d{4}\b', '[CARD]', data) return data litellm.completion( model="gpt-4", messages=[...], metadata={"langfuse_masking_function": mask_pii} ) ``` * fix(langfuse): Isolate masking function from other logging integrations Extract langfuse_masking_function from metadata early in the flow and store it in a dedicated key (_langfuse_masking_function) that only the Langfuse logger knows to look for. This prevents the callable from leaking to other logging integrations (Datadog, S3, etc.) which would serialize it as "<function at 0x...>". Changes: - scrub_sensitive_keys_in_metadata() now extracts and stores the function - Langfuse logger looks in the dedicated key first, falls back to metadata - Added tests to verify isolation works correctly	2025-12-11 15:36:54 -08:00
Ishaan Jaffer	2b069a343b	test_init_custom_logger_compatible_class_as_callback	2025-12-06 16:21:50 -08:00
rioiart	1ac2655b17	Fix/organization max budget not enforced (#17334 ) * test: add failing tests for organization budget enforcement bug Add comprehensive tests exposing that organization-level budgets are retrieved but never enforced during request authentication. Tests verify: 1. Basic org budget exceeded scenario (team under budget, org over) 2. Multiple teams collectively exceeding org budget 3. Organization budget fields exist but are never checked 4. Inconsistency between team budget enforcement (works) and org (doesn't) Tests intentionally fail to document the bug. Will be fixed in next commit. Related to organization_max_budget not being enforced in auth_checks.py * fix: enforce organization budget in auth checks Add organization budget enforcement to common_checks() in auth_checks.py. Previously, organization_max_budget was retrieved from DB but never checked, allowing teams to collectively exceed their organization's budget limit. Changes: - Add _organization_max_budget_check() function following team budget pattern - Call org budget check after team budget check in common_checks() - Add "organization_budget" to budget_alerts type literals - Update tests to verify org budget is enforced Budget hierarchy is now properly enforced: Organization Budget (hard ceiling) └─ Team Budget (sub-allocation) └─ Team Member Budget (per-user within team) └─ Key Budget (per-key) 🤖 Generated with [Claude Code](https://claude.com/claude-code) Co-Authored-By: Claude <noreply@anthropic.com> * fix: add organization_id to budget alerts, fix enum comparison and linting of newly added code - Add organization_id field to CallInfo class for better alert context - Include organization_id in budget alerts (token, soft, team, org) - Fix event_group enum comparison (was comparing enum to string) - Add OrganizationBudgetAlert class for organization budget alerting - Add organization_budget to test parameterizations - Apply Black formatting to slack_alerting.py --------- Co-authored-by: Claude <noreply@anthropic.com>	2025-12-02 22:46:03 -08:00
Ishaan Jaff	427074ac6e	Fix: Datadog callback regression when ddtrace is installed (#17393 ) * fix DD agent host logging * docs fix * test_datadog_agent_configuration * test_datadog_ignores_ddtrace_agent_host	2025-12-02 17:27:50 -08:00
Krish Dholakia	1cb5fcddba	make generic api OSS + support multiple generic API's (#17152 ) * feat(generic_api_callback.py): make generic api OSS + support multiple generic API's Enables https://github.com/BerriAI/litellm/pull/17094#discussion_r2562832967 * feat(callback_utils.py): support custom generic api callbacks * feat(generic_api_callback.py): support specifying which event types to run the generic api for * fix(litellm_logging.py): log system prompt for anthropic messages * feat(generic_api_callback.py): support generic api compatible api's - e.g. rubrik agent cloud * docs(sidebars.js): document new OSS generic api * docs(generic_api.md): document new OSS Generic API * docs(custom_webhook_api.md): document custom webhook api integration tutorial * docs(custom_webhook_api.md): cleanup * docs(custom_webhook_api.md): document what get's logged to custom webhook api * Refactor: Pass callback config to GenericAPILogger Co-authored-by: krrishdholakia <krrishdholakia@gmail.com> * Fix: Handle empty messages list in logging payload Co-authored-by: krrishdholakia <krrishdholakia@gmail.com> * Checkpoint before follow-up message Co-authored-by: krrishdholakia <krrishdholakia@gmail.com> * feat: Cache GenericAPILogger instances to improve performance Co-authored-by: krrishdholakia <krrishdholakia@gmail.com> --------- Co-authored-by: Cursor Agent <cursoragent@cursor.com>	2025-11-26 18:38:38 -08:00
yuneng-jiang	1a9b2d2206	Merge pull request #16560 from BerriAI/litellm_org_usage [Feature] Organization Usage	2025-11-26 13:55:58 -08:00
Ishaan Jaffer	529c56423c	test_opentelemetry_integration	2025-11-26 12:28:47 -08:00
yuneng-jiang	24f90679f8	Changes for CI/CD Tests	2025-11-22 10:21:52 -08:00
Ishaan Jaffer	0699430206	test logging tests + mcp server QA checks	2025-11-15 08:58:46 -08:00
Ishaan Jaffer	11cf22e7b8	add "provider_specific_fields": null	2025-11-14 18:49:35 -08:00
Ishaan Jaffer	65468353d1	provider_specific_fields	2025-11-14 18:17:50 -08:00
yuneng-jiang	898f15c33c	Add Langfuse OTEL and SQS to health check (#16514 )	2025-11-12 18:25:30 -08:00
Ishaan Jaff	abde56391b	[Fix] - Bedrock Knowledge Bases - add support for filtering kb queries (#16543 ) * test_e2e_bedrock_knowledgebase_retrieval_with_llm_api_call_with_tools_and_filters * fix vs registry * fix merging params * test_bedrock_kb_request_body_has_transformed_filters * fix typing / linting	2025-11-12 12:38:50 -08:00
Andrew Maguire	bd15250960	fix: Add atexit handlers to flush callbacks for async completions (#16487 ) Fixes #16486 ## Problem Callbacks configured via litellm.success_callback (e.g., PostHog, LangSmith) were not being invoked for litellm.acompletion() in short-lived scripts. The callbacks worked correctly for synchronous completions but async completions would queue callbacks that were lost when the script exited before GLOBAL_LOGGING_WORKER could process them. Root cause: asyncio.run() closes the event loop immediately after the async function completes, preventing the background worker from processing queued callbacks. ## Solution Implemented a two-level atexit handler approach: 1. GLOBAL_LOGGING_WORKER atexit handler (logging_worker.py): - Processes remaining callbacks from queue before exit - Creates new event loop to run pending coroutines synchronously - Applies time and iteration limits to prevent blocking shutdown 2. Integration-specific atexit handlers (posthog.py as example): - Flushes internal queue to external service - Uses synchronous HTTP client for reliable delivery - Each integration needs its own handler due to varying sync APIs ## Changes - litellm/litellm_core_utils/logging_worker.py: - Added _flush_on_exit() method - Registered atexit handler in __init__ - Processes up to MAX_ITERATIONS_TO_CLEAR_QUEUE events - Time-limited to MAX_TIME_TO_CLEAR_QUEUE seconds - litellm/integrations/posthog.py: - Added _flush_on_exit() method - Registered atexit handler in __init__ - Groups events by credentials for batch sending - Uses sync_client for reliable HTTP delivery - tests/logging_callback_tests/test_posthog.py: - Added test_async_callback_atexit_handler_exists() - Added test_posthog_atexit_flushes_internal_queue() - Added test_sync_callback_not_affected_by_atexit() ## Testing - All existing tests pass - Manual end-to-end testing confirms fix: - Async events now arrive in PostHog - Sync events continue working (no regression) - Unit tests verify atexit handlers registered and functional ## Impact - Fixes async callback delivery for ALL integrations using GLOBAL_LOGGING_WORKER - No breaking changes - only adds missing functionality - Sync path unchanged - no performance impact	2025-11-11 19:12:53 -08:00
Ishaan Jaff	e94186629d	[Fix] Bedrock Knowledge bases - ensure users can access `search_results` for both stream + non stream response to /chat/completions (#16459 ) * fix message with provider_specific_fields * test_provider_specific_fields_in_proxy_http_response * test_provider_specific_fields_in_proxy_http_response	2025-11-11 08:19:57 -08:00

1 2 3 4 5 ...

305 Commits