litellm-mirror

mirror of https://github.com/BerriAI/litellm.git synced 2025-04-27 03:34:10 +00:00

Author	SHA1	Message	Date
Ishaan Jaff	2fa9709af0	stash - langsmith use batching for logging	2024-09-11 08:06:56 -07:00
Ishaan Jaff	c1262addbe	Merge pull request #5623 from BerriAI/litellm_vertex_use_async_for_getting_token [Feat-Vertex Perf] Use async func to get auth credentials	2024-09-10 18:53:48 -07:00
Ishaan Jaff	96fa9d46f5	fix case when gemini is used	2024-09-10 17:06:45 -07:00
Ishaan Jaff	1c6f8b1be2	fix vertex use async func to set auth creds	2024-09-10 16:12:18 -07:00
Ishaan Jaff	3ebff903c3	Merge branch 'main' into litellm_use_helper_to_get_httpx_clients	2024-09-10 15:02:54 -07:00
Ishaan Jaff	f3593aed68	Merge pull request #5619 from BerriAI/litellm_vertex_use_get_httpx_client [Fix-Perf] Vertex AI cache httpx clients	2024-09-10 13:59:39 -07:00
Ishaan Jaff	08f8f9634f	use get async httpx client	2024-09-10 13:08:49 -07:00
Ishaan Jaff	421b857714	pass llm provider when creating async httpx clients	2024-09-10 11:51:42 -07:00
Ishaan Jaff	d4b9a1307d	rename get_async_httpx_client	2024-09-10 10:38:01 -07:00
Ishaan Jaff	1e8cf9f2a6	fix vertex ai use _get_async_client	2024-09-10 10:33:19 -07:00
Ishaan Jaff	428762542c	fix regen keys when no duration is passed	2024-09-10 08:04:18 -07:00
Ishaan Jaff	479b12be09	Merge branch 'main' into litellm_allow_turning_off_message_logging_for_callbacks	2024-09-09 21:59:36 -07:00
Krish Dholakia	2d2282101b	LiteLLM Minor Fixes and Improvements (09/09/2024) (#5602 ) * fix(main.py): pass default azure api version as alternative in completion call Fixes api error caused due to api version Closes https://github.com/BerriAI/litellm/issues/5584 * Fixed gemini-1.5-flash pricing (#5590) * add /key/list endpoint * bump: version 1.44.21 → 1.44.22 * docs architecture * Fixed gemini-1.5-flash pricing --------- Co-authored-by: Ishaan Jaff <ishaanjaffer0324@gmail.com> * fix(bedrock/chat.py): fix converse api stop sequence param mapping Fixes https://github.com/BerriAI/litellm/issues/5592 * fix(databricks/cost_calculator.py): handle databricks model name changes Fixes https://github.com/BerriAI/litellm/issues/5597 * fix(azure.py): support azure api version 2024-08-01-preview Closes https://github.com/BerriAI/litellm/issues/5377 * fix(proxy/_types.py): allow dev keys to call cohere /rerank endpoint Fixes issue where only admin could call rerank endpoint * fix(azure.py): check if model is gpt-4o * fix(proxy/_types.py): support /v1/rerank on non-admin routes as well * fix(cost_calculator.py): fix split on `/` logic in cost calculator --------- Co-authored-by: F1bos <44951186+F1bos@users.noreply.github.com> Co-authored-by: Ishaan Jaff <ishaanjaffer0324@gmail.com>	2024-09-09 21:56:12 -07:00
Krish Dholakia	4ac66bd843	LiteLLM Minor Fixes and Improvements (09/07/2024) (#5580 ) * fix(litellm_logging.py): set completion_start_time_float to end_time_float if none Fixes https://github.com/BerriAI/litellm/issues/5500 * feat(_init_.py): add new 'openai_text_completion_compatible_providers' list Fixes https://github.com/BerriAI/litellm/issues/5558 Handles correctly routing fireworks ai calls when done via text completions * fix: fix linting errors * fix: fix linting errors * fix(openai.py): fix exception raised * fix(openai.py): fix error handling * fix(_redis.py): allow all supported arguments for redis cluster (#5554) * Revert "fix(_redis.py): allow all supported arguments for redis cluster (#5554)" (#5583) This reverts commit `f2191ef4cb`. * fix(router.py): return model alias w/ underlying deployment on router.get_model_list() Fixes https://github.com/BerriAI/litellm/issues/5524#issuecomment-2336410666 * test: handle flaky tests --------- Co-authored-by: Jonas Dittrich <58814480+Kakadus@users.noreply.github.com>	2024-09-09 18:54:17 -07:00
Ishaan Jaff	a6d3bd0ab7	Merge branch 'main' into litellm_tag_routing_fixes	2024-09-09 17:45:18 -07:00
Ishaan Jaff	bbdcc75c60	fix log failures for key based logging	2024-09-09 16:33:06 -07:00
Ishaan Jaff	b60361fca1	fix otel test	2024-09-09 16:20:47 -07:00
Ishaan Jaff	7c9591881c	use callback_settings when intializing otel	2024-09-09 16:05:48 -07:00
Ishaan Jaff	15c761a56b	Merge pull request #5599 from BerriAI/litellm_allow_mounting_prom_callbacks [Feat] support using "callbacks" for prometheus	2024-09-09 15:00:43 -07:00
Ishaan Jaff	a1f0df3cea	fix debug statements	2024-09-09 14:00:17 -07:00
Ishaan Jaff	7ff7028885	fix create script for pre-creating views	2024-09-09 11:03:27 -07:00
Ishaan Jaff	e253c100f4	support using "callbacks" for prometheus	2024-09-09 08:26:03 -07:00
Ishaan Jaff	204c384400	add /key/list endpoint	2024-09-07 16:52:28 -07:00
Ishaan Jaff	c574c729cd	ui new build	2024-09-07 16:24:06 -07:00
Ishaan Jaff	ecb774c3e8	add doc on spend report frequency	2024-09-07 11:54:33 -07:00
Ishaan Jaff	805e4c5754	add spend_report_frequency as a general setting	2024-09-07 11:44:58 -07:00
Krish Dholakia	32d0277f03	Allow client-side credentials to be sent to proxy (accept only if complete credentials are given) (#5575 ) * feat: initial commit * fix(proxy/auth/auth_utils.py): Allow client-side credentials to be given to the proxy (accept only if complete credentials are given)	2024-09-06 19:21:54 -07:00
Ishaan Jaff	516a6b63e1	ui new build	2024-09-06 18:10:46 -07:00
Ishaan Jaff	ff9aafe05d	Merge pull request #5566 from BerriAI/litellm_ui_regen_keys [Feat] Allow setting duration time when regenerating key	2024-09-06 18:05:51 -07:00
Ishaan Jaff	09a4568172	Merge pull request #5574 from BerriAI/litellm_tags_use_views [Feat-Proxy] Use DB Views to Get spend per Tag (Usage endpoints)	2024-09-06 17:33:06 -07:00
Krish Dholakia	72e961af3c	LiteLLM Minor Fixes and Improvements (08/06/2024) (#5567 ) * fix(utils.py): return citations for perplexity streaming Fixes https://github.com/BerriAI/litellm/issues/5535 * fix(anthropic/chat.py): support fallbacks for anthropic streaming (#5542) * fix(anthropic/chat.py): support fallbacks for anthropic streaming Fixes https://github.com/BerriAI/litellm/issues/5512 * fix(anthropic/chat.py): use module level http client if none given (prevents early client closure) * fix: fix linting errors * fix(http_handler.py): fix raise_for_status error handling * test: retry flaky test * fix otel type * fix(bedrock/embed): fix error raising * test(test_openai_batches_and_files.py): skip azure batches test (for now) quota exceeded * fix(test_router.py): skip azure batch route test (for now) - hit batch quota limits --------- Co-authored-by: Ishaan Jaff <ishaanjaffer0324@gmail.com> * All `model_group_alias` should show up in `/models`, `/model/info` , `/model_group/info` (#5539) * fix(router.py): support returning model_alias model names in `/v1/models` * fix(proxy_server.py): support returning model alias'es on `/model/info` * feat(router.py): support returning model group alias for `/model_group/info` * fix(proxy_server.py): fix linting errors * fix(proxy_server.py): fix linting errors * build(model_prices_and_context_window.json): add amazon titan text premier pricing information Closes https://github.com/BerriAI/litellm/issues/5560 * feat(litellm_logging.py): log standard logging response object for pass through endpoints. Allows bedrock /invoke agent calls to be correctly logged to langfuse + s3 * fix(success_handler.py): fix linting error * fix(success_handler.py): fix linting errors * fix(team_endpoints.py): Allows admin to update team member budgets --------- Co-authored-by: Ishaan Jaff <ishaanjaffer0324@gmail.com>	2024-09-06 17:16:24 -07:00
Ishaan Jaff	7f461dbf68	fix linting	2024-09-06 16:54:43 -07:00
Ishaan Jaff	67751d0ecc	fix use view for getting tag usage	2024-09-06 16:28:24 -07:00
Ishaan Jaff	aed59abe35	allow passing expiry time to /key/regenerate	2024-09-06 08:36:34 -07:00
Krish Dholakia	f584021f7c	LiteLLM Minor Fixes and Improvements (#5537 ) * fix(vertex_ai): Fixes issue where multimodal message without text was failing vertex calls Fixes https://github.com/BerriAI/litellm/issues/5515 * fix(azure.py): move to using httphandler for oidc token calls Fixes issue where ssl certificates weren't being picked up as expected Closes https://github.com/BerriAI/litellm/issues/5522 * feat: Allows admin to set a default_max_internal_user_budget in config, and allow setting more specific values as env vars * fix(proxy_server.py): fix read for max_internal_user_budget * build(model_prices_and_context_window.json): add regional gpt-4o-2024-08-06 pricing Closes https://github.com/BerriAI/litellm/issues/5540 * test: skip re-test	2024-09-05 18:03:34 -07:00
Ishaan Jaff	8e25ba8de1	ui new build	2024-09-05 17:05:39 -07:00
Ishaan Jaff	f42a0528db	Merge branch 'main' into litellm_allow_internal_user_view_usage	2024-09-05 16:46:06 -07:00
Ishaan Jaff	c40f6f0437	fix on /user/info show all keys - even expired ones	2024-09-05 15:31:41 -07:00
Ishaan Jaff	491e50f381	fix allow internal user to view their own usage	2024-09-05 12:53:44 -07:00
Ishaan Jaff	fe55563233	fix /global/spend/provider	2024-09-05 12:48:58 -07:00
Ishaan Jaff	2ba2de5e6d	add global/spend/provider	2024-09-05 12:44:44 -07:00
Ishaan Jaff	1b42e53e06	allow internal user to view global/spend/models	2024-09-05 12:38:48 -07:00
Ishaan Jaff	0a05c24a9a	allow internal user to view their own spend	2024-09-05 12:35:04 -07:00
Ishaan Jaff	034de5b3cc	add usage endpoints for internal user	2024-09-05 12:34:41 -07:00
Ishaan Jaff	09894204a5	show /spend/logs for internal users	2024-09-05 12:14:03 -07:00
Ishaan Jaff	e0400accca	fix create view - MonthlyGlobalSpendPerUserPerKey	2024-09-05 12:11:59 -07:00
Ishaan Jaff	e6e5fb5843	add /spend/tags as allowed route for internal user	2024-09-05 10:41:43 -07:00
Krish Dholakia	ca37bb9de5	fix(pass_through_endpoints): support bedrock agents via pass through (#5527 )	2024-09-04 22:22:22 -07:00
Krish Dholakia	1e7e538261	LiteLLM Minor fixes + improvements (08/04/2024) (#5505 ) * Minor IAM AWS OIDC Improvements (#5246) * AWS IAM: Temporary tokens are valid across all regions after being issued, so it is wasteful to request one for each region. * AWS IAM: Include an inline policy, to help reduce misuse of overly permissive IAM roles. * (test_bedrock_completion.py): Ensure we are testing cross AWS region OIDC flow. * fix(router.py): log rejected requests Fixes https://github.com/BerriAI/litellm/issues/5498 * refactor: don't use verbose_logger.exception, if exception is raised User might already have handling for this. But alerting systems in prod will raise this as an unhandled error. * fix(datadog.py): support setting datadog source as an env var Fixes https://github.com/BerriAI/litellm/issues/5508 * docs(logging.md): add dd_source to datadog docs * fix(proxy_server.py): expose `/customer/list` endpoint for showing all customers * (bedrock): Fix usage with Cloudflare AI Gateway, and proxies in general. (#5509) * feat(anthropic.py): support 'cache_control' param for content when it is a string * Revert "(bedrock): Fix usage with Cloudflare AI Gateway, and proxies in gener…" (#5519) This reverts commit `3fac0349c2`. * refactor: ci/cd run again --------- Co-authored-by: David Manouchehri <david.manouchehri@ai.moda>	2024-09-04 22:16:55 -07:00
Ishaan Jaff	20a5bbe6a6	fix allow general guardrails on free tier	2024-09-04 19:59:32 -07:00

1 2 3 4 5 ...

3612 commits