litellm

Author	SHA1	Message	Date
dependabot[bot]	772b2f9cd2	Bump cross-spawn from 7.0.3 to 7.0.6 in /ui/litellm-dashboard (#6865 ) Bumps [cross-spawn](https://github.com/moxystudio/node-cross-spawn) from 7.0.3 to 7.0.6. - [Changelog](https://github.com/moxystudio/node-cross-spawn/blob/master/CHANGELOG.md) - [Commits](https://github.com/moxystudio/node-cross-spawn/compare/v7.0.3...v7.0.6) --- updated-dependencies: - dependency-name: cross-spawn dependency-type: indirect ... Signed-off-by: dependabot[bot] <support@github.com> Co-authored-by: dependabot[bot] <49699333+dependabot[bot]@users.noreply.github.com>	2024-11-22 17:42:08 -08:00
Ishaan Jaff	97cde31113	fix tests (#6875 )	2024-11-22 17:35:38 -08:00
Ishaan Jaff	b2b3e40d13	(feat) use `@google-cloud/vertexai` js sdk with litellm (#6873 ) * stash gemini JS test * add vertex js sdj example * handle vertex pass through separately * tes vertex JS sdk * fix vertex_proxy_route * use PassThroughStreamingHandler * fix PassThroughStreamingHandler * use common _create_vertex_response_logging_payload_for_generate_content * test vertex js * add working vertex jest tests * move basic bass through test * use good name for test * test vertex * test_chunk_processor_yields_raw_bytes * unit tests for streaming * test_convert_raw_bytes_to_str_lines * run unit tests 1st * simplify local * docs add usage example for js * use get_litellm_virtual_key * add unit tests for vertex pass through	2024-11-22 16:50:10 -08:00
Ishaan Jaff	5930c42e74	fix coverage	2024-11-22 16:21:22 -08:00
Ishaan Jaff	377cfeb24f	add pass_through_unit_testing	2024-11-22 16:20:16 -08:00
Krrish Dholakia	d8e5134935	test: skip flaky test	2024-11-22 19:23:36 +05:30
Ishaan Jaff	a6220f7a40	test - also try diff host for langfuse	2024-11-21 23:51:58 -08:00
Ishaan Jaff	701c154e35	fix test_aaateam_logging	2024-11-21 23:47:38 -08:00
Ishaan Jaff	8856256730	fix doc format	2024-11-21 23:29:40 -08:00
Ishaan Jaff	20f2bf4bbd	bump: version 1.52.13 → 1.52.14	2024-11-21 23:19:02 -08:00
Ishaan Jaff	b903134cc9	ci/cd run again	2024-11-21 23:12:54 -08:00
Ishaan Jaff	952dbb9eb7	test_langfuse_masked_input_output	2024-11-21 22:59:36 -08:00
Ishaan Jaff	366a6895e2	test_langfuse_masked_input_output	2024-11-21 22:54:18 -08:00
Ishaan Jaff	be0f0dd345	test_langfuse_masked_input_output	2024-11-21 22:51:19 -08:00
Ishaan Jaff	027967d260	test_langfuse_logging_audio_transcriptions	2024-11-21 22:46:23 -08:00
Ishaan Jaff	f398c9b172	fix test_aaateam_logging	2024-11-21 22:36:44 -08:00
Ishaan Jaff	5a2e5b43c4	fix test_aaapass_through_endpoint_pass_through_keys_langfuse	2024-11-21 22:05:00 -08:00
Ishaan Jaff	e0921da38c	test_team_logging	2024-11-21 22:01:12 -08:00
Ishaan Jaff	f77bd9a99c	test_aaalangfuse_logging_metadata	2024-11-21 21:56:36 -08:00
Ishaan Jaff	14124bab45	docs - Send `litellm_metadata` (tags)	2024-11-21 21:46:49 -08:00
Ishaan Jaff	6717929206	(Feat) Allow passing `litellm_metadata` to pass through endpoints + Add e2e tests for /anthropic/ usage tracking (#6864 ) * allow passing _litellm_metadata in pass through endpoints * fix _create_anthropic_response_logging_payload * include litellm_call_id in logging * add e2e testing for anthropic spend logs * add testing for spend logs payload * add example with anthropic python SDK	2024-11-21 21:41:05 -08:00
Ishaan Jaff	b8af46e1a2	(feat) Add usage tracking for streaming `/anthropic` passthrough routes (#6842 ) * use 1 file for AnthropicPassthroughLoggingHandler * add support for anthropic streaming usage tracking * ci/cd run again * fix - add real streaming for anthropic pass through * remove unused function stream_response * working anthropic streaming logging * fix code quality * fix use 1 file for vertex success handler * use helper for _handle_logging_vertex_collected_chunks * enforce vertex streaming to use sse for streaming * test test_basic_vertex_ai_pass_through_streaming_with_spendlog * fix type hints * add comment * fix linting * add pass through logging unit testing	2024-11-21 19:36:03 -08:00
Ishaan Jaff	920f4c9f82	(fix) add linting check to ban creating `AsyncHTTPHandler` during LLM calling (#6855 ) * fix triton * fix TEXT_COMPLETION_CODESTRAL * fix REPLICATE * fix CLARIFAI * fix HUGGINGFACE * add test_no_async_http_handler_usage * fix PREDIBASE * fix anthropic use get_async_httpx_client * fix vertex fine tuning * fix dbricks get_async_httpx_client * fix get_async_httpx_client vertex * fix get_async_httpx_client * fix get_async_httpx_client * fix make_async_azure_httpx_request * fix check_for_async_http_handler * test: cleanup mistral model * add check for AsyncClient * fix check_for_async_http_handler * fix get_async_httpx_client * fix tests using in_memory_llm_clients_cache * fix langfuse import * fix import --------- Co-authored-by: Krrish Dholakia <krrishdholakia@gmail.com>	2024-11-21 19:03:02 -08:00
Ishaan Jaff	71ebf47cef	fix latency issues on google ai studio (#6852 )	2024-11-21 19:02:08 -08:00
Krrish Dholakia	2903fd4164	docs: update json mode docs	2024-11-22 03:00:45 +05:30
Krrish Dholakia	b8edef389c	bump: version 1.52.12 → 1.52.13	2024-11-22 02:29:16 +05:30
Krish Dholakia	7e5085dc7b	Litellm dev 11 21 2024 (#6837 ) * Fix Vertex AI function calling invoke: use JSON format instead of protobuf text format. (#6702) * test: test tool_call conversion when arguments is empty dict Fixes https://github.com/BerriAI/litellm/issues/6833 * fix(openai_like/handler.py): return more descriptive error message Fixes https://github.com/BerriAI/litellm/issues/6812 * test: skip overloaded model * docs(anthropic.md): update anthropic docs to show how to route to any new model * feat(groq/): fake stream when 'response_format' param is passed Groq doesn't support streaming when response_format is set * feat(groq/): add response_format support for groq Closes https://github.com/BerriAI/litellm/issues/6845 * fix(o1_handler.py): remove fake streaming for o1 Closes https://github.com/BerriAI/litellm/issues/6801 * build(model_prices_and_context_window.json): add groq llama3.2b model pricing Closes https://github.com/BerriAI/litellm/issues/6807 * fix(utils.py): fix handling ollama response format param Fixes https://github.com/BerriAI/litellm/issues/6848#issuecomment-2491215485 * docs(sidebars.js): refactor chat endpoint placement * fix: fix linting errors * test: fix test * test: fix test * fix(openai_like/handler): handle max retries * fix(streaming_handler.py): fix streaming check for openai-compatible providers * test: update test * test: correctly handle model is overloaded error * test: update test * test: fix test * test: mark flaky test --------- Co-authored-by: Guowang Li <Guowang@users.noreply.github.com>	2024-11-22 01:53:52 +05:30
Ishaan Jaff	a7d5536872	(fix) passthrough - allow internal users to access /anthropic (#6843 ) * fix /anthropic/ * test llm_passthrough_router * fix test_gemini_pass_through_endpoint	2024-11-21 11:46:50 -08:00
Krrish Dholakia	50d2510b60	test: cleanup mistral model	2024-11-21 23:44:50 +05:30
Ishaan Jaff	ddfe687b13	(fix) don't block proxy startup if license check fails & using prometheus (#6839 ) * fix - don't block proxy startup if not a premium user * test_litellm_proxy_server_config_with_prometheus * add test for proxy startup * fix remove unused test * fix startup test * add comment on bad-license	2024-11-20 17:55:39 -08:00
Ishaan Jaff	cc1f8ff0ba	(testing) - add e2e tests for anthropic pass through endpoints (#6840 ) * tests - add e2e tests for anthropic pass through * fix swagger * fix pass through tests	2024-11-20 17:55:13 -08:00
Ishaan Jaff	c107bae7ae	(feat) add usage / cost tracking for Anthropic passthrough routes (#6835 ) * move _process_response in transformation * fix AnthropicConfig test * add AnthropicConfig * fix anthropic_passthrough_handler * fix get_response_body * fix check for streaming response * use 1 helper to return stream_response on passthrough	2024-11-20 17:25:12 -08:00
Ishaan Jaff	434b1d3d86	(refactor) anthropic - move _process_response in transformation.py (#6834 ) * move _process_response in transformation * fix AnthropicConfig test	2024-11-20 17:24:19 -08:00
Krish Dholakia	b11bc0374e	Litellm dev 11 20 2024 (#6838 ) * feat(customer_endpoints.py): support passing budget duration via `/customer/new` endpoint Closes https://github.com/BerriAI/litellm/issues/5651 * docs: add missing params to swagger + api documentation test * docs: add documentation for all key endpoints documents all params on swagger * docs(internal_user_endpoints.py): document all /user/new params Ensures all params are documented * docs(team_endpoints.py): add missing documentation for team endpoints Ensures 100% param documentation on swagger * docs(organization_endpoints.py): document all org params Adds documentation for all params in org endpoint * docs(customer_endpoints.py): add coverage for all params on /customer endpoints ensures all /customer/* params are documented * ci(config.yml): add endpoint doc testing to ci/cd * fix: fix internal_user_endpoints.py * fix(internal_user_endpoints.py): support 'duration' param * fix(partner_models/main.py): fix anthropic re-raise exception on vertex * fix: fix pydantic obj * build(model_prices_and_context_window.json): add new vertex claude model names vertex claude changed model names - causes cost tracking errors	2024-11-21 05:20:37 +05:30
Krrish Dholakia	0b0253f7ad	build: update ui build	2024-11-21 05:16:58 +05:30
Krrish Dholakia	746881485f	bump: version 1.52.11 → 1.52.12	2024-11-21 04:38:04 +05:30
Krish Dholakia	689cd677c6	Litellm dev 11 20 2024 (#6831 ) * feat(customer_endpoints.py): support passing budget duration via `/customer/new` endpoint Closes https://github.com/BerriAI/litellm/issues/5651 * docs: add missing params to swagger + api documentation test * docs: add documentation for all key endpoints documents all params on swagger * docs(internal_user_endpoints.py): document all /user/new params Ensures all params are documented * docs(team_endpoints.py): add missing documentation for team endpoints Ensures 100% param documentation on swagger * docs(organization_endpoints.py): document all org params Adds documentation for all params in org endpoint * docs(customer_endpoints.py): add coverage for all params on /customer endpoints ensures all /customer/* params are documented * ci(config.yml): add endpoint doc testing to ci/cd * fix: fix internal_user_endpoints.py * fix(internal_user_endpoints.py): support 'duration' param * fix(partner_models/main.py): fix anthropic re-raise exception on vertex * fix: fix pydantic obj	2024-11-21 04:06:06 +05:30
David Manouchehri	a1f06de53d	Add gpt-4o-2024-11-20. (#6832 )	2024-11-21 03:48:29 +05:30
Krish Dholakia	b0be5bf3a1	LiteLLM Minor Fixes & Improvements (11/19/2024) (#6820 ) * fix(anthropic/chat/transformation.py): add json schema as values: json_schema fixes passing pydantic obj to anthropic Fixes https://github.com/BerriAI/litellm/issues/6766 * (feat): Add timestamp_granularities parameter to transcription API (#6457) * Add timestamp_granularities parameter to transcription API * add param to the local test * fix(databricks/chat.py): handle max_retries optional param handling for openai-like calls Fixes issue with calling finetuned vertex ai models via databricks route * build(ui/): add team admins via proxy ui * fix: fix linting error * test: fix test * docs(vertex.md): refactor docs * test: handle overloaded anthropic model error * test: remove duplicate test * test: fix test * test: update test to handle model overloaded error --------- Co-authored-by: Show <35062952+BrunooShow@users.noreply.github.com>	2024-11-21 00:57:58 +05:30
Krrish Dholakia	7d0e1f05ac	build: run new build	2024-11-20 19:48:57 +05:30
Krrish Dholakia	6a816bceee	test: fix test	2024-11-20 14:13:14 +05:30
Ishaan Jaff	132569dafc	ci/cd run again	2024-11-19 22:38:45 -08:00
Ishaan Jaff	8631f3bb60	use correct name for test file	2024-11-19 22:11:52 -08:00
Ishaan Jaff	8b92e4f77a	fix test_prometheus_metric_tracking	2024-11-19 22:11:30 -08:00
Ishaan Jaff	7463dab9c6	(feat) provider budget routing improvements (#6827 ) * minor fix for provider budget * fix raise good error message when budget crossed for provider budget * fix test provider budgets * test provider budgets * feat - emit llm provider spend on prometheus * test_prometheus_metric_tracking * doc provider budgets	2024-11-19 21:25:08 -08:00
Ishaan Jaff	3c6fe21935	(Feat) Add provider specific budget routing (#6817 ) * add ProviderBudgetConfig * working test_provider_budgets_e2e_test * test_provider_budgets_e2e_test_expect_to_fail * use 1 cache read for getting provider spend * test_provider_budgets_e2e_test * add doc on provider budgets * clean up provider budgets * unit testing for provider budget routing * use as flag, not routing strat * fix init provider budget routing * use async_filter_deployments * fix test provider budgets * doc provider budget routing * doc provider budget routing * fix docs changes * fix comment	2024-11-19 20:25:27 -08:00
Krrish Dholakia	59a9b71d21	build: fix test	2024-11-20 05:50:08 +05:30
Krish Dholakia	cf579fe644	Litellm stable pr 10 30 2024 (#6821 ) * Update organization_endpoints.py to be able to list organizations (#6473) * Update organization_endpoints.py to be able to list organizations * Update test_organizations.py * Update test_organizations.py add test for list * Update test_organizations.py correct indentation * Add unreleased Claude 3.5 Haiku models. (#6476) --------- Co-authored-by: superpoussin22 <vincent.nadal@orange.fr> Co-authored-by: David Manouchehri <david.manouchehri@ai.moda>	2024-11-20 05:03:42 +05:30
Ishaan Jaff	98c7889013	feat - add qwen2p5-coder-32b-instruct (#6818 )	2024-11-19 14:50:51 -08:00
Ishaan Jaff	1890fde3f3	(Proxy) add support for DOCS_URL and REDOC_URL (#6806 ) * add support for DOCS_URL and REDOC_URL * document env vars * add unit tests for docs url and redocs url	2024-11-19 07:02:12 -08:00

1 2 3 4 5 ...

18422 commits