Support budget/rate limit tiers for keys (#7429)

mirror of https://github.com/BerriAI/litellm.git synced 2025-04-26 03:04:13 +00:00

* feat(proxy/utils.py): get associated litellm budget from db in combined_view for key

allows user to create rate limit tiers and associate those to keys

* feat(proxy/_types.py): update the value of key-level tpm/rpm/model max budget metrics with the associated budget table values if set

allows rate limit tiers to be easily applied to keys

* docs(rate_limit_tiers.md): add doc on setting rate limit / budget tiers

make feature discoverable

* feat(key_management_endpoints.py): return litellm_budget_table value in key generate

make it easy for user to know associated budget on key creation

* fix(key_management_endpoints.py): document 'budget_id' param in `/key/generate`

* docs(key_management_endpoints.py): document budget_id usage

* refactor(budget_management_endpoints.py): refactor budget endpoints into separate file - makes it easier to run documentation testing against it

* docs(test_api_docs.py): add budget endpoints to ci/cd doc test + add missing param info to docs

* fix(customer_endpoints.py): use new pydantic obj name

* docs(user_management_heirarchy.md): add simple doc explaining teams/keys/org/users on litellm

* Litellm dev 12 26 2024 p2 (#7432)

* (Feat) Add logging for `POST v1/fine_tuning/jobs`  (#7426)

* init commit ft jobs logging

* add ft logging

* add logging for FineTuningJob

* simple FT Job create test

* (docs) - show all supported Azure OpenAI endpoints in overview  (#7428)

* azure batches

* update doc

* docs azure endpoints

* docs endpoints on azure

* docs azure batches api

* docs azure batches api

* fix(key_management_endpoints.py): fix key update to actually work

* test(test_key_management.py): add e2e test asserting ui key update call works

* fix: proxy/_types - fix linting erros

* test: update test

---------

Co-authored-by: Ishaan Jaff <ishaanjaffer0324@gmail.com>

* fix: test

* fix(parallel_request_limiter.py): enforce tpm/rpm limits on key from tiers

* fix: fix linting errors

* test: fix test

* fix: remove unused import

* test: update test

* docs(customer_endpoints.py): document new model_max_budget param

* test: specify unique key alias

* docs(budget_management_endpoints.py): document new model_max_budget param

* test: fix test

* test: fix tests

---------

Co-authored-by: Ishaan Jaff <ishaanjaffer0324@gmail.com>

This commit is contained in:

Krish Dholakia

2024-12-26 19:05:27 -08:00

• committed by

GitHub

parent 12e8fe72f9

commit d6a2beb342

25 changed files with 764 additions and 376 deletions

									
										2

litellm/router.py
									
										View file
										
				@ -98,7 +98,6 @@ from litellm.types.router import (

				    CustomRoutingStrategyBase,

				    Deployment,

				    DeploymentTypedDict,

				    GenericBudgetConfigType,

				    LiteLLM_Params,

				    ModelGroupInfo,

				    OptionalPreCallChecks,

				@ -111,6 +110,7 @@ from litellm.types.router import (

				    RoutingStrategy,

				)

				from litellm.types.services import ServiceTypes

				from litellm.types.utils import GenericBudgetConfigType

				from litellm.types.utils import ModelInfo as ModelMapInfo

				from litellm.types.utils import StandardLoggingPayload

				from litellm.utils import (

Rows
Columns

Support budget/rate limit tiers for keys (#7429)

2 litellm/router.py Unescape Escape View file

2

litellm/router.py

View file