llama-stack-mirror

mirror of https://github.com/meta-llama/llama-stack.git synced 2025-12-30 21:50:01 +00:00

History

Ben Browning ac5dc8fae2 Add prompt_logprobs and guided_choice to OpenAI completions This adds the vLLM-specific extra_body parameters of prompt_logprobs and guided_choice to our openai_completion inference endpoint. The plan here would be to expand this to support all common optional parameters of any of the OpenAI providers, allowing each provider to use or ignore these parameters based on whether their server supports them. Signed-off-by: Ben Browning <bbrownin@redhat.com>		2025-04-09 15:47:02 -04:00
..
agents	feat(telemetry): clean up spans (#1760 )	2025-03-21 20:05:11 -07:00
batch_inference	fix: solve ruff B008 warnings (#1444 )	2025-03-06 16:48:35 -08:00
benchmarks	fix: return 4xx for non-existent resources in GET requests (#1635 )	2025-03-18 14:06:53 -07:00
common	refactor: extract pagination logic into shared helper function (#1770 )	2025-03-31 13:08:29 -07:00
datasetio	refactor: extract pagination logic into shared helper function (#1770 )	2025-03-31 13:08:29 -07:00
datasets	chore: Don't set type variables from register_schema() (#1713 )	2025-03-19 20:29:00 -07:00
eval	fix: fix jobs api literal return type (#1757 )	2025-03-21 14:04:21 -07:00
files	feat(api): don't return a payload on file delete (#1640 )	2025-03-25 17:12:36 -07:00
inference	Add prompt_logprobs and guided_choice to OpenAI completions	2025-04-09 15:47:02 -04:00
inspect	chore: deprecate /v1/inspect/providers (#1678 )	2025-03-19 20:27:06 -07:00
models	Use our own pydantic models for OpenAI Server APIs	2025-04-09 15:47:02 -04:00
post_training	fix: Restore discriminator for AlgorithmConfig (#1706 )	2025-03-20 07:33:26 -07:00
providers	fix: OpenAPI with provider get (#1627 )	2025-03-13 19:56:32 -07:00
safety	chore: move all Llama Stack types from llama-models to llama-stack (#1098 )	2025-02-14 09:10:59 -08:00
scoring	docs: api documentation for agents/eval/scoring/datasets (#1400 )	2025-03-05 09:40:24 -08:00
scoring_functions	chore: Don't set type variables from register_schema() (#1713 )	2025-03-19 20:29:00 -07:00
shields	fix: return 4xx for non-existent resources in GET requests (#1635 )	2025-03-18 14:06:53 -07:00
synthetic_data_generation	chore: move all Llama Stack types from llama-models to llama-stack (#1098 )	2025-02-14 09:10:59 -08:00
telemetry	chore: Don't set type variables from register_schema() (#1713 )	2025-03-19 20:29:00 -07:00
tools	fix(api): don't return list for runtime tools (#1686 )	2025-04-01 09:53:11 +02:00
vector_dbs	fix: return 4xx for non-existent resources in GET requests (#1635 )	2025-03-18 14:06:53 -07:00
vector_io	chore: mypy violations cleanup for inline::{telemetry,tool_runtime,vector_io} (#1711 )	2025-03-20 10:01:10 -07:00
__init__.py	API Updates (#73 )	2024-09-17 19:51:35 -07:00
datatypes.py	feat(api): don't return a payload on file delete (#1640 )	2025-03-25 17:12:36 -07:00
resource.py	fix!: update eval-tasks -> benchmarks (#1032 )	2025-02-13 16:40:58 -08:00
version.py	llama-stack version alpha -> v1	2025-01-15 05:58:09 -08:00