llama-stack-mirror

mirror of https://github.com/meta-llama/llama-stack.git synced 2025-12-03 09:53:45 +00:00

History

Ben Browning 8dfce2f596 feat: OpenAI Responses API (#1989 ) # What does this PR do? This provides an initial [OpenAI Responses API](https://platform.openai.com/docs/api-reference/responses) implementation. The API is not yet complete, and this is more a proof-of-concept to show how we can store responses in our key-value stores and use them to support the Responses API concepts like `previous_response_id`. ## Test Plan I've added a new `tests/integration/openai_responses/test_openai_responses.py` as part of a test-driven development for this new API. I'm only testing this locally with the remote-vllm provider for now, but it should work with any of our inference providers since the only API it requires out of the inference provider is the `openai_chat_completion` endpoint. ``` VLLM_URL="http://localhost:8000/v1" \ INFERENCE_MODEL="meta-llama/Llama-3.2-3B-Instruct" \ llama stack build --template remote-vllm --image-type venv --run ``` ``` LLAMA_STACK_CONFIG="http://localhost:8321" \ python -m pytest -v \ tests/integration/openai_responses/test_openai_responses.py \ --text-model "meta-llama/Llama-3.2-3B-Instruct" ``` --------- Signed-off-by: Ben Browning <bbrownin@redhat.com> Co-authored-by: Ashwin Bharambe <ashwin.bharambe@gmail.com>		2025-04-28 14:06:00 -07:00
..
apis	feat: OpenAI Responses API (#1989 )	2025-04-28 14:06:00 -07:00
cli	feat(cli): add interactive tab completion for image type selection (#2027 )	2025-04-25 16:57:42 +02:00
distribution	feat: Add Kubernetes authentication (#1778 )	2025-04-28 22:24:58 +02:00
models	docs: update prompt_format.md for llama4 (#2035 )	2025-04-25 15:52:15 -07:00
providers	feat: OpenAI Responses API (#1989 )	2025-04-28 14:06:00 -07:00
strong_typing	feat: OpenAI Responses API (#1989 )	2025-04-28 14:06:00 -07:00
templates	feat: Add NVIDIA NeMo datastore (#1852 )	2025-04-28 09:41:59 -07:00
__init__.py	export LibraryClient	2024-12-13 12:08:00 -08:00
env.py	refactor(test): move tools, evals, datasetio, scoring and post training tests (#1401 )	2025-03-04 14:53:47 -08:00
log.py	chore: Remove style tags from log formatter (#1808 )	2025-03-27 10:18:21 -04:00
schema_utils.py	fix: dont check protocol compliance for experimental methods	2025-04-12 16:26:32 -07:00