llama-stack

forked from phoenix-oss/llama-stack-mirror

History

Ben Browning fa5dfee07b fix: Return HTTP 400 for OpenAI API validation errors (#2002 ) # What does this PR do? When clients called the Open AI API with invalid input that wasn't caught by our own Pydantic API validation but instead only caught by the backend inference provider, that backend inference provider was returning a HTTP 400 error. However, we were wrapping that into a HTTP 500 error, obfuscating the actual issue from calling clients and triggering OpenAI client retry logic. This change adjusts our existing `translate_exception` method in `server.py` to wrap `openai.BadRequestError` as HTTP 400 errors, passing through the string representation of the error message to the calling user so they can see the actual input validation error and correct it. I tried changing this in a few other places, but ultimately `translate_exception` was the only real place to handle this for both streaming and non-streaming requests across all inference providers that use the OpenAI server APIs. This also tightens up our validation a bit for the OpenAI chat completions API, to catch empty `messages` parameters, invalid `tool_choice` parameters, invalid `tools` items, or passing `tool_choice` when `tools` isn't given. Lastly, this extends our OpenAI API chat completions verifications to also check for consistent input validation across providers. Providers behind Llama Stack should automatically pass all the new tests due to the input validation added here, but some of the providers fail this test when not run behind Llama Stack due to differences in how they handle input validation and errors. (Closes #1951) ## Test Plan To test this, start an OpenAI API verification stack: ``` llama stack run --image-type venv tests/verifications/openai-api-verification-run.yaml ``` Then, run the new verification tests with your provider(s) of choice: ``` python -m pytest -s -v \ tests/verifications/openai_api/test_chat_completion.py \ --provider openai-llama-stack python -m pytest -s -v \ tests/verifications/openai_api/test_chat_completion.py \ --provider together-llama-stack ``` Signed-off-by: Ben Browning <bbrownin@redhat.com>		2025-04-23 17:48:32 +02:00
..
routers	fix: Return HTTP 400 for OpenAI API validation errors (#2002 )	2025-04-23 17:48:32 +02:00
server	fix: Return HTTP 400 for OpenAI API validation errors (#2002 )	2025-04-23 17:48:32 +02:00
store	fix: handle registry errors gracefully (#1732 )	2025-03-20 15:24:07 -07:00
ui	feat: add tool name to chat output in playground (#1996 )	2025-04-23 15:57:54 +02:00
utils	feat: add health to all providers through providers endpoint (#1418 )	2025-04-14 11:59:36 +02:00
__init__.py	API Updates (#73 )	2024-09-17 19:51:35 -07:00
access_control.py	feat: make sure agent sessions are under access control (#1737 )	2025-03-21 07:31:16 -07:00
build.py	feat: allow building distro with external providers (#1967 )	2025-04-18 17:18:28 +02:00
build_conda_env.sh	chore: remove straggler references to llama-models (#1345 )	2025-03-01 14:26:03 -08:00
build_container.sh	feat: allow building distro with external providers (#1967 )	2025-04-18 17:18:28 +02:00
build_venv.sh	chore: remove straggler references to llama-models (#1345 )	2025-03-01 14:26:03 -08:00
client.py	chore: move all Llama Stack types from llama-models to llama-stack (#1098 )	2025-02-14 09:10:59 -08:00
common.sh	fix: Fixing some small issues with the build scripts (#1132 )	2025-02-19 22:20:49 -08:00
configure.py	feat: add provider API for listing and inspecting provider info (#1429 )	2025-03-13 15:07:21 -07:00
datatypes.py	feat: allow building distro with external providers (#1967 )	2025-04-18 17:18:28 +02:00
distribution.py	feat: allow building distro with external providers (#1967 )	2025-04-18 17:18:28 +02:00
inspect.py	feat: add health to all providers through providers endpoint (#1418 )	2025-04-14 11:59:36 +02:00
library_client.py	feat: add health to all providers through providers endpoint (#1418 )	2025-04-14 11:59:36 +02:00
providers.py	feat: add health to all providers through providers endpoint (#1418 )	2025-04-14 11:59:36 +02:00
request_headers.py	feat(server): add attribute based access control for resources (#1703 )	2025-03-19 21:28:52 -07:00
resolver.py	feat: add health to all providers through providers endpoint (#1418 )	2025-04-14 11:59:36 +02:00
stack.py	feat: add health to all providers through providers endpoint (#1418 )	2025-04-14 11:59:36 +02:00
start_stack.sh	docs: Update docs and fix warning in start-stack.sh (#1937 )	2025-04-11 16:26:17 -07:00