llama-stack-mirror

mirror of https://github.com/meta-llama/llama-stack.git synced 2025-10-10 13:28:40 +00:00

History

Eric Huang a93130e323 test # What does this PR do? ## Test Plan # What does this PR do? ## Test Plan # What does this PR do? ## Test Plan Completes the refactoring started in previous commit by: 1. Fix library client (critical): Add logic to detect Pydantic model parameters and construct them properly from request bodies. The key fix is to NOT exclude any params when converting the body for Pydantic models - we need all fields to pass to the Pydantic constructor. Before: _convert_body excluded all params, leaving body empty for Pydantic construction After: Check for Pydantic params first, skip exclusion, construct model with full body 2. Update remaining providers to use new Pydantic-based signatures: - litellm_openai_mixin: Extract extra fields via __pydantic_extra__ - databricks: Use TYPE_CHECKING import for params type - llama_openai_compat: Use TYPE_CHECKING import for params type - sentence_transformers: Update method signatures to use params 3. Update unit tests to use new Pydantic signature: - test_openai_mixin.py: Use OpenAIChatCompletionRequestParams This fixes test failures where the library client was trying to construct Pydantic models with empty dictionaries. The previous fix had a bug: it called _convert_body() which only keeps fields that match function parameter names. For Pydantic methods with signature: openai_chat_completion(params: OpenAIChatCompletionRequestParams) The signature only has 'params', but the body has 'model', 'messages', etc. So _convert_body() returned an empty dict. Fix: Skip _convert_body() entirely for Pydantic params. Use the raw body directly to construct the Pydantic model (after stripping NOT_GIVENs). This properly fixes the ValidationError where required fields were missing. The streaming code path (_call_streaming) had the same issue as non-streaming: it called _convert_body() which returned empty dict for Pydantic params. Applied the same fix as commit 7476c0ae: - Detect Pydantic model parameters before body conversion - Skip _convert_body() for Pydantic params - Construct Pydantic model directly from raw body (after stripping NOT_GIVENs) This fixes streaming endpoints like openai_chat_completion with stream=True. The streaming code path (_call_streaming) had the same issue as non-streaming: it called _convert_body() which returned empty dict for Pydantic params. Applied the same fix as commit 7476c0ae: - Detect Pydantic model parameters before body conversion - Skip _convert_body() for Pydantic params - Construct Pydantic model directly from raw body (after stripping NOT_GIVENs) This fixes streaming endpoints like openai_chat_completion with stream=True.		2025-10-09 13:53:33 -07:00
..
anthropic	chore: turn OpenAIMixin into a pydantic.BaseModel (#3671 )	2025-10-06 11:33:19 -04:00
azure	chore: turn OpenAIMixin into a pydantic.BaseModel (#3671 )	2025-10-06 11:33:19 -04:00
bedrock	test	2025-10-09 13:53:33 -07:00
cerebras	chore: turn OpenAIMixin into a pydantic.BaseModel (#3671 )	2025-10-06 11:33:19 -04:00
databricks	test	2025-10-09 13:53:33 -07:00
fireworks	chore: turn OpenAIMixin into a pydantic.BaseModel (#3671 )	2025-10-06 11:33:19 -04:00
gemini	chore: turn OpenAIMixin into a pydantic.BaseModel (#3671 )	2025-10-06 11:33:19 -04:00
groq	chore: turn OpenAIMixin into a pydantic.BaseModel (#3671 )	2025-10-06 11:33:19 -04:00
llama_openai_compat	test	2025-10-09 13:53:33 -07:00
nvidia	chore: remove dead code (#3729 )	2025-10-07 20:26:02 -07:00
ollama	feat: add refresh_models support to inference adapters (default: false) (#3719 )	2025-10-07 15:19:56 +02:00
openai	chore: turn OpenAIMixin into a pydantic.BaseModel (#3671 )	2025-10-06 11:33:19 -04:00
passthrough	test	2025-10-09 13:53:33 -07:00
runpod	test	2025-10-09 13:53:33 -07:00
sambanova	chore: turn OpenAIMixin into a pydantic.BaseModel (#3671 )	2025-10-06 11:33:19 -04:00
tgi	chore: turn OpenAIMixin into a pydantic.BaseModel (#3671 )	2025-10-06 11:33:19 -04:00
together	feat: add refresh_models support to inference adapters (default: false) (#3719 )	2025-10-07 15:19:56 +02:00
vertexai	chore: turn OpenAIMixin into a pydantic.BaseModel (#3671 )	2025-10-06 11:33:19 -04:00
vllm	test	2025-10-09 13:53:33 -07:00
watsonx	fix: Update watsonx.ai provider to use LiteLLM mixin and list all models (#3674 )	2025-10-08 07:29:43 -04:00
__init__.py	`impls` -> `inline`, `adapters` -> `remote` (#381 )	2024-11-06 14:54:05 -08:00