llama-stack-mirror

mirror of https://github.com/meta-llama/llama-stack.git synced 2025-12-08 03:00:56 +00:00

History

Akram Ben Aissi a548169b99 fix: allow skipping model availability check for vLLM (#3739 ) # What does this PR do? <!-- Provide a short summary of what this PR does and why. Link to relevant issues if applicable. --> Allows model check to fail gracefully instead of crashing on startup. <!-- If resolving an issue, uncomment and update the line below --> <!-- Closes #[issue-number] --> ## Test Plan <!-- Describe the tests you ran to verify your changes with result summaries. Provide clear instructions so the plan can be easily re-executed. --> set VLLM_URL to your VLLM server ``` (base) akram@Mac llama-stack % LAMA_STACK_LOGGING="all=debug" VLLM_ENABLE_MODEL_DISCOVERY=false MILVUS_DB_PATH=./milvus.db INFERENCE_MODEL=vllm uv run --with llama-stack llama stack build --distro starter --image-type venv --run ``` ``` INFO 2025-10-08 20:11:24,637 llama_stack.providers.utils.inference.inference_store:74 inference: Write queue disabled for SQLite to avoid concurrency issues INFO 2025-10-08 20:11:24,866 llama_stack.providers.utils.responses.responses_store:96 openai_responses: Write queue disabled for SQLite to avoid concurrency issues ERROR 2025-10-08 20:11:26,160 llama_stack.providers.utils.inference.openai_mixin:439 providers::utils: VLLMInferenceAdapter.list_provider_model_ids() failed with: <a href="https://oauth.akram.a1ey.p3.openshiftapps.com:443/oauth/authorize?approval_prompt=force&client_id=system%3Aserviceaccount%3Arhoai-30-genai%3Adefault&redirect_uri=ht tps%3A%2F%2Fvllm-rhoai-30-genai.apps.rosa.akram.a1ey.p3.openshiftapps.com%2Foauth%2Fcallback&response_type=code&scope=user%3Ainfo+user%3Acheck-access&state=9fba207425 5851c718aca717a5887d76%3A%2Fmodels">Found</a>. [...] INFO 2025-10-08 20:11:26,295 uvicorn.error:84 uncategorized: Started server process [83144] INFO 2025-10-08 20:11:26,296 uvicorn.error:48 uncategorized: Waiting for application startup. INFO 2025-10-08 20:11:26,297 llama_stack.core.server.server:170 core::server: Starting up INFO 2025-10-08 20:11:26,297 llama_stack.core.stack:399 core: starting registry refresh task INFO 2025-10-08 20:11:26,311 uvicorn.error:62 uncategorized: Application startup complete. INFO 2025-10-08 20:11:26,312 uvicorn.error:216 uncategorized: Uvicorn running on http://['::', '0.0.0.0']:8321 (Press CTRL+C to quit) ERROR 2025-10-08 20:11:26,791 llama_stack.providers.utils.inference.openai_mixin:439 providers::utils: VLLMInferenceAdapter.list_provider_model_ids() failed with: <a href="https://oauth.akram.a1ey.p3.openshiftapps.com:443/oauth/authorize?approval_prompt=force&client_id=system%3Aserviceaccount%3Arhoai-30-genai%3Adefault&redirect_uri=ht tps%3A%2F%2Fvllm-rhoai-30-genai.apps.rosa.akram.a1ey.p3.openshiftapps.com%2Foauth%2Fcallback&response_type=code&scope=user%3Ainfo+user%3Acheck-access&state=8ef0cba3e1 71a4f8b04cb445cfb91a4c%3A%2Fmodels">Found</a>. ```		2025-10-10 07:23:13 -07:00
..
apis	feat(responses): add usage types to inference and responses APIs (#3764 )	2025-10-10 09:22:59 -04:00
cli	chore!: remove model mgmt from CLI for Hugging Face CLI (#3700 )	2025-10-09 16:50:33 -07:00
core	fix(inference): propagate 401/403 errors from remote providers (#3762 )	2025-10-09 18:34:39 -07:00
distributions	chore!: remove model mgmt from CLI for Hugging Face CLI (#3700 )	2025-10-09 16:50:33 -07:00
models	chore: remove dead code (#3729 )	2025-10-07 20:26:02 -07:00
providers	fix: allow skipping model availability check for vLLM (#3739 )	2025-10-10 07:23:13 -07:00
strong_typing	feat: Add OpenAI Conversations API (#3429 )	2025-10-03 08:47:18 -07:00
testing	fix(testing): improve api_recorder error messages for missing recordings (#3760 )	2025-10-09 15:04:16 -07:00
ui	chore(ui-deps): bump react-dom and @types/react-dom in /llama_stack/ui (#3693 )	2025-10-06 00:02:31 -04:00
__init__.py	chore(rename): move llama_stack.distribution to llama_stack.core (#2975 )	2025-07-30 23:30:53 -07:00
env.py	refactor(test): move tools, evals, datasetio, scoring and post training tests (#1401 )	2025-03-04 14:53:47 -08:00
log.py	fix(tests): ensure test isolation in server mode (#3737 )	2025-10-08 12:03:36 -07:00
schema_utils.py	feat(api): add extra_body parameter support with shields example (#3670 )	2025-10-03 13:25:09 -07:00