llama-stack-mirror

mirror of https://github.com/meta-llama/llama-stack.git synced 2025-12-03 09:53:45 +00:00

Author	SHA1	Message	Date
Sébastien Han	71da65ae1a	fix: post is paginated not get Signed-off-by: Sébastien Han <seb@redhat.com>	2025-11-14 15:21:43 +01:00
Sébastien Han	c921a37200	fix: revert "fix: pagination config" This reverts commit `9a36669ef3`. Signed-off-by: Sébastien Han <seb@redhat.com>	2025-11-14 15:08:32 +01:00
Sébastien Han	3dd252ef3e	chore: fix missing endpoint on stainless config delete /v1/scoring-functions/{scoring_fn_id} exists in the OpenAPI spec, but isn't specified in the Stainless config, so code will not be generated for it. delete /v1alpha/eval/benchmarks/{benchmark_id} exists in the OpenAPI spec, but isn't specified in the Stainless config, so code will not be generated for it Signed-off-by: Sébastien Han <seb@redhat.com>	2025-11-14 11:59:13 +01:00
Sébastien Han	738d4bfd7e	chore: add paginated false to chat/completions Signed-off-by: Sébastien Han <seb@redhat.com>	2025-11-14 11:55:35 +01:00
Sébastien Han	9a36669ef3	fix: pagination config use paginated endpoint for example and mark input_items.list as non-paginated Signed-off-by: Sébastien Han <seb@redhat.com>	2025-11-14 11:20:42 +01:00
Sébastien Han	da37b2a847	chore: revert "fix: Exclude deprecated endpoints from stainless config" This reverts commit `06acbdab6f`. Signed-off-by: Sébastien Han <seb@redhat.com>	2025-11-14 11:19:22 +01:00
Sébastien Han	eabd248ea0	chore: remove RAG Signed-off-by: Sébastien Han <seb@redhat.com>	2025-11-14 11:12:35 +01:00
Sébastien Han	7908a1026a	chore: revert "fix: remove unused endpoint and outdate code" This reverts commit `7bc9aeaf9c`. Signed-off-by: Sébastien Han <seb@redhat.com>	2025-11-14 11:10:58 +01:00
Sébastien Han	6de63d5de1	chore: revert "fix: remove unregister shield" This reverts commit `84277988b8`. Signed-off-by: Sébastien Han <seb@redhat.com>	2025-11-14 11:09:51 +01:00
Sébastien Han	1b982ff2b6	chore: re-add missing decorator Signed-off-by: Sébastien Han <seb@redhat.com>	2025-11-14 11:01:13 +01:00
Sébastien Han	bb34f7a4d4	chore: re-add deprecated routes to the combined spec Matches https://github.com/llamastack/llama-stack/pull/4156 Signed-off-by: Sébastien Han <seb@redhat.com>	2025-11-14 10:15:53 +01:00
Sébastien Han	2a257dbdea	chore: rebase on main Signed-off-by: Sébastien Han <seb@redhat.com>	2025-11-14 10:03:44 +01:00
Sébastien Han	912ee24bdf	fix: convert anyOf with const values to enum types in OpenAPI schema Add a post-processing step that converts anyOf schemas containing multiple const string values into proper enum types. This fixes the Schema/EnumDescriptionNotValid error from Stainless by ensuring enum schemas are properly formatted instead of using anyOf with const values. Signed-off-by: Sébastien Han <seb@redhat.com>	2025-11-14 09:56:45 +01:00
Sébastien Han	769cfe4654	fix: pagination config use paginated endpoint for example and mark input_items.list as non-paginated Signed-off-by: Sébastien Han <seb@redhat.com>	2025-11-14 09:56:45 +01:00
Sébastien Han	f7d0494927	fix: Added the missing endpoints to the Stainless config Signed-off-by: Sébastien Han <seb@redhat.com>	2025-11-14 09:56:45 +01:00
Sébastien Han	e0a69f2709	fix: remove unsused ressources removed in https://github.com/llamastack/llama-stack/pull/4067 Signed-off-by: Sébastien Han <seb@redhat.com>	2025-11-14 09:56:45 +01:00
Sébastien Han	84277988b8	fix: remove unregister shield https://github.com/llamastack/llama-stack/pull/4099 Signed-off-by: Sébastien Han <seb@redhat.com>	2025-11-14 09:56:45 +01:00
Sébastien Han	827cc9b9b8	fix: deprecated endpoint in Stainless config example Replace deprecated `post /v1/models` with `get /v1/models` in the headline example to fix Stainless Endpoint/NotFound error. Signed-off-by: Sébastien Han <seb@redhat.com>	2025-11-14 09:56:44 +01:00
Sébastien Han	7bc9aeaf9c	fix: remove unused endpoint and outdate code The register/unregister were removed in https://github.com/llamastack/llama-stack/pull/4099 Signed-off-by: Sébastien Han <seb@redhat.com>	2025-11-14 09:56:44 +01:00
Sébastien Han	06acbdab6f	fix: Exclude deprecated endpoints from stainless config Filter out deprecated endpoints from the combined OpenAPI spec and remove their references from the Stainless config to fix Endpoint/NotFound errors. Signed-off-by: Sébastien Han <seb@redhat.com>	2025-11-14 09:56:44 +01:00
Sébastien Han	24b275d0dd	fix: revert "chore: add deprecated to combined schema" This reverts commit 53fc2a05812ebf24d5598a70972c86d72c50fd4e. Signed-off-by: Sébastien Han <seb@redhat.com>	2025-11-14 09:56:44 +01:00
Sébastien Han	01f441b3ac	fix: duplicate union type declarations for Stainless codegen Extract duplicate union types to shared schema references and remove duplicate references within unions to fix Stainless duplicate declaration warnings. Fixes: https://www.stainless.com/docs/reference/diagnostics#Python/DuplicateDeclaration Signed-off-by: Sébastien Han <seb@redhat.com>	2025-11-14 09:56:44 +01:00
Sébastien Han	09280301de	fix: Query default values can't be set in Annotated The error is that Query default values can't be set in Annotated; they must be set with = in the function signature. See the error: ``` The error is that Query default values can't be set in Annotated; they must be set with = in the function signature. Searching for where include_embeddings is defined: ``` Signed-off-by: Sébastien Han <seb@redhat.com>	2025-11-14 09:56:44 +01:00
Sébastien Han	221f28b685	chore: fix missing titles for unions Added _add_titles_to_unions() to: Recursively scan all schemas for anyOf/oneOf unions Generate descriptive titles from the union members Add those titles to help code generators infer names Signed-off-by: Sébastien Han <seb@redhat.com>	2025-11-14 09:56:30 +01:00
Sébastien Han	500804f0eb	chore: add deprecated to combined schema The _filter_combined_schema function was excluding deprecated operations. I updated it to include all operations (deprecated and non-deprecated) for the combined/stainless spec, so these deprecated endpoints are now included. Signed-off-by: Sébastien Han <seb@redhat.com>	2025-11-14 09:56:30 +01:00
Sébastien Han	73861b504d	chore: re-add missing endpoints _filter_combined_schema was using path-level filtering with _is_path_deprecated, which excluded entire paths if any operation was deprecated. Since /v1/toolgroups has both GET (not deprecated) and POST (deprecated), the entire path was excluded, removing the GET operation and its response schema. Updated _filter_combined_schema to use operation-level filtering, matching _filter_schema_by_version Signed-off-by: Sébastien Han <seb@redhat.com>	2025-11-14 09:56:29 +01:00
Sébastien Han	2cb0c31edd	chore: re-add missing endpoints Signed-off-by: Sébastien Han <seb@redhat.com>	2025-11-14 09:56:18 +01:00
Sébastien Han	3d33291f23	chore: refactor code to reduce generator script length Signed-off-by: Sébastien Han <seb@redhat.com>	2025-11-14 09:56:18 +01:00
Sébastien Han	de4ed29310	chore: replace JSON requestBody block with query params Signed-off-by: Sébastien Han <seb@redhat.com>	2025-11-14 09:56:18 +01:00
Sébastien Han	e3d831f504	chore: re-add text/event-stream media type Signed-off-by: Sébastien Han <seb@redhat.com>	2025-11-14 09:56:18 +01:00
Sébastien Han	66056ddb87	chore: re-add x-llama-stack-extra-body-params Signed-off-by: Sébastien Han <seb@redhat.com>	2025-11-14 09:56:18 +01:00
Sébastien Han	c4cad890cc	chore: regen scehma with main Signed-off-by: Sébastien Han <seb@redhat.com>	2025-11-14 09:56:17 +01:00
Sébastien Han	8e1f89b32e	chore: update generator script location Signed-off-by: Sébastien Han <seb@redhat.com>	2025-11-14 09:56:04 +01:00
Sébastien Han	e3cb8ed74a	chore: use Pydantic to generate OpenAPI schema Removes the need for the strong_typing and pyopenapi packages and purely use Pydantic for schema generation. Our generator now purely relies on Pydantic and FastAPI, it is available at `scripts/fastapi_generator.py`, you can run it like so: ``` uv run ./scripts/run_openapi_generator.sh ``` The generator will: * Generate the deprecated, experimental, stable and combined specs * Validate all the spec it generates against OpenAPI standards A few changes in the schema required for oasdiff some updates so I've made the following ignore rules. The new Pydantic-based generator is likely more correct and follows OpenAPI standards better than the old pyopenapi generator. Instead of trying to make the new generator match the old one's quirks, we should focus on what's actually correct according to OpenAPI standards. These are non-critical changes: * response-property-became-nullable: Backward compatible: existing non-null values still work, now also accepts null * response-required-property-removed: oasdiff reports a false positive because it doesn't resolve $refs inside anyOf; we could use tool like 'redocly' to flatten the schema to a single file. * response-property-type-changed: properties are still object types, but oasdiff doesn't resolve $refs, so it flags the missing inline type: object even though the referenced schemas define type: object * request-property-one-of-removed: These are false positives caused by schema restructuring (wrapping in anyOf for nullability, using -Input variants, or simplifying nested oneOf structures) that don't change the actual API contract - the same data types are still accepted, just represented differently in the schema. * request-parameter-enum-value-removed: These are false positives caused by oasdiff not resolving $refs - the enum values (asc, desc, assistants, batch) are still present in the referenced schemas (Order and OpenAIFilePurpose), just represented via schema references instead of inline enums. * request-property-enum-value-removed: this is a false positive caused by oasdiff not resolving $refs - the enum values (llm, embedding, rerank) are still present in the referenced ModelType schema, just represented via schema reference instead of inline enums. * request-property-type-changed: These are schema quality issues where type information is missing (due to Any fallback in dynamic model creation), but the API contract remains unchanged - properties still exist with correct names and defaults, so the same requests will work. * response-body-type-changed: These are false positives caused by schema representation changes (from inferred/empty types to explicit $ref schemas, or vice versa) - the actual response types an API contract remain unchanged, just how they're represented in the OpenAPI spec. * response-media-type-removed: This is a false positive caused by FastAPI's OpenAPI generator not documenting union return types with AsyncIterator - the streaming functionality with text/event-stream media type still works when stream=True is passed, it's just not reflected in the generated OpenAPI spec. * request-body-type-changed: This is a schema correction - the old spec incorrectly represented the request body as an object, but the function signature shows chunks: list[Chunk], so the new spec correctly shows it as an array, matching the actual API implementation. Signed-off-by: Sébastien Han <seb@redhat.com>	2025-11-14 09:56:02 +01:00
Ashwin Bharambe	2441ca9389	fix(api): ensure openapi spec has deprecated routes (#4156 ) Some checks failed SqlStore Integration Tests / test-postgres (3.12) (push) Failing after 0s Details Integration Auth Tests / test-matrix (oauth2_token) (push) Failing after 1s Details SqlStore Integration Tests / test-postgres (3.13) (push) Failing after 0s Details Test Llama Stack Build / generate-matrix (push) Successful in 3s Details Integration Tests (Replay) / generate-matrix (push) Successful in 5s Details Test External Providers Installed via Module / test-external-providers-from-module (venv) (push) Has been skipped Details Test llama stack list-deps / generate-matrix (push) Successful in 3s Details Python Package Build Test / build (3.12) (push) Failing after 4s Details API Conformance Tests / check-schema-compatibility (push) Successful in 19s Details Python Package Build Test / build (3.13) (push) Failing after 17s Details Test External API and Providers / test-external (venv) (push) Failing after 30s Details Test llama stack list-deps / list-deps-from-config (push) Successful in 36s Details Test Llama Stack Build / build-single-provider (push) Successful in 40s Details Test llama stack list-deps / show-single-provider (push) Successful in 48s Details Vector IO Integration Tests / test-matrix (push) Failing after 55s Details Test Llama Stack Build / build (push) Successful in 48s Details UI Tests / ui-tests (22) (push) Successful in 54s Details Test llama stack list-deps / list-deps (push) Failing after 1m34s Details Test Llama Stack Build / build-custom-container-distribution (push) Successful in 2m6s Details Unit Tests / unit-tests (3.13) (push) Failing after 2m38s Details Integration Tests (Replay) / Integration Tests (, , , client=, ) (push) Failing after 2m38s Details Unit Tests / unit-tests (3.12) (push) Failing after 2m44s Details Test Llama Stack Build / build-ubi9-container-distribution (push) Successful in 2m50s Details Pre-commit / pre-commit (push) Successful in 3m51s Details Deprecated doesn't mean it's "gone", it just means it is "going away" in the next major version of the package.	2025-11-13 13:16:02 -08:00
Francisco Arceo	eb3f9ac278	feat: allow returning embeddings and metadata from `/vector_stores/` methods; disallow changing Provider ID (#4046 ) # What does this PR do? - Updates `/vector_stores/{vector_store_id}/files/{file_id}/content` to allow returning `embeddings` and `metadata` using the `extra_query` - Updates the UI accordingly to display them. - Update UI to support CRUD operations in the Vector Stores section and adds a new modal exposing the functionality. - Updates Vector Store update to fail if a user tries to update Provider ID (which doesn't make sense to allow) ```python In [1]: client.vector_stores.files.content( vector_store_id=vector_store.id, file_id=file.id, extra_query={"include_embeddings": True, "include_metadata": True} ) Out [1]: FileContentResponse(attributes={}, content=[Content(text='This is a test document to check if embeddings are generated properly.\n', type='text', embedding=[0.33760684728622437, ...,], chunk_metadata={'chunk_id': '62a63ae0-c202-f060-1b86-0a688995b8d3', 'document_id': 'file-27291dbc679642ac94ffac6d2810c339', 'source': None, 'created_timestamp': 1762053437, 'updated_timestamp': 1762053437, 'chunk_window': '0-13', 'chunk_tokenizer': 'DEFAULT_TIKTOKEN_TOKENIZER', 'chunk_embedding_model': 'sentence-transformers/nomic -ai/nomic-embed-text-v1.5', 'chunk_embedding_dimension': 768, 'content_token_count': 13, 'metadata_token_count': 9}, metadata={'filename': 'test-embedding.txt', 'chunk_id': '62a63ae0-c202-f060-1b86-0a688995b8d3', 'document_id': 'file-27291dbc679642ac94ffac6d2810c339', 'token_count': 13, 'metadata_token_count': 9})], file_id='file-27291dbc679642ac94ffac6d2810c339', filename='test-embedding.txt') ``` Screenshots of UI are displayed below: ### List Vector Store with Added "Create New Vector Store" <img width="1912" height="491" alt="Screenshot 2025-11-06 at 10 47 25 PM" src="https://github.com/user-attachments/assets/a3a3ddd9-758d-4005-ac9c-5047f03916f3" /> ### Create New Vector Store <img width="1918" height="1048" alt="Screenshot 2025-11-06 at 10 47 49 PM" src="https://github.com/user-attachments/assets/b4dc0d31-696f-4e68-b109-27915090f158" /> ### Edit Vector Store <img width="1916" height="1355" alt="Screenshot 2025-11-06 at 10 48 32 PM" src="https://github.com/user-attachments/assets/ec879c63-4cf7-489f-bb1e-57ccc7931414" /> ### Vector Store Files Contents page (with Embeddings) <img width="1914" height="849" alt="Screenshot 2025-11-06 at 11 54 32 PM" src="https://github.com/user-attachments/assets/3095520d-0e90-41f7-83bd-652f6c3fbf27" /> ### Vector Store Files Contents Details page (with Embeddings) <img width="1916" height="1221" alt="Screenshot 2025-11-06 at 11 55 00 PM" src="https://github.com/user-attachments/assets/e71dbdc5-5b49-472b-a43a-5785f58d196c" /> <!-- If resolving an issue, uncomment and update the line below --> <!-- Closes #[issue-number] --> ## Test Plan Tests added for Middleware extension and Provider failures. --------- Signed-off-by: Francisco Javier Arceo <farceo@redhat.com>	2025-11-12 09:59:48 -08:00
Sam El-Borai	63137f9af1	chore(stainless): add config for file header (#4126 ) # What does this PR do? <!-- Provide a short summary of what this PR does and why. Link to relevant issues if applicable. --> <!-- If resolving an issue, uncomment and update the line below --> <!-- Closes #[issue-number] --> This PR adds Stainless config to specify the Meta copyright file header for generated files. Doing it via config instead of custom code will reduce the probability of git conflict. ## Test Plan <!-- Describe the tests you ran to verify your changes with result summaries. Provide clear instructions so the plan can be easily re-executed. --> - review preview builds	2025-11-12 11:39:21 -05:00
Nathan Weinberg	97ccfb5e62	refactor: inspect routes now shows all non-deprecated APIs (#4116 ) Some checks failed SqlStore Integration Tests / test-postgres (3.12) (push) Failing after 1s Details SqlStore Integration Tests / test-postgres (3.13) (push) Failing after 0s Details Integration Auth Tests / test-matrix (oauth2_token) (push) Failing after 1s Details Pre-commit / pre-commit (push) Failing after 1s Details Integration Tests (Replay) / generate-matrix (push) Successful in 2s Details Vector IO Integration Tests / test-matrix (push) Failing after 4s Details Test Llama Stack Build / generate-matrix (push) Successful in 4s Details Test Llama Stack Build / build-single-provider (push) Failing after 3s Details Test External Providers Installed via Module / test-external-providers-from-module (venv) (push) Has been skipped Details Test Llama Stack Build / build-ubi9-container-distribution (push) Failing after 3s Details Test Llama Stack Build / build-custom-container-distribution (push) Failing after 4s Details Python Package Build Test / build (3.12) (push) Failing after 2s Details Python Package Build Test / build (3.13) (push) Failing after 1s Details Test llama stack list-deps / generate-matrix (push) Successful in 4s Details Test llama stack list-deps / list-deps-from-config (push) Failing after 4s Details API Conformance Tests / check-schema-compatibility (push) Successful in 10s Details Test llama stack list-deps / show-single-provider (push) Failing after 5s Details Test External API and Providers / test-external (venv) (push) Failing after 4s Details Unit Tests / unit-tests (3.12) (push) Failing after 4s Details Unit Tests / unit-tests (3.13) (push) Failing after 4s Details Integration Tests (Replay) / Integration Tests (, , , client=, ) (push) Failing after 4s Details Test llama stack list-deps / list-deps (push) Failing after 3s Details Test Llama Stack Build / build (push) Failing after 21s Details UI Tests / ui-tests (22) (push) Successful in 46s Details # What does this PR do? the inspect API lacked any mechanism to get all non-deprecated APIs (v1, v1alpha, v1beta) change default to this behavior 'v1' filter can be used for user' wanting a list of stable APIs ## Test Plan 1. pull the PR 2. launch a LLS server 3. run `curl http://beanlab3.bss.redhat.com:8321/v1/inspect/routes` 4. note there are APIs for `v1`, `v1alpha`, and `v1beta` but no deprecated APIs Signed-off-by: Nathan Weinberg <nweinber@redhat.com>	2025-11-10 15:57:17 -08:00
Shabana Baig	433438cfc0	feat: Implement the 'max_tool_calls' parameter for the Responses API (#4062 ) # Problem Responses API uses max_tool_calls parameter to limit the number of tool calls that can be generated in a response. Currently, LLS implementation of the Responses API does not support this parameter. # What does this PR do? This pull request adds the max_tool_calls field to the response object definition and updates the inline provider. it also ensures that: - the total number of calls to built-in and mcp tools do not exceed max_tool_calls - an error is thrown if max_tool_calls < 1 (behavior seen with the OpenAI Responses API, but we can change this if needed) Closes #[3563](https://github.com/llamastack/llama-stack/issues/3563) ## Test Plan - Tested manually for change in model response w.r.t supplied max_tool_calls field. - Added integration tests to test invalid max_tool_calls parameter. - Added integration tests to check max_tool_calls parameter with built-in and function tools. - Added integration tests to check max_tool_calls parameter in the returned response object. - Recorded OpenAI Responses API behavior using a sample script: https://github.com/s-akhtar-baig/llama-stack-examples/blob/main/responses/src/max_tool_calls.py Co-authored-by: Ashwin Bharambe <ashwin.bharambe@gmail.com>	2025-11-10 13:21:27 -08:00
Ashwin Bharambe	fadf17daf3	feat(api)!: deprecate register/unregister resource APIs (#4099 ) Some checks failed SqlStore Integration Tests / test-postgres (3.13) (push) Failing after 0s Details Integration Auth Tests / test-matrix (oauth2_token) (push) Failing after 1s Details Python Package Build Test / build (3.12) (push) Failing after 1s Details Test External Providers Installed via Module / test-external-providers-from-module (venv) (push) Has been skipped Details Python Package Build Test / build (3.13) (push) Failing after 1s Details Integration Tests (Replay) / generate-matrix (push) Successful in 3s Details Pre-commit / pre-commit (push) Failing after 3s Details SqlStore Integration Tests / test-postgres (3.12) (push) Failing after 6s Details Vector IO Integration Tests / test-matrix (push) Failing after 4s Details API Conformance Tests / check-schema-compatibility (push) Successful in 8s Details Integration Tests (Replay) / Integration Tests (, , , client=, ) (push) Failing after 4s Details Unit Tests / unit-tests (3.12) (push) Failing after 3s Details Test External API and Providers / test-external (venv) (push) Failing after 5s Details Unit Tests / unit-tests (3.13) (push) Failing after 3s Details UI Tests / ui-tests (22) (push) Successful in 1m10s Details Mark all register_* / unregister_* APIs as deprecated across models, shields, tool groups, datasets, benchmarks, and scoring functions. This is the first step toward moving resource mutations to an `/admin` namespace as outlined in https://github.com/llamastack/llama-stack/issues/3809#issuecomment-3492931585. The deprecation flag will be reflected in the OpenAPI schema to warn API users that these endpoints are being phased out. Next step will be implementing the `/admin` route namespace for these resource management operations. - `register_model` / `unregister_model` - `register_shield` / `unregister_shield` - `register_tool_group` / `unregister_toolgroup` - `register_dataset` / `unregister_dataset` - `register_benchmark` / `unregister_benchmark` - `register_scoring_function` / `unregister_scoring_function`	2025-11-10 10:36:33 -08:00
ehhuang	d4ecbfd092	fix(vector store)!: fix file content API (#4105 ) # What does this PR do? - changed to match https://app.stainless.com/api/spec/documented/openai/openapi.documented.yml ## Test Plan updated test CI	2025-11-10 10:16:35 -08:00
Sam El-Borai	8f4c431370	chore(ci): setup automated stainless builds (#3557 ) Some checks failed SqlStore Integration Tests / test-postgres (3.13) (push) Failing after 0s Details Integration Auth Tests / test-matrix (oauth2_token) (push) Failing after 1s Details SqlStore Integration Tests / test-postgres (3.12) (push) Failing after 0s Details Python Package Build Test / build (3.12) (push) Failing after 1s Details Test External Providers Installed via Module / test-external-providers-from-module (venv) (push) Has been skipped Details Vector IO Integration Tests / test-matrix (push) Failing after 4s Details Integration Tests (Replay) / generate-matrix (push) Successful in 6s Details Unit Tests / unit-tests (3.13) (push) Failing after 4s Details Python Package Build Test / build (3.13) (push) Failing after 9s Details API Conformance Tests / check-schema-compatibility (push) Successful in 15s Details Unit Tests / unit-tests (3.12) (push) Failing after 13s Details Pre-commit / pre-commit (push) Failing after 21s Details Test External API and Providers / test-external (venv) (push) Failing after 22s Details Integration Tests (Replay) / Integration Tests (, , , client=, ) (push) Failing after 18s Details UI Tests / ui-tests (22) (push) Successful in 1m7s Details # What does this PR do? <!-- Provide a short summary of what this PR does and why. Link to relevant issues if applicable. --> This pull request adds a new workflow that does 2 things: 1. generate [SDK preview builds](https://www.stainless.com/docs/guides/automate-updates#set-up-automatic-preview-builds) whenever the OpenAPI spec file is modified in a PR 2. on PR merge, generate SDK builds that will be pushed to the different SDK repos (i.e start the release process) > [!NOTE] > No repo secret `STAINLESS_API_KEY` is needed, the authentication is done automatically via GitHub OIDC. <!-- If resolving an issue, uncomment and update the line below --> <!-- Closes #[issue-number] --> ## Test Plan <!-- Describe the tests you ran to verify your changes with result summaries. Provide clear instructions so the plan can be easily re-executed. --> I tested in my fork: https://github.com/stainless-api/llama-stack/pull/3	2025-11-07 12:15:26 -08:00
Aakanksha Duggal	b83184f7ef	feat(responses)!: Add web_search_2025_08_26 to the WebSearchToolTypes (#4103 ) # What does this PR do? Resolves #4102 1. Added `web_search_2025_08_26` to the `WebSearchToolTypes` list and the `OpenAIResponseInputToolWebSearch.type` Literal union 2. No changes needed to tool execution logic - all `web_search` types map to the same underlying tool 3. Backward compatibility is maintained - existing `web_search`, `web_search_preview`, and `web_search_preview_2025_03_11` types continue to work 4. Added an integration test case using {"type": "web_search_2025_08_26"} to verify it works correctly 5. Updated `docs/docs/providers/openai_responses_limitations.mdx` to reflect that `web_search_2025_08_26` is now supported. 6. Removed incorrect references to `MOD1/MOD2/MOD3` (which don't exist in the codebase) <!-- If resolving an issue, uncomment and update the line below --> <!-- Closes #[issue-number] --> ## Test Plan <!-- Describe the tests you ran to verify your changes with result summaries. Provide clear instructions so the plan can be easily re-executed. --> --------- Signed-off-by: Aakanksha Duggal <aduggal@redhat.com> Co-authored-by: github-actions[bot] <github-actions[bot]@users.noreply.github.com>	2025-11-07 10:01:12 -08:00
Sébastien Han	939a2db58f	chore: update stainless config (#4096 ) # What does this PR do? Removed in https://github.com/llamastack/llama-stack/pull/4067 Signed-off-by: Sébastien Han <seb@redhat.com>	2025-11-06 15:58:13 -05:00
ehhuang	9d5c34af27	fix!: BREAKING CHANGE: vector_store: search API response fix (#4080 ) # What does this PR do? - search_query in the vector store search API should be a list, according to https://github.com/openai/openai-openapi ## Test Plan modified tests --- [//]: # (BEGIN SAPLING FOOTER) Stack created with [Sapling](https://sapling-scm.com). Best reviewed with [ReviewStack](https://reviewstack.dev/llamastack/llama-stack/pull/4080). * #4086 * __->__ #4080	2025-11-05 15:01:48 -08:00
Ashwin Bharambe	392e01dc79	chore: add stainless config Some checks failed SqlStore Integration Tests / test-postgres (3.12) (push) Failing after 0s Details SqlStore Integration Tests / test-postgres (3.13) (push) Failing after 1s Details Integration Auth Tests / test-matrix (oauth2_token) (push) Failing after 1s Details Python Package Build Test / build (3.12) (push) Failing after 2s Details Pre-commit / pre-commit (push) Failing after 2s Details Test External Providers Installed via Module / test-external-providers-from-module (venv) (push) Has been skipped Details Python Package Build Test / build (3.13) (push) Failing after 2s Details Integration Tests (Replay) / Integration Tests (, , , client=, ) (push) Failing after 5s Details Vector IO Integration Tests / test-matrix (push) Failing after 6s Details Test External API and Providers / test-external (venv) (push) Failing after 4s Details Unit Tests / unit-tests (3.12) (push) Failing after 5s Details API Conformance Tests / check-schema-compatibility (push) Successful in 13s Details Unit Tests / unit-tests (3.13) (push) Failing after 7s Details UI Tests / ui-tests (22) (push) Successful in 1m13s Details name it to indicate it is not yet source of truth to avoid confusion	2025-11-04 15:44:07 -08:00
Ashwin Bharambe	0c49a53c97	chore(api)!: remove tool_runtime.rag_tool from the API surface (#4067 ) RAG aka file search is implemented via the Responses API by specifying the file-search tool. The backend implementation remains unchanged. This PR merely removes the directly exposed API surface which allowed users to directly perform searches from the client. This facility is now available via the `client.vector_store.search()` OpenAI compatible API.	2025-11-04 14:50:54 -08:00
Ashwin Bharambe	a8a8aa56c0	chore!: remove the agents (sessions and turns) API (#4055 ) Some checks failed SqlStore Integration Tests / test-postgres (3.12) (push) Failing after 0s Details Integration Auth Tests / test-matrix (oauth2_token) (push) Failing after 1s Details Test External Providers Installed via Module / test-external-providers-from-module (venv) (push) Has been skipped Details Pre-commit / pre-commit (push) Failing after 3s Details Python Package Build Test / build (3.12) (push) Failing after 2s Details Python Package Build Test / build (3.13) (push) Failing after 2s Details Vector IO Integration Tests / test-matrix (push) Failing after 4s Details Integration Tests (Replay) / Integration Tests (, , , client=, ) (push) Failing after 5s Details Test External API and Providers / test-external (venv) (push) Failing after 5s Details SqlStore Integration Tests / test-postgres (3.13) (push) Failing after 9s Details Unit Tests / unit-tests (3.13) (push) Failing after 5s Details Unit Tests / unit-tests (3.12) (push) Failing after 6s Details API Conformance Tests / check-schema-compatibility (push) Successful in 13s Details UI Tests / ui-tests (22) (push) Successful in 1m10s Details - Removes the deprecated agents (sessions and turns) API that was marked alpha in 0.3.0 - Cleans up unused imports and orphaned types after the API removal - Removes `SessionNotFoundError` and `AgentTurnInputType` which are no longer needed The agents API is completely superseded by the Responses + Conversations APIs, and the client SDK Agent class already uses those implementations. Corresponding client-side PR: https://github.com/llamastack/llama-stack-client-python/pull/295	2025-11-04 09:38:39 -08:00
Ashwin Bharambe	053fc0ac39	chore!: remove all deprecated routes (including /openai/v1/ ones) (#4054 ) Some checks failed SqlStore Integration Tests / test-postgres (3.12) (push) Failing after 0s Details SqlStore Integration Tests / test-postgres (3.13) (push) Failing after 0s Details Integration Auth Tests / test-matrix (oauth2_token) (push) Failing after 1s Details Python Package Build Test / build (3.12) (push) Failing after 2s Details Python Package Build Test / build (3.13) (push) Failing after 2s Details Test External Providers Installed via Module / test-external-providers-from-module (venv) (push) Has been skipped Details Pre-commit / pre-commit (push) Failing after 2s Details Integration Tests (Replay) / Integration Tests (, , , client=, ) (push) Failing after 4s Details Vector IO Integration Tests / test-matrix (push) Failing after 6s Details Test External API and Providers / test-external (venv) (push) Failing after 4s Details Unit Tests / unit-tests (3.12) (push) Failing after 5s Details Unit Tests / unit-tests (3.13) (push) Failing after 5s Details API Conformance Tests / check-schema-compatibility (push) Successful in 13s Details UI Tests / ui-tests (22) (push) Successful in 1m13s Details This PR removes all routes which we had marked deprecated for the 0.3.0 release. This includes: - all the `/v1/openai/v1/` routes (the corresponding /v1 routes still exist of course) - the /agents API (which is superseded completely by Responses + Conversations) - several alpha routes which had a "v1" route to aide transitioning to "v1alpha" This is the corresponding client-python change: https://github.com/llamastack/llama-stack-client-python/pull/294	2025-11-03 19:00:59 -08:00
Sébastien Han	4a5ef65286	chore!: remove SDG API (#4035 ) # What does this PR do? This API hasn't received any traction and close to zero interest from the community. Let's revisit in the future if things change. Signed-off-by: Sébastien Han <seb@redhat.com> Co-authored-by: Ashwin Bharambe <ashwin.bharambe@gmail.com>	2025-11-03 16:12:06 -08:00

1 2

66 commits