llama-stack-mirror

mirror of https://github.com/meta-llama/llama-stack.git synced 2025-06-29 11:24:19 +00:00

History

Ben Browning fa34468308 feat: File search tool for Responses API This is an initial working prototype of wiring up the `file_search` builtin tool for the Responses API to our existing rag knowledge search tool. I stubbed in a new test (that uses a hardcoded url hybrid of the OpenAI and Llama Stack clients for now, only until we finish landing the vector store APIs and insertion support). Note that this is currently under tests/verification only because it sometimes flakes with tool calling of the small Llama-3.2-3B model we run in CI (and that I use as an example below). We'd want to make the test a bit more robust in some way if we moved this over to tests/integration and ran it in CI. ``` ollama run llama3.2:3b INFERENCE_MODEL="meta-llama/Llama-3.2-3B-Instruct" \ llama stack run ./llama_stack/templates/ollama/run.yaml \ --image-type venv \ --env OLLAMA_URL="http://0.0.0.0:11434" pytest -sv 'tests/verifications/openai_api/test_responses.py::test_response_non_streaming_file_search' \ --base-url=http://localhost:8321/v1/openai/v1 \ --model meta-llama/Llama-3.2-3B-Instruct ``` Signed-off-by: Ben Browning <bbrownin@redhat.com>		2025-06-13 09:36:04 -04:00
..
agents	feat: File search tool for Responses API	2025-06-13 09:36:04 -04:00
batch_inference	chore: more API validators (#2165 )	2025-05-15 11:22:51 -07:00
benchmarks	chore: more API validators (#2165 )	2025-05-15 11:22:51 -07:00
common	chore: removed unused class (#2268 )	2025-05-26 08:41:37 -07:00
datasetio	chore: more API validators (#2165 )	2025-05-15 11:22:51 -07:00
datasets	chore: more API validators (#2165 )	2025-05-15 11:22:51 -07:00
eval	chore: more API validators (#2165 )	2025-05-15 11:22:51 -07:00
files	feat: openai files api (#2321 )	2025-06-02 11:45:53 -07:00
inference	feat: New OpenAI compat embeddings API (#2314 )	2025-05-31 22:11:47 -07:00
inspect	chore: more API validators (#2165 )	2025-05-15 11:22:51 -07:00
models	chore: more API validators (#2165 )	2025-05-15 11:22:51 -07:00
post_training	chore: more API validators (#2165 )	2025-05-15 11:22:51 -07:00
providers	chore: more API validators (#2165 )	2025-05-15 11:22:51 -07:00
safety	chore: more API validators (#2165 )	2025-05-15 11:22:51 -07:00
scoring	chore: more API validators (#2165 )	2025-05-15 11:22:51 -07:00
scoring_functions	chore: more API validators (#2165 )	2025-05-15 11:22:51 -07:00
shields	chore: more API validators (#2165 )	2025-05-15 11:22:51 -07:00
synthetic_data_generation	chore: enable pyupgrade fixes (#1806 )	2025-05-01 14:23:50 -07:00
telemetry	chore: more API validators (#2165 )	2025-05-15 11:22:51 -07:00
tools	fix(tools): do not index tools, only index toolgroups (#2261 )	2025-05-25 13:27:52 -07:00
vector_dbs	chore: more API validators (#2165 )	2025-05-15 11:22:51 -07:00
vector_io	feat: update search for vector_stores (#2441 )	2025-06-12 15:34:22 -07:00
__init__.py	API Updates (#73 )	2024-09-17 19:51:35 -07:00
datatypes.py	chore: enable pyupgrade fixes (#1806 )	2025-05-01 14:23:50 -07:00
resource.py	chore: more mypy fixes (#2029 )	2025-05-06 09:52:31 -07:00
version.py	llama-stack version alpha -> v1	2025-01-15 05:58:09 -08:00