llama-stack-mirror

mirror of https://github.com/meta-llama/llama-stack.git synced 2025-12-28 07:41:59 +00:00

History

Ben Browning 8ede67b809 More work on file_search verification test This gets the file_search verification test working against ollama, fireworks, and api.openai.com. We don't have the entirety of the vector store API implemented in Llama Stack yet, so this still has a bit of a hack to swap between using only OpenAI-compatible APIs versus using the LlamaStackClient to insert content into our vector stores. Outside of actually inserting file contents, the rest of the test works the same and uses only the OpenAI client for all of these providers. How to run the tests: Ollama (sometimes flakes with small model): ``` ollama run llama3.2:3b INFERENCE_MODEL="meta-llama/Llama-3.2-3B-Instruct" \ llama stack run ./llama_stack/templates/ollama/run.yaml \ --image-type venv \ --env OLLAMA_URL="http://0.0.0.0:11434" pytest -sv \ 'tests/verifications/openai_api/test_responses.py::test_response_non_streaming_file_search' \ --base-url=http://localhost:8321/v1/openai/v1 \ --model meta-llama/Llama-3.2-3B-Instruct ``` Fireworks via Llama Stack: ``` llama stack run llama_stack/templates/fireworks/run.yaml pytest -sv \ 'tests/verifications/openai_api/test_responses.py::test_response_non_streaming_file_search' \ --base-url=http://localhost:8321/v1/openai/v1 \ --model meta-llama/Llama-3.3-70B-Instruct ``` OpenAI directly: ``` pytest -sv \ 'tests/verifications/openai_api/test_responses.py::test_response_non_streaming_file_search' \ --base-url=https://api.openai.com/v1 \ --model gpt-4o ``` Signed-off-by: Ben Browning <bbrownin@redhat.com>		2025-06-13 09:36:04 -04:00
..
k8s	chore(ui): use proxy server for backend API calls; simplified k8s deployment (#2350 )	2025-06-03 14:57:10 -07:00
ondevice_distro	docs: 0.2.2 doc updates (#1961 )	2025-04-15 13:26:17 -07:00
remote_hosted_distro	fix: replace all instances of --yaml-config with --config (#2196 )	2025-05-16 14:31:12 -07:00
self_hosted_distro	More work on file_search verification test	2025-06-13 09:36:04 -04:00
building_distro.md	refactor: remove container from list of run image types (#2178 )	2025-06-02 09:57:55 +02:00
configuration.md	feat(auth): allow token to be provided for use against jwks endpoint (#2394 )	2025-06-13 10:13:41 +02:00
importing_as_library.md	docs: update importing_as_library.md (#1863 )	2025-04-07 12:31:04 +02:00
index.md	docs: Updated documentation and Sphinx configuration (#1845 )	2025-03-31 13:08:05 -07:00
kubernetes_deployment.md	fix: replace all instances of --yaml-config with --config (#2196 )	2025-05-16 14:31:12 -07:00
list_of_distributions.md	docs: Updated documentation and Sphinx configuration (#1845 )	2025-03-31 13:08:05 -07:00
starting_llama_stack_server.md	docs: Update quickstart page to structure things a little more for the novices (#1873 )	2025-04-10 14:09:00 -07:00