llama-stack-mirror

mirror of https://github.com/meta-llama/llama-stack.git synced 2025-12-17 16:32:38 +00:00

History

Matthew Farrellee 1f7e87c647 feat: update Cerebras inference provider to support dynamic model listing - update Cerebras to use OpenAIMixin - enable openai completions tests - enable openai chat completions tests - disable with n > 1 tests - add recording for --setup cerebras --subdirs inference --pattern openai test with: `./scripts/integration-tests.sh --stack-config server:ci-tests --setup cerebras --subdirs inference --pattern openai`		2025-09-18 06:54:05 -04:00
..
anthropic	chore: update the anthropic inference impl to use openai-python for openai-compat functions (#3366 )	2025-09-07 14:00:42 -07:00
azure	feat: add Azure OpenAI inference provider support (#3396 )	2025-09-11 13:48:38 +02:00
bedrock	fix: AWS Bedrock inference profile ID conversion for region-specific endpoints (#3386 )	2025-09-11 11:41:53 +02:00
cerebras	feat: update Cerebras inference provider to support dynamic model listing	2025-09-18 06:54:05 -04:00
databricks	feat(starter)!: simplify starter distro; litellm model registry changes (#2916 )	2025-07-25 15:02:04 -07:00
fireworks	refactor(logging): rename llama_stack logger categories (#3065 )	2025-08-21 17:31:04 -07:00
gemini	chore: update the gemini inference impl to use openai-python for openai-compat functions (#3351 )	2025-09-06 12:22:20 -07:00
groq	chore: update the groq inference impl to use openai-python for openai-compat functions (#3348 )	2025-09-06 15:36:27 -07:00
llama_openai_compat	chore: indicate to mypy that InferenceProvider.rerank is concrete (#3238 )	2025-08-22 12:02:13 -07:00
nvidia	docs: add VLM NIM example (#3277 )	2025-08-29 16:23:52 -07:00
ollama	feat(tests): auto-merge all model list responses and unify recordings (#3320 )	2025-09-03 11:33:03 -07:00
openai	refactor(logging): rename llama_stack logger categories (#3065 )	2025-08-21 17:31:04 -07:00
passthrough	chore(rename): move llama_stack.distribution to llama_stack.core (#2975 )	2025-07-30 23:30:53 -07:00
runpod	ci: test safety with starter (#2628 )	2025-07-09 16:53:50 +02:00
sambanova	chore: update the sambanova inference impl to use openai-python for openai-compat functions (#3345 )	2025-09-06 12:25:13 -07:00
tgi	feat: add dynamic model registration support to TGI inference (#3417 )	2025-09-15 15:52:40 -04:00
together	feat: add embedding and dynamic model support to Together inference adapter (#3458 )	2025-09-16 11:53:41 -07:00
vertexai	ci: Re-enable pre-commit to fail (#3399 )	2025-09-10 10:00:46 -04:00
vllm	feat: Add dynamic authentication token forwarding support for vLLM (#3388 )	2025-09-18 11:13:55 +02:00
watsonx	chore: various watsonx fixes (#3428 )	2025-09-16 13:55:10 +02:00
__init__.py	`impls` -> `inline`, `adapters` -> `remote` (#381 )	2024-11-06 14:54:05 -08:00