llama-stack-mirror

mirror of https://github.com/meta-llama/llama-stack.git synced 2025-12-18 07:49:48 +00:00

History

Matthew Farrellee ae804ed5a8 feat: (re-)enable Databricks inference adapter Databricks inference adapter was broken, would not start, see #3486 - remove deprecated completion / chat_completion endpoints - enable dynamic model listing w/o refresh, listing is not async - use SecretStr instead of str for token - backward incompatible change: for consistency with databricks docs, env DATABRICKS_URL -> DATABRICKS_HOST and DATABRICKS_API_TOKEN -> DATABRICKS_TOKEN - databricks urls are custom per user/org, add special recorder handling for databricks urls - add integration test --setup databricks - enable chat completions tests - enable embeddings tests - disable n > 1 tests - disable embeddings base64 tests - disable embeddings dimensions tests note: reasoning models, e.g. gpt oss, fail because databricks has a custom, incompatible response format test with: ./scripts/integration-tests.sh --stack-config server:ci-tests --setup databricks --subdirs inference --pattern openai note: databricks needs to be manually added to the ci-tests distro for replay testing		2025-09-20 05:05:05 -04:00
..
anthropic	chore: update the anthropic inference impl to use openai-python for openai-compat functions (#3366 )	2025-09-07 14:00:42 -07:00
azure	feat: add Azure OpenAI inference provider support (#3396 )	2025-09-11 13:48:38 +02:00
bedrock	fix: AWS Bedrock inference profile ID conversion for region-specific endpoints (#3386 )	2025-09-11 11:41:53 +02:00
cerebras	feat(starter)!: simplify starter distro; litellm model registry changes (#2916 )	2025-07-25 15:02:04 -07:00
databricks	feat: (re-)enable Databricks inference adapter	2025-09-20 05:05:05 -04:00
fireworks	refactor(logging): rename llama_stack logger categories (#3065 )	2025-08-21 17:31:04 -07:00
gemini	chore: update the gemini inference impl to use openai-python for openai-compat functions (#3351 )	2025-09-06 12:22:20 -07:00
groq	chore: update the groq inference impl to use openai-python for openai-compat functions (#3348 )	2025-09-06 15:36:27 -07:00
llama_openai_compat	chore: indicate to mypy that InferenceProvider.rerank is concrete (#3238 )	2025-08-22 12:02:13 -07:00
nvidia	docs: add VLM NIM example (#3277 )	2025-08-29 16:23:52 -07:00
ollama	chore: update the ollama inference impl to use OpenAIMixin for openai-compat functions (#3395 )	2025-09-18 13:09:57 +02:00
openai	refactor(logging): rename llama_stack logger categories (#3065 )	2025-08-21 17:31:04 -07:00
passthrough	chore(rename): move llama_stack.distribution to llama_stack.core (#2975 )	2025-07-30 23:30:53 -07:00
runpod	ci: test safety with starter (#2628 )	2025-07-09 16:53:50 +02:00
sambanova	chore: update the sambanova inference impl to use openai-python for openai-compat functions (#3345 )	2025-09-06 12:25:13 -07:00
tgi	feat: add dynamic model registration support to TGI inference (#3417 )	2025-09-15 15:52:40 -04:00
together	feat: add embedding and dynamic model support to Together inference adapter (#3458 )	2025-09-16 11:53:41 -07:00
vertexai	ci: Re-enable pre-commit to fail (#3399 )	2025-09-10 10:00:46 -04:00
vllm	feat: Add dynamic authentication token forwarding support for vLLM (#3388 )	2025-09-18 11:13:55 +02:00
watsonx	chore: various watsonx fixes (#3428 )	2025-09-16 13:55:10 +02:00
__init__.py	`impls` -> `inline`, `adapters` -> `remote` (#381 )	2024-11-06 14:54:05 -08:00