llama-stack-mirror

mirror of https://github.com/meta-llama/llama-stack.git synced 2025-12-04 10:10:36 +00:00

History

Matthew Farrellee 53b15725b6 chore(apis): unpublish deprecated /v1/inference apis (#3297 ) # What does this PR do? unpublish (make unavailable to users) the following apis - - `/v1/inference/completion`, replaced by `/v1/openai/v1/completions` - `/v1/inference/chat-completion`, replaced by `/v1/openai/v1/chat/completions` - `/v1/inference/embeddings`, replaced by `/v1/openai/v1/embeddings` - `/v1/inference/batch-completion`, replaced by `/v1/openai/v1/batches` - `/v1/inference/batch-chat-completion`, replaced by `/v1/openai/v1/batches` note: the implementations are still available for internal use, e.g. agents uses chat-completion.		2025-09-27 11:20:06 -07:00
..
__init__.py	chore: enable pyupgrade fixes (#1806 )	2025-05-01 14:23:50 -07:00
embedding_mixin.py	fix: Make SentenceTransformer embedding operations non-blocking (#3335 )	2025-09-04 13:58:41 -04:00
inference_store.py	chore: simplify authorized sqlstore (#3496 )	2025-09-19 16:13:56 -07:00
litellm_openai_mixin.py	feat: add static embedding metadata to dynamic model listings for providers using OpenAIMixin (#3547 )	2025-09-25 17:17:00 -04:00
model_registry.py	chore: prune mypy exclude list (#3561 )	2025-09-26 11:44:43 -04:00
openai_compat.py	refactor(logging): rename llama_stack logger categories (#3065 )	2025-08-21 17:31:04 -07:00
openai_mixin.py	feat(internal): add image_url download feature to OpenAIMixin (#3516 )	2025-09-26 17:32:16 -04:00
prompt_adapter.py	chore(apis): unpublish deprecated /v1/inference apis (#3297 )	2025-09-27 11:20:06 -07:00