llama-stack-mirror

mirror of https://github.com/meta-llama/llama-stack.git synced 2025-10-04 04:04:14 +00:00

History

r3v5 1d0f0a0d8e refactor: switch to the new default nomic-embed-text-v1.5 embedding model in LS		2025-10-02 14:50:00 +01:00
..
agents	refactor(agents): migrate to OpenAI chat completions API (#3323 )	2025-10-02 06:50:32 -04:00
batches	feat(batches, completions): add /v1/completions support to /v1/batches (#3309 )	2025-09-05 11:59:57 -07:00
datasetio	chore(misc): make tests and starter faster (#3042 )	2025-08-05 14:55:05 -07:00
eval	feat: update eval runner to use openai endpoints (#3588 )	2025-09-29 13:13:53 -07:00
files/localfs	fix(expires_after): make sure multipart/form-data is properly parsed (#3612 )	2025-09-30 16:14:03 -04:00
inference	refactor: switch to the new default nomic-embed-text-v1.5 embedding model in LS	2025-10-02 14:50:00 +01:00
ios/inference	chore: removed executorch submodule (#1265 )	2025-02-25 21:57:21 -08:00
post_training	chore(pre-commit): add pre-commit hook to enforce llama_stack logger usage (#3061 )	2025-08-20 07:15:35 -04:00
safety	feat: use /v1/chat/completions for safety model inference (#3591 )	2025-09-30 11:01:44 -07:00
scoring	chore: use openai_chat_completion for llm as a judge scoring (#3635 )	2025-10-01 09:44:31 -04:00
telemetry	chore: Remove debug logging from telemetry adapter (#3643 )	2025-10-01 15:16:23 -07:00
tool_runtime	chore: Updating documentation, adding exception handling for Vector Stores in RAG Tool, more tests on migration, and migrate off of inference_api for context_retriever for RAG (#3367 )	2025-09-11 14:20:11 +02:00
vector_io	refactor: use generic WeightedInMemoryAggregator for hybrid search in SQLiteVecIndex (#3303 )	2025-09-02 10:38:35 -07:00
__init__.py	`impls` -> `inline`, `adapters` -> `remote` (#381 )	2024-11-06 14:54:05 -08:00