llama-stack-mirror

mirror of https://github.com/meta-llama/llama-stack.git synced 2025-12-04 02:03:44 +00:00

History

Derek Higgins 5bbca56cfc fix: Make SentenceTransformer embedding operations non-blocking (#3335 ) - Wrap model loading with asyncio.to_thread() to prevent blocking during model download/initialization - Wrap encoding operations with asyncio.to_thread() to run in background thread - Convert _load_sentence_transformer_model() to async method This ensures the async event loop remains responsive during embedding operations. Closes: #3332 Signed-off-by: Derek Higgins <derekh@redhat.com> Co-authored-by: Francisco Arceo <arceofrancisco@gmail.com>		2025-09-04 13:58:41 -04:00
..
bedrock	feat: drop python 3.10 support (#2469 )	2025-06-19 12:07:14 +05:30
common	chore(rename): move llama_stack.distribution to llama_stack.core (#2975 )	2025-07-30 23:30:53 -07:00
datasetio	chore(misc): make tests and starter faster (#3042 )	2025-08-05 14:55:05 -07:00
inference	fix: Make SentenceTransformer embedding operations non-blocking (#3335 )	2025-09-04 13:58:41 -04:00
kvstore	refactor(logging): rename llama_stack logger categories (#3065 )	2025-08-21 17:31:04 -07:00
memory	chore(migrate apis): move VectorDBWithIndex from embeddings to openai_embeddings (#3294 )	2025-08-31 14:48:35 -07:00
responses	chore(rename): move llama_stack.distribution to llama_stack.core (#2975 )	2025-07-30 23:30:53 -07:00
scoring	chore: enable pyupgrade fixes (#1806 )	2025-05-01 14:23:50 -07:00
sqlstore	chore(dev): add inequality support to sqlstore where clause (#3272 )	2025-08-28 14:49:36 -07:00
telemetry	feat: implement query_metrics (#3074 )	2025-08-22 14:19:24 -07:00
tools	chore(rename): move llama_stack.distribution to llama_stack.core (#2975 )	2025-07-30 23:30:53 -07:00
vector_io	feat: implement keyword, vector and hybrid search inside vector stores for PGVector provider (#3064 )	2025-08-29 16:30:12 +02:00
__init__.py	API Updates (#73 )	2024-09-17 19:51:35 -07:00
pagination.py	chore(refact): move paginate_records fn outside of datasetio (#2137 )	2025-05-12 10:56:14 -07:00
scheduler.py	refactor(logging): rename llama_stack logger categories (#3065 )	2025-08-21 17:31:04 -07:00