llama-stack-mirror/llama_stack/providers/utils/inference
Ashwin Bharambe 540fc4d717
Fix Meta reference GPU implementation (#663)
By performing in-place mutations, we lost. Never in life do that.
2024-12-19 14:09:45 -08:00
..
__init__.py Added support for llama 3.3 model (#601) 2024-12-10 20:03:31 -08:00
embedding_mixin.py Update the "InterleavedTextMedia" type (#635) 2024-12-17 11:18:31 -08:00
model_registry.py add embedding model by default to distribution templates (#617) 2024-12-13 12:48:00 -08:00
openai_compat.py Update the "InterleavedTextMedia" type (#635) 2024-12-17 11:18:31 -08:00
prompt_adapter.py Fix Meta reference GPU implementation (#663) 2024-12-19 14:09:45 -08:00