Added Ollama as an inference impl (#20)

* fix non-streaming api in inference server * unit test for inline inference * Added non-streaming ollama inference impl * add streaming support for ollama inference with tests * addressing comments --------- Co-authored-by: Hardik Shah <hjshah@fb.com>
2025-10-04 04:04:14 +00:00 · 2024-07-31 22:08:37 -07:00 · 2024-07-31 22:08:37 -07:00 · 156bfa0e15
commit 156bfa0e15
parent c253c1c9ad
9 changed files with 921 additions and 33 deletions
--- a/requirements.txt
+++ b/requirements.txt
@ -13,6 +13,7 @@ hydra-zen
 json-strong-typing
 llama-models
 matplotlib
+ollama
 omegaconf
 pandas
 Pillow