llama-stack-mirror

mirror of https://github.com/meta-llama/llama-stack.git synced 2025-12-03 09:53:45 +00:00

History

Jash Gulabrai 1a770cf8ac fix: Pass model parameter as config name to NeMo Customizer (#2218 ) # What does this PR do? When launching a fine-tuning job, an upcoming version of NeMo Customizer will expect the `config` name to be formatted as `namespace/name@version`. Here, `config` is a reference to a model + additional metadata. There could be multiple `config`s that reference the same base model. This PR updates NVIDIA's `supervised_fine_tune` to simply pass the `model` param as-is to NeMo Customizer. Currently, it expects a specific, allowlisted llama model (i.e. `meta/Llama3.1-8B-Instruct`) and converts it to the provider format (`meta/llama-3.1-8b-instruct`). [//]: # (If resolving an issue, uncomment and update the line below) [//]: # (Closes #[issue-number]) ## Test Plan From a notebook, I built an image with my changes: ``` !llama stack build --template nvidia --image-type venv from llama_stack.distribution.library_client import LlamaStackAsLibraryClient client = LlamaStackAsLibraryClient("nvidia") client.initialize() ``` And could successfully launch a job: ``` response = client.post_training.supervised_fine_tune( job_uuid="", model="meta/llama-3.2-1b-instruct@v1.0.0+A100", # Model passed as-is to Customimzer ... ) job_id = response.job_uuid print(f"Created job with ID: {job_id}") Output: Created job with ID: cust-Jm4oGmbwcvoufaLU4XkrRU ``` [//]: # (## Documentation) --------- Co-authored-by: Jash Gulabrai <jgulabrai@nvidia.com>		2025-05-20 09:51:39 -07:00
..
apis	feat: introduce APIs for retrieving chat completion requests (#2145 )	2025-05-18 21:43:19 -07:00
cli	fix: Pass external_config_dir to BuildConfig (#2190 )	2025-05-19 14:01:28 +02:00
distribution	feat: Propagate W3C trace context headers from clients (#2153 )	2025-05-19 18:56:54 -07:00
models	fix: llama4 tool use prompt fix (#2103 )	2025-05-06 22:18:31 -07:00
providers	fix: Pass model parameter as config name to NeMo Customizer (#2218 )	2025-05-20 09:51:39 -07:00
strong_typing	chore: enable pyupgrade fixes (#1806 )	2025-05-01 14:23:50 -07:00
templates	feat: add huggingface post_training impl (#2132 )	2025-05-16 14:41:28 -07:00
ui	chore: Updated readme (#2219 )	2025-05-20 17:06:20 +02:00
__init__.py	export LibraryClient	2024-12-13 12:08:00 -08:00
env.py	refactor(test): move tools, evals, datasetio, scoring and post training tests (#1401 )	2025-03-04 14:53:47 -08:00
log.py	chore: enable pyupgrade fixes (#1806 )	2025-05-01 14:23:50 -07:00
schema_utils.py	chore: enable pyupgrade fixes (#1806 )	2025-05-01 14:23:50 -07:00