forked from phoenix-oss/llama-stack-mirror
		
	fix!: update eval-tasks -> benchmarks (#1032)
# What does this PR do? - Update `/eval-tasks` to `/benchmarks` - ⚠️ Remove differentiation between `app` v.s. `benchmark` eval task config. Now we only have `BenchmarkConfig`. The overloaded `benchmark` is confusing and do not add any value. Backward compatibility is being kept as the "type" is not being used anywhere. [//]: # (If resolving an issue, uncomment and update the line below) [//]: # (Closes #[issue-number]) ## Test Plan - This change is backward compatible - Run notebook test with ``` pytest -v -s --nbval-lax ./docs/getting_started.ipynb pytest -v -s --nbval-lax ./docs/notebooks/Llama_Stack_Benchmark_Evals.ipynb ``` <img width="846" alt="image" src="https://github.com/user-attachments/assets/d2fc06a7-593a-444f-bc1f-10ab9b0c843d" /> [//]: # (## Documentation) [//]: # (- [ ] Added a Changelog entry if the change is significant) --------- Signed-off-by: Ihar Hrachyshka <ihar.hrachyshka@gmail.com> Signed-off-by: Ben Browning <bbrownin@redhat.com> Signed-off-by: Sébastien Han <seb@redhat.com> Signed-off-by: reidliu <reid201711@gmail.com> Co-authored-by: Ihar Hrachyshka <ihar.hrachyshka@gmail.com> Co-authored-by: Ben Browning <ben324@gmail.com> Co-authored-by: Sébastien Han <seb@redhat.com> Co-authored-by: Reid <61492567+reidliu41@users.noreply.github.com> Co-authored-by: reidliu <reid201711@gmail.com> Co-authored-by: Yuan Tang <terrytangyuan@gmail.com>
This commit is contained in:
		
							parent
							
								
									225dd38e5c
								
							
						
					
					
						commit
						8b655e3cd2
					
				
					 60 changed files with 2622 additions and 1910 deletions
				
			
		|  | @ -64,7 +64,7 @@ Interactive pages for users to play with and explore Llama Stack API capabilitie | |||
|     ``` | ||||
| 
 | ||||
|     ```bash | ||||
|     $ llama-stack-client eval_tasks register \ | ||||
|     $ llama-stack-client benchmarks register \ | ||||
|     --eval-task-id meta-reference-mmlu \ | ||||
|     --provider-id meta-reference \ | ||||
|     --dataset-id mmlu \ | ||||
|  | @ -86,7 +86,7 @@ Interactive pages for users to play with and explore Llama Stack API capabilitie | |||
|   - Under the hood, it uses Llama Stack's `/providers` API to get information about the providers. | ||||
| 
 | ||||
| - **API Resources**: Inspect Llama Stack API resources | ||||
|   - This page allows you to inspect Llama Stack API resources (`models`, `datasets`, `memory_banks`, `eval_tasks`, `shields`). | ||||
|   - This page allows you to inspect Llama Stack API resources (`models`, `datasets`, `memory_banks`, `benchmarks`, `shields`). | ||||
|   - Under the hood, it uses Llama Stack's `/<resources>/list` API to get information about each resources. | ||||
|   - Please visit [Core Concepts](https://llama-stack.readthedocs.io/en/latest/concepts/index.html) for more details about the resources. | ||||
| 
 | ||||
|  |  | |||
		Loading…
	
	Add table
		Add a link
		
	
		Reference in a new issue