forked from phoenix-oss/llama-stack-mirror
* rename evals related stuff * fix datasetio * fix scoring test * localfs -> LocalFS * refactor scoring * refactor scoring * remove 8b_correctness scoring_fn from tests * tests w/ eval params * scoring fn braintrust fixture * import |
||
|---|---|---|
| .. | ||
| agents | ||
| datasetio/localfs | ||
| eval/meta_reference | ||
| inference | ||
| ios/inference | ||
| memory | ||
| meta_reference | ||
| safety | ||
| scoring | ||
| __init__.py | ||