mirror of https://github.com/meta-llama/llama-stack.git synced 2025-08-13 05:17:26 +00:00

Composable building blocks to build Llama Apps https://llama-stack.readthedocs.io

Find a file

Sébastien Han 0942329ec6 docs: move sections from README to docs To maintain a clean and uncluttered README for the project’s landing page, the API providers and distributions are now included in the documentation. Relevant sections have been updated as needed. Signed-off-by: Sébastien Han <seb@redhat.com>		2025-02-25 17:33:03 +01:00
.github	ci: improve GitHub Actions workflow for website builds (#1151 )	2025-02-20 21:37:37 -08:00
distributions	test: add a ci-tests distro template for running e2e tests (#1237 )	2025-02-24 14:43:21 -08:00
docs	docs: move sections from README to docs	2025-02-25 17:33:03 +01:00
llama_stack	refactor: combine start scripts for each env (#1139 )	2025-02-24 16:53:31 -08:00
rfcs	docs: Fix url to the llama-stack-spec yaml/html files (#1081 )	2025-02-13 12:39:26 -08:00
tests/client-sdk	test fix for sometimes tools get called more than once	2025-02-24 13:16:40 -08:00
.gitignore	github: ignore non-hidden python virtual environments (#939 )	2025-02-03 11:53:05 -08:00
.gitmodules	`impls` -> `inline`, `adapters` -> `remote` (#381 )	2024-11-06 14:54:05 -08:00
.pre-commit-config.yaml	fix: update virtualenv building so llamastack- prefix is not added, make notebook experience easier (#1225 )	2025-02-23 16:57:11 -08:00
.python-version	build: hint on Python version for uv venv (#1172 )	2025-02-25 10:37:45 -05:00
.readthedocs.yaml	first version of readthedocs (#278 )	2024-10-22 10:15:58 +05:30
CODE_OF_CONDUCT.md	Initial commit	2024-07-23 08:32:33 -07:00
CONTRIBUTING.md	docs: Add missing uv command and clarify website rebuild (#1199 )	2025-02-21 11:29:32 -05:00
LICENSE	Update LICENSE (#47 )	2024-08-29 07:39:50 -07:00
MANIFEST.in	Add test jsons to MANIFEST for now	2025-02-21 16:25:51 -08:00
pyproject.toml	Bump version to 0.1.4	2025-02-24 15:59:26 -08:00
README.md	docs: move sections from README to docs	2025-02-25 17:33:03 +01:00
requirements.txt	fix: pre-commit updates (#1243 )	2025-02-24 17:20:29 -08:00
SECURITY.md	Create SECURITY.md	2024-10-08 13:30:40 -04:00
uv.lock	fix: pre-commit updates (#1243 )	2025-02-24 17:20:29 -08:00

README.md

Llama Stack

Quick Start | Documentation | Colab Notebook

Llama Stack standardizes the core building blocks that simplify AI application development. It codifies best practices across the Llama ecosystem. More specifically, it provides

Unified API layer for Inference, RAG, Agents, Tools, Safety, Evals, and Telemetry.
Plugin architecture to support the rich ecosystem of different API implementations in various environments, including local development, on-premises, cloud, and mobile.
Prepackaged verified distributions which offer a one-stop solution for developers to get started quickly and reliably in any environment.
Multiple developer interfaces like CLI and SDKs for Python, Typescript, iOS, and Android.
Standalone applications as examples for how to build production-grade AI applications with Llama Stack.

Llama Stack Benefits

Flexible Options: Developers can choose their preferred infrastructure without changing APIs and enjoy flexible deployment choices.
Consistent Experience: With its unified APIs, Llama Stack makes it easier to build, test, and deploy AI applications with consistent application behavior.
Robust Ecosystem: Llama Stack is already integrated with distribution partners (cloud providers, hardware vendors, and AI-focused companies) that offer tailored infrastructure, software, and services for deploying Llama models.

By reducing friction and complexity, Llama Stack empowers developers to focus on what they do best: building transformative generative AI applications.

Installation

You have two ways to install this repository:

Install as a package: You can install the repository directly from PyPI by running the following command:
```
pip install llama-stack
```
Install from source: If you prefer to install from the source code, we recommend using uv. Then, run the following commands:
```
 git clone git@github.com:meta-llama/llama-stack.git
 cd llama-stack

 uv sync
 uv pip install -e .
```

Documentation

Please checkout our Documentation page for more details.

CLI references
- llama (server-side) CLI Reference: Guide for using the llama CLI to work with Llama models (download, study prompts), and building/starting a Llama Stack distribution.
- llama (client-side) CLI Reference: Guide for using the llama-stack-client CLI, which allows you to query information about the distribution.
Getting Started
- Quick guide to start a Llama Stack server.
- Jupyter notebook to walk-through how to use simple text and vision inference llama_stack_client APIs
- The complete Llama Stack lesson Colab notebook of the new Llama 3.2 course on Deeplearning.ai.
- A Zero-to-Hero Guide that guide you through all the key components of llama stack with code samples.
Contributing
- Adding a new API Provider to walk-through how to add a new API provider.

Llama Stack Client SDKs

Language	Client SDK	Package
Python	llama-stack-client-python
Swift	llama-stack-client-swift
Typescript	llama-stack-client-typescript
Kotlin	llama-stack-client-kotlin

Check out our client SDKs for connecting to a Llama Stack server in your preferred language, you can choose from python, typescript, swift, and kotlin programming languages to quickly build your applications.

You can find more example scripts with client SDKs to talk with the Llama Stack server in our llama-stack-apps repo.