mirror of
https://github.com/meta-llama/llama-stack.git
synced 2025-12-22 20:50:00 +00:00
- Add new Vertex AI remote inference provider with litellm integration - Support for Gemini models through Google Cloud Vertex AI platform - Uses Google Cloud Application Default Credentials (ADC) for authentication - Added VertexAI models: gemini-2.5-flash, gemini-2.5-pro, gemini-2.0-flash. - Updated provider registry to include vertexai provider - Add vertexai to INFERENCE_PROVIDER_IDS in starter distribution template - Update VertexAI provider to be conditionally included in starter template when VERTEX_AI_PROJECT env var is set - Added comprehensive documentation and sample configuration Signed-off-by: Eran Cohen <eranco@redhat.com>
55 lines
1.7 KiB
YAML
55 lines
1.7 KiB
YAML
version: 2
|
|
distribution_spec:
|
|
description: Quick start template for running Llama Stack with several popular providers
|
|
providers:
|
|
inference:
|
|
- provider_type: remote::cerebras
|
|
- provider_type: remote::ollama
|
|
- provider_type: remote::vllm
|
|
- provider_type: remote::tgi
|
|
- provider_type: remote::fireworks
|
|
- provider_type: remote::together
|
|
- provider_type: remote::bedrock
|
|
- provider_type: remote::nvidia
|
|
- provider_type: remote::openai
|
|
- provider_type: remote::anthropic
|
|
- provider_type: remote::gemini
|
|
- provider_type: remote::vertexai
|
|
- provider_type: remote::groq
|
|
- provider_type: remote::sambanova
|
|
- provider_type: inline::sentence-transformers
|
|
vector_io:
|
|
- provider_type: inline::faiss
|
|
- provider_type: inline::sqlite-vec
|
|
- provider_type: inline::milvus
|
|
- provider_type: remote::chromadb
|
|
- provider_type: remote::pgvector
|
|
files:
|
|
- provider_type: inline::localfs
|
|
safety:
|
|
- provider_type: inline::llama-guard
|
|
agents:
|
|
- provider_type: inline::meta-reference
|
|
telemetry:
|
|
- provider_type: inline::meta-reference
|
|
post_training:
|
|
- provider_type: inline::huggingface
|
|
eval:
|
|
- provider_type: inline::meta-reference
|
|
datasetio:
|
|
- provider_type: remote::huggingface
|
|
- provider_type: inline::localfs
|
|
scoring:
|
|
- provider_type: inline::basic
|
|
- provider_type: inline::llm-as-judge
|
|
- provider_type: inline::braintrust
|
|
tool_runtime:
|
|
- provider_type: remote::brave-search
|
|
- provider_type: remote::tavily-search
|
|
- provider_type: inline::rag-runtime
|
|
- provider_type: remote::model-context-protocol
|
|
image_type: venv
|
|
additional_pip_packages:
|
|
- aiosqlite
|
|
- asyncpg
|
|
- sqlalchemy[asyncio]
|