Predictable pricing at any scale
Your costs are based on the resources your application uses, not how many queries you run. Scale compute, memory, storage, and GPUs as your workload grows.
Choose how you run Vespa
Every option runs the same Vespa engine. Choose based on where it runs, who operates it, and the level of control your organization needs.
Select the support level that fits your workload
Choose the level of support that fits your needs. Higher tiers provide faster response times, stronger operational guarantees, and more direct access to Vespa experts. You can change tiers as your needs evolve.
-
Basic
Non-critical production workloads
- Next business day response
- CI validation & continuous deployment
-
Professional
Customer-facing production workloads.
- 1 hour response
- 24/7 support
- CI validation & continuous deployment
- Backups & DR
-
Premium
Critical workloads requiring the fastest response and named experts.
- 15-minute response
- 24/7 support
- CI validation & continuous deployment
- Backups & DR
$20k/mo minimum
Visit the Developer Center for additional helpful resources.
Compare Deployment Options
Choose deployment based on where Vespa needs to run and how much of the operational work you want Vespa to handle.
Deployment
|
|
Open Source
Self-managed |
Cloud
Fully managed |
Enclave
Managed in your cloud |
Enterprise
Supported self-hosted |
|
Runs in
|
Your infrastructure
|
Vespa Cloud
|
Your cloud account (AWS, Azure, GCP)
|
Your infrastructure
|
|
Managed by
|
Your team
|
Vespa
|
Vespa
|
Your team, with Vespa support
|
|
Cloud / environment options
|
Cloud or on-premises
|
Supported Vespa Cloud regions
|
Your supported cloud environment
|
Cloud or on-premises
|
Core Engine
|
|
Open Source
Self-managed |
Cloud
Fully managed |
Enclave
Managed in your cloud |
Enterprise
Supported self-hosted |
|
Vector, lexical & hybrid search
|
|
|
|
|
|
Filtered vector search
|
|
|
|
|
|
Multi-phase ranking & reranking
|
|
|
|
|
|
Tensor computation & ML inference
|
|
|
|
|
|
Custom query processing
|
|
|
|
|
|
Grouping, aggregation, & faceting
|
|
|
|
|
|
Streaming search (personal mode)
|
|
|
|
|
|
Real-time indexing & partial updates
|
|
|
|
|
Platform & management
|
|
Open Source
Self-managed |
Cloud
Fully managed |
Enclave
Managed in your cloud |
Enterprise
Supported self-hosted |
|
Managed embedding/inference
|
-
|
|
|
-
|
|
Autoscaling controller
|
-
|
|
|
-
|
|
Application package validation in CLI
|
-
|
|
|
|
|
Performance & latency dashboards
|
-
|
|
|
|
|
Vendor-supported long-term release line
|
-
|
|
|
|
Operations & reliability
|
|
Open Source
Self-managed |
Cloud
Fully managed |
Enclave
Managed in your cloud |
Enterprise
Supported self-hosted |
|
Cluster provisioning & setup
|
Your team
|
Automated
|
Automated
|
Automated, with playbooks
|
|
Software upgrades & patches
|
Manual
|
Continuous, zero downtime
|
Continuous, zero downtime
|
Vendor supported
|
|
Multi-AZ high availability
|
-
|
Included with Basic+ Support
|
Included with Basic+ Support
|
Your team
|
|
Multi-region failover
|
-
|
Included with Premium Support
|
Included with Professional+ Support
|
Your team
|
|
Backups & disaster recovery
|
-
|
Included with Professional+ Support
|
Included with Professional+ Support
|
Your team
|
|
Monitoring & observability
|
-
|
Included
|
Included
|
Included
|
|
Uptime SLA
|
-
|
Based on support plan
|
Based on support plan
|
Based on contract
|
Security & compliance
|
|
Open Source
Self-managed |
Cloud
Fully managed |
Enclave
Managed in your cloud |
Enterprise
Supported self-hosted |
|
SOC 2 Type II
|
-
|
|
|
|
|
GDPR-aligned data residency
|
-
|
EU Regions
|
Your VPC
|
|
|
Customer-managed keys (BYOK)
|
-
|
Add-on
|
|
|
|
PrivateLink / VPC peering
|
-
|
Available
|
|
|
|
SSO / SAML
|
-
|
|
|
|
|
Audit Logging
|
-
|
|
|
|
Frequently asked questions