Vespa
An open-source platform for search, recommendation and personalization that serves vectors, tensors, text and structured data with machine-learned ranking at any scale.
- GitHub stars
- 7.1k
- Last commit
- today
- Latest release
- v8.753.16
- Licence
- Apache-2.0
- Self-hosted
- Yes
- Hosted version
- Available
Vespa is a platform for applications that must pick a subset of data from a large, changing corpus, evaluate machine-learned models over it, organize and aggregate the results and return them quickly. Typical examples are search, recommendation and personalization, and the platform adds vector search and retrieval-augmented generation as well, as shown in its topics.
Doing this over large data sets distributed across many nodes and evaluated in parallel is hard, and Vespa handles it with high availability and performance. It handles searching, running inference over and organizing vectors, tensors, text and structured data while serving. According to the README, Vespa has been developed over many years and runs behind several large internet services.
The repository contains all the code needed to build and run Vespa yourself under the Apache 2.0 licence, and new releases are made from the master branch on weekday mornings. You can deploy applications to the Vespa Cloud service, which offers a free trial, or run your own instance following the getting started guide. It is written in Java and C++ and suits teams building large-scale search and recommendation systems.
Key features
- Search over text, vectors and tensors
- Machine-learned model inference at serving time
- Structured data filtering and aggregation
- Distributed, highly available serving
- Real-time updates while serving queries
- Cloud service with a free trial
Pricing: Free and open source under Apache 2.0; Vespa Cloud is a managed service with a free trial.

