Skip to content
aicoolies logo
Vespa logo

Vespa

Hybrid search and ML ranking engine at scale

Vespa is an open-source serving engine with 6K+ GitHub stars for hybrid search combining vector similarity, BM25 text ranking, and structured filtering in a single query. Built by Yahoo for web-scale, it handles billions of documents with millisecond latency. Features real-time indexing, ML model serving, tensor computation, and ACID-compliant writes. Supports custom ranking models, query federation, and geographic search. Used for recommendation systems, personalization, and RAG.

About Vespa

Vespa combines vector, text, and structured search in one engine. Billions of documents, millisecond latency. Built by Yahoo for web-scale.

Real-time indexing, built-in ML model serving, tensor computation. ACID-compliant partial updates.

Custom ranking models, query federation, geographic search. Self-hosted or Vespa Cloud managed.

Pricing & Platform Specs

Pricing Summary

Open-source big data search and AI serving engine licensed under Apache-2.0 (Vespa.ai, originally developed by Yahoo). 100% free with zero software licensing costs for unlimited nodes, clusters, and data volumes when self-hosted on bare-metal, VMs, or Kubernetes. Managed cloud deployments are available via Vespa Cloud with usage-based billing per allocated resource (vCPU-hours, GB-memory-hours, disk GB-hours, and GPU hours), automated zero-downtime deployments, multi-zone replication, and enterprise SLAs.

full pricing breakdown →

Supported Platforms

Self-hosted, Docker, Vespa Cloud

Explore categories, tags & use cases

Embedded vector database for multimodal AI with petabyte scale

LanceDB is an open-source embedded vector database built on the Lance columnar format for multimodal AI. It delivers near in-memory performance from disk with zero-copy architecture, supporting vector search, full-text search, and SQL. Native SDKs for Python, TypeScript, and Rust integrate with LangChain, LlamaIndex, and DuckDB. Backed by a $30M Series A, used by Harvey AI and Runway, with 18,000+ GitHub stars.

freemiumOpen Source

Serverless vector and full-text search on object storage

turbopuffer is a serverless vector and full-text search engine built on object storage and vendor-positioned as roughly 10x cheaper than traditional vector databases. Used by Anthropic, Cursor, Notion, and Atlassian for production search workloads. Official site reports 4T+ documents, 10M+ writes/s, and 25k+ queries/s in production systems. Funded by Thrive Capital.

paid

Fully managed RAG-as-a-Service platform for enterprise AI applications

Ragie is a managed retrieval-augmented generation platform that handles document ingestion, indexing, and retrieval so developers can build grounded AI applications without managing vector databases or chunking pipelines. It connects to Google Drive, Notion, Slack, Confluence, and other enterprise data sources with simple APIs for hybrid search and entity extraction.

freemium

Community experience

Sources & verification

Sources checked
Content verified

Verification dates are editorial checks. Routine CMS saves and automatic updatedAt timestamps do not advance them.

FAQ

What is Vespa?

Vespa is an open-source serving engine with 6K+ GitHub stars for hybrid search combining vector similarity, BM25 text ranking, and structured filtering in a single query. Built by Yahoo for web-scale, it handles billions of documents with millisecond latency. Features real-time indexing, ML model serving, tensor computation, and ACID-compliant writes. Supports custom ranking models, query federation, and geographic search. Used for recommendation systems, personalization, and RAG.

Is Vespa free?

Yes — Vespa is open source and free to use. Open-source big data search and AI serving engine licensed under Apache-2.0 (Vespa.ai, originally developed by Yahoo). 100% free with zero software licensing costs for unlimited nodes, clusters, and data volumes when self-hosted on bare-metal, VMs, or Kubernetes. Managed cloud deployments are available via Vespa Cloud with usage-based billing per allocated resource (vCPU-hours, GB-memory-hours, disk GB-hours, and GPU hours), automated zero-downtime deployments, multi-zone replication, and enterprise SLAs.

Is Vespa open source?

Yes — Vespa is open source.

Is Vespa still maintained?

Yes — Vespa is active. Its listing was last verified on September 6, 2026.

What are the best Vespa alternatives?

The first editor-selected Vespa alternatives are LanceDB, turbopuffer, Ragie.

How does Vespa score in our review?

The published editorial review lists Vespa at 82/100 overall across speed, privacy, and developer experience. Check the review's evidence status and test metadata for its verification level.