aicoolies logo
Holori logo
Holori logo

Holori

FOCUS-native multi-cloud cost management and FinOps platform

freemiumupdated Aug 16, 2026

Holori is a multi-cloud cost management platform built on the FOCUS billing data standard. It provides unified cost visibility across AWS, Azure, GCP, and other cloud providers with automated tagging, budget alerts, and optimization recommendations. Features interactive infrastructure diagrams that link architecture visualization directly to cost data for contextual spending analysis.

Holori differentiates in the FinOps space by building natively on the FOCUS (FinOps Cost and Usage Specification) billing data standard, enabling consistent cost analysis across multiple cloud providers without the normalization headaches that plague multi-cloud environments. The platform ingests billing data from AWS, Azure, GCP, and other providers, presenting unified dashboards that break down spending by service, team, project, or any custom dimension regardless of the underlying cloud provider's billing format.

The interactive infrastructure diagram feature sets Holori apart from spreadsheet-oriented cost tools by visualizing cloud architecture and overlaying cost data directly on infrastructure components. Engineers and finance teams can explore spending in the context of how resources relate to each other architecturally, making it easier to understand why costs increased and which design decisions drive the most significant spending. This visual approach bridges the communication gap between engineering teams who understand architecture and finance teams who manage budgets.

Holori provides automated tagging recommendations to improve cost allocation accuracy, budget alerts with customizable thresholds and notification channels, and optimization recommendations for right-sizing, Reserved Instance coverage, and idle resource cleanup. The platform targets mid-market organizations that need multi-cloud cost visibility without the complexity and price tag of enterprise FinOps platforms, offering a streamlined experience focused on actionable insights rather than exhaustive feature sets.

Pricing

Free tier available; paid plans based on cloud spend

Platforms

SaaS, AWS/Azure/GCP, FOCUS billing standard

Categories

Tags

Use Cases

Related Tools

computed discovery: shared active categories · kept separate from editor-verified Alternatives

KTransformers parent kvcache-ai logo

KTransformers

Heterogeneous CPU-GPU inference and SFT for large MoE models

Open-source framework for running and fine-tuning large Mixture-of-Experts models with heterogeneous CPU-GPU execution, optimized kernels, limited VRAM and SGLang or LLaMA-Factory integrations.

Open Source
vLLM Production Stack parent vLLM logo

vLLM Production Stack

Official Kubernetes and Helm reference stack built on the vLLM inference engine

Official vLLM reference implementation for scaling the existing inference engine on Kubernetes with Helm, request routing, KV-cache offload, autoscaling and Prometheus/Grafana observability.

Open Source
Dynamo logo

NVIDIA Dynamo

Distributed inference orchestration above vLLM, SGLang and TensorRT-LLM

Open-source, datacenter-scale orchestration layer that coordinates vLLM, SGLang and TensorRT-LLM across nodes with disaggregated serving, KV-aware routing, multi-tier cache management and automatic scaling.

Open Source
GPUStack logo

GPUStack

Open-source GPU control plane for scalable AI model serving

Open-source GPU cluster manager that configures vLLM, SGLang, TensorRT-LLM or custom engines, serves models through compatible APIs, and provisions SSH-accessible GPU instances across on-premises, Kubernetes and cloud environments.

Open Source
Mooncake logo

Mooncake

Disaggregated KV cache storage and transfer for LLM serving

Open-source infrastructure for disaggregated LLM serving that pools KV caches across prefill and decode workers, with high-performance transfer, distributed storage and integrations for vLLM and SGLang.

Open Source
LMCache logo

LMCache

Reusable KV cache infrastructure for scalable LLM inference

Open-source KV cache management layer that persists, offloads and reuses model key-value caches across requests and serving engines to reduce repeated prefill work and improve inference throughput.

Open Source

Used in Stacks

FAQ

What is Holori?

Holori is a multi-cloud cost management platform built on the FOCUS billing data standard. It provides unified cost visibility across AWS, Azure, GCP, and other cloud providers with automated tagging, budget alerts, and optimization recommendations. Features interactive infrastructure diagrams that link architecture visualization directly to cost data for contextual spending analysis.

Is Holori free?

Holori offers a free tier alongside paid plans. Free tier available; paid plans based on cloud spend

What are the best Holori alternatives?

The top editor-verified Holori alternatives are Zesty, Kubecost.