aicoolies logo
Dstack logo
Dstack logo

Dstack

Open-source control plane for AI workloads across multi-cloud GPU infrastructure

open sourceupdated Apr 21, 2026

dstack is an open-source platform that orchestrates AI training and inference workloads across heterogeneous GPU infrastructure spanning multiple clouds, Kubernetes clusters, and bare-metal servers. It abstracts away cloud-specific APIs so teams define GPU requirements declaratively and dstack automatically provisions the cheapest available resources from AWS, GCP, Azure, Lambda, or on-premises hardware.

dstack is a control plane for AI infrastructure that solves the operational complexity of running training and inference workloads across diverse GPU environments. Modern AI teams face a fragmented landscape where GPU availability, pricing, and APIs differ across every cloud provider and on-premises setup. dstack provides a single declarative interface where developers specify what they need — GPU type, count, memory, and framework — and the platform handles provisioning, scheduling, and lifecycle management across all configured backends.

The platform supports NVIDIA, AMD, and Google TPU accelerators across AWS, GCP, Azure, Lambda Cloud, and self-managed Kubernetes or bare-metal clusters. Workloads are defined in YAML configuration files that specify resource requirements, Docker images, and execution commands. dstack's fleet management automatically discovers available GPUs, tracks utilization, and schedules jobs to minimize cost and maximize throughput. The auto-scaling engine provisions and deprovisisions cloud instances based on queue depth.

dstack has raised venture funding and maintains an active open-source project with over 2,000 GitHub stars. The MPL-2.0 license allows commercial use while requiring modifications to the core to be shared. For AI teams that have outgrown the workflow of manually SSH-ing into GPU instances or navigating cloud console UIs, dstack provides the infrastructure abstraction layer that makes multi-cloud GPU orchestration as straightforward as container orchestration with Kubernetes.

Pricing

Free open-source core; commercial managed offering

Platforms

Multi-cloud (AWS, GCP, Azure, Lambda), Kubernetes, bare metal

Categories

Tags

Use Cases

Daytona logo

Daytona

Open-source dev environment management with AI integration

Daytona is secure, elastic infrastructure for running AI-generated code in isolated sandboxes. It gives agents and developer workflows programmable environments with dedicated kernel, filesystem, network, vCPU, memory, and disk, backed by OCI/Docker compatibility, SDK/API access, and under-90ms sandbox startup. The project has 72,000+ GitHub stars and is AGPL-3.0 licensed.

Open Source
Railway logo

Railway

Infrastructure, instantly

Modern cloud platform for deploying full-stack apps, databases, and workers with instant provisioning and usage-based pricing. Deploy from GitHub or CLI with zero config for Node.js, Python, Go, Rust, and Docker. Built-in PostgreSQL, MySQL, Redis, and MongoDB with auto backups. Features private networking, environment management, cron jobs, TCP proxying, and real-time logs. Popular with indie hackers and startups for fast MVPs with a generous free trial including $5 monthly credits.

freemium
Coolify logo

Coolify

Self-hosted Heroku/Vercel alternative

Open-source, self-hostable PaaS alternative to Heroku, Vercel, and Netlify with 44K+ GitHub stars. Deploy static sites, APIs, full-stack apps, databases, and 280+ one-click services on your own VPS or bare metal via SSH. Features auto Let's Encrypt SSL, Git integration (GitHub/GitLab/Bitbucket/Gitea), S3 backups, Docker Swarm support, and a REST API for CI/CD automation. Self-hosted version is free forever with no features behind paywalls.

Open Source
DeepInfra logo

DeepInfra

Cost-effective AI inference platform with 86+ models from $0.02/M tokens

DeepInfra is an AI inference platform offering 86+ LLM models with pricing starting at $0.02 per million tokens. Backed by $20.6M in funding including an $18M Series A from Felicis Ventures, it provides OpenAI-compatible endpoints for models including DeepSeek, Llama, and Mistral with pay-as-you-go pricing.

api-usage-based

Related Tools

computed discovery: shared active categories · kept separate from editor-verified Alternatives

KTransformers parent kvcache-ai logo

KTransformers

Heterogeneous CPU-GPU inference and SFT for large MoE models

Open-source framework for running and fine-tuning large Mixture-of-Experts models with heterogeneous CPU-GPU execution, optimized kernels, limited VRAM and SGLang or LLaMA-Factory integrations.

Open Source
vLLM Production Stack parent vLLM logo

vLLM Production Stack

Official Kubernetes and Helm reference stack built on the vLLM inference engine

Official vLLM reference implementation for scaling the existing inference engine on Kubernetes with Helm, request routing, KV-cache offload, autoscaling and Prometheus/Grafana observability.

Open Source
Dynamo logo

NVIDIA Dynamo

Distributed inference orchestration above vLLM, SGLang and TensorRT-LLM

Open-source, datacenter-scale orchestration layer that coordinates vLLM, SGLang and TensorRT-LLM across nodes with disaggregated serving, KV-aware routing, multi-tier cache management and automatic scaling.

Open Source
GPUStack logo

GPUStack

Open-source GPU control plane for scalable AI model serving

Open-source GPU cluster manager that configures vLLM, SGLang, TensorRT-LLM or custom engines, serves models through compatible APIs, and provisions SSH-accessible GPU instances across on-premises, Kubernetes and cloud environments.

Open Source
Mooncake logo

Mooncake

Disaggregated KV cache storage and transfer for LLM serving

Open-source infrastructure for disaggregated LLM serving that pools KV caches across prefill and decode workers, with high-performance transfer, distributed storage and integrations for vLLM and SGLang.

Open Source
LMCache logo

LMCache

Reusable KV cache infrastructure for scalable LLM inference

Open-source KV cache management layer that persists, offloads and reuses model key-value caches across requests and serving engines to reduce repeated prefill work and improve inference throughput.

Open Source

FAQ

What is Dstack?

dstack is an open-source platform that orchestrates AI training and inference workloads across heterogeneous GPU infrastructure spanning multiple clouds, Kubernetes clusters, and bare-metal servers. It abstracts away cloud-specific APIs so teams define GPU requirements declaratively and dstack automatically provisions the cheapest available resources from AWS, GCP, Azure, Lambda, or on-premises hardware.

Is Dstack free?

Yes — Dstack is open source and free to use. Free open-source core; commercial managed offering

Is Dstack open source?

Yes — Dstack is open source.

What are the best Dstack alternatives?

The top editor-verified Dstack alternatives are Daytona, Railway, Coolify, and more.