Open source, cloud native, Postgres platform with copy-on-write branching and scale-to-zero
-
Updated
Jul 30, 2026 - Go
Open source, cloud native, Postgres platform with copy-on-write branching and scale-to-zero
Kubernetes-native scale-to-zero with zero traffic loss, no code changes, and direct integration with kubernetes resources
KEDA External gRPC Scaler for GPU workloads - native NVML metrics via DaemonSet, no Prometheus required
CNPG plugin to manage scale to zero functionality
Multi-tenant AI assistant platform on Amazon EKS. One-command deploy, scale-to-zero per user, powered by Amazon Bedrock.
A scale to zero Minecraft server running on Fly.io
Scale-to-zero NAT instances for AWS
Kubernetes-native control plane for scale-to-zero serving of long-tail LLMs
Personal search for technical screenshots and notes. Hybrid vector and full-text retrieval on Amazon EKS, with Firn and S3 as the storage and index layer, and GPU capacity that scales to zero.
Scale idle apps to zero and wake them up when they receive traffic
Scale-to-zero with wake-on-request for Kubernetes. Sleep idle services on a schedule, wake them instantly on HTTP access. Single pod, no CRDs, no sidecars.
Wake-on-request and scale-to-zero for HashiCorp Nomad services—Rust proxy for Traefik, Consul, and Redis that cuts idle cost without breaking long-running requests
Multi-tenant Kubernetes operator for self-hosted GitHub Actions runners. Scale-to-zero workers, per-tenant egress IP pools, and GPU priority scheduling across a shared ResourceQuota — an Actions Runner Controller (ARC) alternative.
On-demand TCP+UDP proxy for Docker containers.
A fully automated, scale-to-zero AWS ECS Fargate platform — wake-on-demand via API Gateway + Lambda, auto-sleep via EventBridge, Terraform IaC, and GitHub Actions OIDC CI/CD. Zero idle cost. Clean, modern, conference-ready architecture.
Scale-to-zero GPU inference on Kubernetes with KEDA pod autoscaling and Cluster Autoscaler node provisioning
Distributed, cloud-native Roaring Bitmaps — query and intersect billion-scale integer sets over tiered cloud storage (RAM → NoSQL → object store), at a fraction of an always-on cache.
Queue-driven, scale-from-zero GPU inference for any Kubernetes — bursts to cross-region VMs when GPUs run dry
Serverless-GPU LLM serving: scale-to-zero with fast GPU snapshot/restore (cuda-checkpoint), multi-tenant packing, and an OpenAI-compatible API — built on vLLM.
Multi-tenant AI assistant platform on Amazon EKS. One-command deploy, scale-to-zero per user, powered by Amazon Bedrock.
Add a description, image, and links to the scale-to-zero topic page so that developers can more easily learn about it.
To associate your repository with the scale-to-zero topic, visit your repo's landing page and select "manage topics."