Checkpointing cuts cold starts 80%, one global URL routes in 9 ms, and protected compute tiers arrive. Action needed before August 1.
͏‌ ͏‌ ͏‌ ͏‌ ͏‌ ͏‌ ͏‌ ͏‌ ͏‌ ͏‌ ͏‌ ͏‌ ͏‌ ͏‌ ͏‌ ͏‌ ͏‌ ͏‌ ͏‌ ͏‌ ͏‌ ͏‌ ͏‌ ͏‌ ͏‌ ͏‌ ͏‌ ͏‌ ͏‌ ͏‌ 
Lumenrack
Product update · June 2026

Lumenrack Functions goes global

June shipped four features that take Functions real-time workloads to a global footprint: checkpointing, global routing, global storage, and protected compute tiers.

80%
lower cold starts with checkpointing
9 ms
routing decisions on the global URL
2x
interruptible rate for protected capacity

Hello from the Lumenrack team. June is the culmination of many months of work on Lumenrack Functions, the serverless container platform for CPU and GPU workloads. Here is what is new.

CPU and GPU checkpointing

Snapshot a fully warmed CPU and GPU container after model weights and kernels are loaded, then restore new replicas from that snapshot instead of starting from scratch. For vLLM, SGLang, and other GPU-heavy workloads this turns multi-minute cold starts into seconds and cuts wasted GPU time during scale-up.

Checkpointing docs →

A deep navy band with diagonal lime stripes and a ring holding a restore-from-history glyph

Global routing

Deploy to a single URL and Lumenrack routes every request to the closest healthy cluster across multiple regions and providers. If a cluster degrades, traffic fails over automatically.

Global routing docs →

Action needed before August 1

From August 1 the legacy single-cluster URLs are no longer supported. They keep working through a proxy, which can add latency; switching to the global URL is a host-only swap that takes about five minutes. Follow the migration guide.

Global storage

One persistent storage layer shared across all your apps and regions. Model weights and shared files are stored once and read everywhere, instead of being duplicated per app or per region.

Global storage docs →

A deep navy band with diagonal lime stripes and a ring holding a database glyph

Protected compute tiers

Each deployment now chooses its compute tier.

Lightning bolt icon

Interruptible (default)
Cheaper preemptible capacity. Nothing changes for existing deployments.

Shield with a check mark icon

Protected
On-demand capacity Lumenrack does not reclaim while it is serving a request. For real-time and latency-sensitive workloads.

From July 1, protected is billed at 2x the interruptible rate. Interruptible pricing is unchanged, and enterprise customers with negotiated pricing are not affected.

Compute tiers docs →

Happy building. The team reads every reply, so send ideas on how the platform could better fit your workloads.

From the blog

SOC 2 Type II

Lumenrack has completed its SOC 2 Type II audit and supports GDPR, HIPAA, and ISO 27001 programmes.

Beacon, the Lumenrack global router

How Beacon picks a healthy, cheaper, lower-latency cluster for every request in single-digit milliseconds, with automatic failover.

Restoring GPU workloads from memory snapshots

How Lumenrack snapshots warmed CPU and GPU memory inside its sandbox runtime and restores workloads like vLLM in seconds.

Lumenrack on X Lumenrack on GitHub Lumenrack on LinkedIn Lumenrack community Slack
You are receiving this because you have a Lumenrack Functions account.
Lumenrack, Inc., 1420 Halyard Row, Oakland, CA 94607
Copyright © 2026 Lumenrack, Inc. All rights reserved.
Privacy  ·  Terms  ·  Email preferences  ·  Unsubscribe